EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 27 · 28 per page
text-to-image
falREVIEW REQUIRED

Phota Text to Image

fal-ai/phota

Phota's model empowers developers, photographers, and creators with personalized photograph generation and editing.

stylizedtransformtypographyphota
llm
ByteDanceREVIEW REQUIRED

Bytedance Seed V2 Mini

fal-ai/bytedance/seed/v2/mini

Seed 2.0 Mini is a high-performance multimodal model optimized for low latency and high concurrency. It supports text, image, and video input with 256K context and configurable thinking/reasoning modes.

text-to-image
Black Forest LabsREVIEW REQUIRED

Juggernaut Flux Lightning

rundiffusion-fal/juggernaut-flux/lightning

Juggernaut Lightning Flux by RunDiffusion provides blazing-fast, high-quality images rendered at five times the speed of Flux. Perfect for mood boards and mass ideation, this model excels in both realism and prompt adherence.

image generation
video-to-video
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/render-to-real

Transform your 3D video render into realistic using first frame with Ltx 2.3

3dvideo
image-to-image
falREVIEW REQUIRED

Stable Diffusion with LoRAs

fal-ai/lora/image-to-image

Run Any Stable Diffusion model with customizable LoRA weights.

diffusionloracustomizationfine-tuning
text-to-image
falREVIEW REQUIRED

Wan 2.5 Text to Image

fal-ai/wan-25-preview/text-to-image

Wan 2.5 text-to-image model.

text-to-video
KlingREVIEW REQUIRED

Kling Video

fal-ai/kling-video/o3/4k/text-to-video

Kling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling

stylizedtransformlipsync
text-to-video
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.2-5b/text-to-video/fast-wan

Wan 2.2's 5B FastVideo model produces up to 5 seconds of video 720p at 24FPS with fluid motion and powerful prompt understanding

text to videomotion
image-to-image
falREVIEW REQUIRED

Bria GenFill

fal-ai/bria/genfill

Bria GenFill enables high-quality object addition or visual transformation. Trained exclusively on licensed data for safe and risk-free commercial use. Access the model's source code and weights: https://bria.ai/contact-us

image editing
text-to-image
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.2-a14b/text-to-image

Wan 2.2's 14B model generates high-resolution, photorealistic images with powerful prompt understanding and fine-grained visual detail

image-to-image
MiniMaxREVIEW REQUIRED

Minimax Image Subject Reference

fal-ai/minimax/image-01/subject-reference

Generate images from text and a reference image using MiniMax Image-01 for consistent character appearance.

stylizedtransform
training
AlibabaREVIEW REQUIRED

Qwen Image 2512 Trainer

fal-ai/qwen-image-2512-trainer

Qwen Image 2512 LoRA training

lorapersonalization
video-to-video
clarityaiREVIEW REQUIRED

Crystal Upscaler [Video]

clarityai/crystal-video-upscaler

Do high precision video upscaling that respects the original video perfectly using Crystal Upscaler's new video upscaling method!

upscalevideo-to-video
text-to-video
falREVIEW REQUIRED

LTX Video-0.9.7 13B Distilled

fal-ai/ltx-video-13b-distilled

Generate videos from prompts using LTX Video-0.9.7 13B Distilled and custom LoRA

videoltx-videotext-to-video
vision
falREVIEW REQUIRED

Moondream3 Preview [Caption]

fal-ai/moondream3-preview/caption

Moondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.

Vision
image-to-image
falREVIEW REQUIRED

Firered Image Edit V1.1

fal-ai/firered-image-edit-v1.1

FireRed Image Edit v1.1 is an updated version of FireRed Image Edit, with improved image editing capabilities.

firered-image-edit
video-to-video
decartREVIEW REQUIRED

Lucy 2.1 VTON Realtime

decart/lucy2-vton/realtime

Realtime Try On experience with Decart Lucy 2.1 VTON

image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] Control LoRA Depth

fal-ai/flux-control-lora-depth/image-to-image

FLUX Control LoRA Depth is a high-performance endpoint that uses a control image using a depth map to transfer structure to the generated image and another initial image to guide color.

lorastyle transfer
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Max

fal-ai/qwen-image-max/edit

Image editing endpoint for Qwen-Image-Max. Qwen Image Max improves upon the Qwen Image Plus series by enhancing the realism and naturalness of images.

qwen-imagemax
image-to-image
topazREVIEW REQUIRED

Topaz Denoise Image

topaz/denoise/image

Professional photo denoising powered by Topaz Labs. Normal, Strong and Extreme presets clean noise at source resolution; Denoise Max adds generative detail recovery. Best for high-ISO and night photography.

restoreimage
image-to-video
KlingREVIEW REQUIRED

Kling O1 Reference Image to Video [Standard]

fal-ai/kling-video/o1/standard/reference-to-video

Transform images, elements, and text into consistent, high-quality video scenes, ensuring stable character identity, object details, and environments.

image-to-image
falREVIEW REQUIRED

Live Portrait

fal-ai/live-portrait/image

Transfer expression from a video to a portrait.

expressionanimation
text-to-speech
falREVIEW REQUIRED

Orpheus TTS

fal-ai/orpheus-tts

Orpheus TTS is a state-of-the-art, Llama-based Speech-LLM designed for high-quality, empathetic text-to-speech generation. This model has been finetuned to deliver human-level speech synthesis, achieving exceptional clarity, expressiveness, and real-time performances.

text to speechvoice synthesishigh-fidelity
text-to-3d
falREVIEW REQUIRED

Hyper3D - Rodin V2.5 - Text to 3D - Fast

fal-ai/hyper3d/rodin/v2.5/text-to-3d/fast

Rodin V2.5 by Hyper3D generates realistic and production ready 3D models from text or images. Do fast prototyping using the fast model.

text-to-3d
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit 2509

fal-ai/qwen-image-edit-2509

Endpoint for Qwen's Image Editing Plus model also known as Qwen-Image-Edit-2509. Has superior text editing capabilities and multi-image support.

image-editingimage-to-imagehigh-quality-text
video-to-video
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/reference-video-to-video

Generate high-quality video with audio from reference video, text and images using LTX-2.3

image-to-json
briaREVIEW REQUIRED

Bria Ad Delayer: Convert Flat Ads into Editable Layers | fal

bria/ad-delayer

Turn any flat ad image into fully editable layers —background, product and logo cutouts, live text with typography, and vector shapes. Commercial-safe, structured JSON output

utilityediting
image-to-video
falREVIEW REQUIRED

LongCat Video Distilled

fal-ai/longcat-video/distilled/image-to-video/720p

Generate long videos in 720p/30fps from images using LongCat Video Distilled