EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 9 · 28 per page
image-to-image
ByteDanceREVIEW REQUIRED

Seedream

bytedance/seedream/v5/lite/edit

Image editing endpoint for the fast Lite version of Seedream 5.0, supporting high quality intelligent image editing with multiple inputs.

bytedanceseedream-5.0-liteedit
video-to-video
KlingREVIEW REQUIRED

Kling O1 Edit Video [Pro]

fal-ai/kling-video/o1/video-to-video/edit

Edit an existing video using natural-language instructions, transforming subjects, settings, and style while retaining the original motion structure.

image-to-image
falREVIEW REQUIRED

Sam 3 1

fal-ai/sam-3-1/image

SAM 3.1 builds comes with Object Multiplex, a shared-memory approach for joint multi-object tracking that delivers faster speeds with larger number of objects tracked.

segmentationmaskreal-time
text-to-image
RecraftREVIEW REQUIRED

Recraft V4.1 Text to Image

fal-ai/recraft/v4.1/text-to-image

Recraft V4.1 builds on the design-first foundation of V4 with sharper prompt control and cleaner composition. Tuned for brand systems and editorial work, it delivers production-ready raster images that hold up next to a designer's hand.

stylizedtransformtypography
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Kontext [max]

fal-ai/flux-pro/kontext/max/multi

Experimental version of FLUX.1 Kontext [max] with multi image handling capabilities

image-to-image
falREVIEW REQUIRED

FASHN Virtual Try-On V1.6

fal-ai/fashn/tryon/v1.6

FASHN v1.6 delivers precise virtual try-on capabilities, accurately rendering garment details like text and patterns at 864x1296 resolution from both on-model and flat-lay photo references.

try-onfashionclothing
text-to-image
GoogleREVIEW REQUIRED

Gemini 2.5 Flash Image

fal-ai/gemini-25-flash-image

Google's famous original image generation and editing model, a.k.a Nano Banana

text-to-image
video-to-video
topazREVIEW REQUIRED

Topaz Upscale Video Generative

topaz/upscale/video/generative

Professional generative video upscaling powered by Topaz Labs. Starlight models rebuild detail that is not in the source, with Fast variants at half the price. Best for low-quality, compressed or archive footage.

upscalevideo
video-to-video
falREVIEW REQUIRED

Workflow Utilities Trim Video

fal-ai/workflow-utilities/trim-video

FFMPEG Utility for Trim Video

video-to-video
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX 2 Pro Outpaint

fal-ai/flux-2-pro/outpaint

Outpainting generation with FLUX.2 [pro] from Black Forest Labs. Optimized for maximum quality, exceptional photorealism and artistic images.

image-to-imageoutpaintoutpainting
text-to-speech
AlibabaREVIEW REQUIRED

Qwen 3 TTS - Text to Speech [1.7B]

fal-ai/qwen-3-tts/text-to-speech/1.7b

Bring speech to your texts using Qwen3-TTS Custom-Voice model with pre-trained voices or use your custom voice with Qwen3-TTS Clone Voice model

text-to-speech
image-to-video
falREVIEW REQUIRED

Heygen

fal-ai/heygen/avatar4/image-to-video

Heygen Photo Avatar 4 Model

image-to-video
image-to-video
falREVIEW REQUIRED

Wan 2.5 Image to Video

fal-ai/wan-25-preview/image-to-video

Wan 2.5 image-to-video model.

image-to-image
Black Forest LabsREVIEW REQUIRED

PuLID Flux

fal-ai/flux-pulid

An endpoint for personalized image generation using Flux as per given description.

personalizationstyle transfer
video-to-text
openrouterREVIEW REQUIRED

OpenRouter [Video]

openrouter/router/video

Run any video-capable LLM with fal. Analyze, summarize, and understand video files using Gemini (Google) models. Supports mp4, mpeg, mov, webm, and YouTube links. Powered by OpenRouter.

image-to-image
falREVIEW REQUIRED

EVF-SAM2 Segmentation

fal-ai/evf-sam

EVF-SAM2 combines natural language understanding with advanced segmentation capabilities, allowing you to precisely mask image regions using intuitive positive and negative text prompts.

segmentationmask
image-to-video
AlibabaREVIEW REQUIRED

Wan v2.2 A14B

fal-ai/wan/v2.2-a14b/image-to-video

fal-ai/wan/v2.2-A14B/image-to-video

text-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 Text to Video

blackforestlabs/flux-3/text-to-video

FLUX 3 is Black Forest Labs' frontier video model. This endpoint generates video directly from a text prompt, translating a written description into motion, composition, and scene.

stylizedtransformlipsync
image-to-video
GoogleREVIEW REQUIRED

Veo3.1 Lite FLF

fal-ai/veo3.1/lite/first-last-frame-to-video

Veo 3.1 Lite balances practical utility with professional capabilities, supporting Text-to-Video and Image-to-Video

stylizedtransformlipsync
image-to-video
KlingREVIEW REQUIRED

Kling Video

fal-ai/kling-video/v3/4k/image-to-video

Kling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling

stylizedtransformlipsync
text-to-video
KlingREVIEW REQUIRED

Kling Video v2.6 Text to Video

fal-ai/kling-video/v2.6/pro/text-to-video

Kling 2.6 Pro: Top-tier text-to-video with cinematic visuals, fluid motion, and native audio generation.

image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX 2 Turbo Edit

fal-ai/flux-2/turbo/edit

Image-to-image editing with FLUX.2 [dev] from Black Forest Labs. Precise modifications using natural language descriptions and hex color control—all at turbo speed.

text-to-image
RecraftREVIEW REQUIRED

Recraft V4

fal-ai/recraft/v4/text-to-image

Recraft V4 was developed with designers to bring true visual taste to AI image generation. Built for brand systems and production-ready workflows, it goes beyond prompt accuracy delivering stronger composition, refined lighting, realistic materials, and a cohesive aesthetic. The result is imagery shaped by professional design judgment, ready for immediate real-world use without additional post-processing.

text-to-image
video-to-video
veedREVIEW REQUIRED

Lipsync

veed/lipsync

Generate realistic lipsync from any audio using VEED's model.

lipsyncvideo-to-videoavatar
image-to-3d
tripo3dREVIEW REQUIRED

Tripo H3.1 Multiview to 3D

tripo3d/h3.1/multiview-to-3d

Generate 3D models from multiple view images using Tripo H3.1.

3dmultiview-to-3d3d-generationtripo
video-to-video
KlingREVIEW REQUIRED

Kling Video v2.6 Motion Control [Pro]

fal-ai/kling-video/v2.6/pro/motion-control

Transfer movements from a reference video to any character image. Pro mode delivers higher quality output, ideal for complex dance moves and gestures.

text-to-audio
falREVIEW REQUIRED

Stable Audio 3

fal-ai/stable-audio-3/medium/text-to-audio

Stable Audio 3 Medium is a 1.4 billion parameter latent diffusion model that generates high-quality stereo music up to 6 minutes from text prompts, trained on fully licensed data for safe commercial use.

musicaudiostereo