EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 5 · 28 per page
speech-to-text
ElevenLabsREVIEW REQUIRED

ElevenLabs Speech to Text

fal-ai/elevenlabs/speech-to-text

Generate text from speech using ElevenLabs advanced speech-to-text model.

speech
video-to-video
MiniMaxREVIEW REQUIRED

H3 Max Recast

minimax/h3-max/recast

Recast the people in a video using reference photos with H3 Max, while preserving the source motion, camera, cuts, and audio.

video-to-videovideo-editingcharacter-replacement
video-to-video
KlingREVIEW REQUIRED

Kling Video

fal-ai/kling-video/v3/pro/motion-control

Transfer movements from a reference video to any character image. Cost-effective mode for motion transfer, perfect for portraits and simple animations.

stylizedtransformediting
image-to-video
KlingREVIEW REQUIRED

Kling O3 Image to Video [Pro]

fal-ai/kling-video/o3/standard/image-to-video

Generate a video by taking a start frame and an end frame, animating the transition between them while following text-driven style and scene guidance.

image-to-video
image-to-image
pixelcutREVIEW REQUIRED

Pixelcut Background Remover

pixelcut/background-removal

Pixelcut’s Background Remover enables fast, ultra high-quality removal of backgrounds from images. Perfect for e-commerce and image editing workflows. Powered by advanced AI for clean, perfect cutouts every time.

background removalutilityremove background
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Fast Image to Video

bytedance/seedance-2.0/fast/image-to-video

ByteDance's most advanced image-to-video model, fast tier. Lower latency and cost with synchronized audio, start and end frame control, and motion prompts.

stylizedtransformlipsync
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX 2 Edit

fal-ai/flux-2/edit

Image-to-image editing with FLUX.2 [dev] from Black Forest Labs. Precise modifications using natural language descriptions and hex color control.

text-to-video
GoogleREVIEW REQUIRED

Veo 3.1 Fast

fal-ai/veo3.1/fast

Faster and more cost effective version of Google's Veo 3.1!

text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX 2 Flash

fal-ai/flux-2/flash

Text-to-image generation with FLUX.2 [dev] from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities— in a flash.

image-to-video
GoogleREVIEW REQUIRED

Gemini Omni Flash 1.1 Reference to Video

google/gemini-omni-flash/v1.1/reference-to-video

Gemini Omni Flash 1.1 is Google's multimodal video model. This endpoint generates video from combined multimodal references, images, videos and text together. Reasoning across all inputs to produce a single coherent result, with characters retaining their face, clothing, and voice throughout

stylizedtransformlipsync
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Kontext [max]

fal-ai/flux-pro/kontext/max

FLUX.1 Kontext [max] is a model with greatly improved prompt adherence and typography generation meet premium consistency for editing without compromise on speed.

audio-to-audio
falREVIEW REQUIRED

Demucs

fal-ai/demucs

SOTA stemming model for voice, drums, bass, guitar and more.

audio
video-to-video
topazREVIEW REQUIRED

Topaz Upscale Video Precision

topaz/upscale/video/precision

Professional video upscaling powered by Topaz Labs. Precision models (Proteus, Artemis, Iris, Dione, Theia, Gaia, Rhea) enhance footage up to 4x while staying faithful to the source. Best for clean, natural upscales of real-world footage.

upscalevideo
video-to-video
falREVIEW REQUIRED

Ffmpeg Api

fal-ai/ffmpeg-api/merge-videos

Use ffmpeg capabilities to merge 2 or more videos.

image-to-3d
falREVIEW REQUIRED

Trellis 2

fal-ai/trellis-2

Generate 3D models from your images using Trellis 2. A native 3D generative model enabling versatile and high-quality 3D asset creation.

image-to-3D
text-to-image
RecraftREVIEW REQUIRED

Recraft V3

fal-ai/recraft/v3/text-to-image

Recraft V3 is a text-to-image model with the ability to generate long texts, vector art, images in brand style, and much more. As of today, it is SOTA in image generation, proven by Hugging Face's industry-leading Text-to-Image Benchmark by Artificial Analysis.

vectortypographystyle
text-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Text to Video API

bytedance/seedance-2.0/text-to-video

ByteDance's most advanced text-to-video model. Cinematic output with native audio, multi-shot editing, real-world physics, and director-level camera control.

stylizedtransformlipsync
image-to-video
KlingREVIEW REQUIRED

Kling Video V3 Turbo Pro Image to Video

fal-ai/kling-video/v3/turbo/pro/image-to-video

Generate high quality 1080p videos from images using Kling's Turbo 3.0 model, with improved lipsync and multishot generation capabilities.

klingv3turbo1080p
image-to-image
falREVIEW REQUIRED

Ffmpeg Api

fal-ai/ffmpeg-api/extract-frame

ffmpeg endpoint for first, middle and last frame extraction from videos

utilityediting
video-to-video
ByteDanceREVIEW REQUIRED

Bytedance Upscaler Upscale Video

fal-ai/bytedance-upscaler/upscale/video

Upscale videos with Bytedance's video upscaler.

upscalervideobytedance
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit 2511

fal-ai/qwen-image-edit-2511

Endpoint for Qwen's Image Editing 2511 model.

stylizedtransform
video-to-video
falREVIEW REQUIRED

Sync Lipsync

fal-ai/sync-lipsync/v2/pro

Generate high-quality realistic lipsync animations from audio while preserving unique details like natural teeth and unique facial features using the state-of-the-art Sync Lipsync 2 Pro model.

animationlip synchigh-quality
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Fast Reference to Video

bytedance/seedance-2.0/fast/reference-to-video

ByteDance's most advanced reference-to-video model, fast tier. Lower latency and cost with up to 9 images, 3 videos, and 3 audio clips as inputs.

stylizedtransformlipsync
image-to-image
ByteDanceREVIEW REQUIRED

Seedream

bytedance/seedream/v5/flash/edit

Seedream 5.0 Flash is a fast image generation and editing model, built for workflows where speed and budget matter.

realismtypographystylized
image-to-image
IdeogramREVIEW REQUIRED

Ideogram Remove Background

fal-ai/ideogram/remove-background

Remove backgrounds from existing images with Ideogram's remove background feature. Isolate subjects cleanly for compositing and creative reuse.

image-to-image
falREVIEW REQUIRED

Remove Background

fal-ai/imageutils/rembg

Remove the background from an image.

background removalutilityediting
image-to-video
KlingREVIEW REQUIRED

Kling Video

fal-ai/kling-video/v2.5-turbo/standard/image-to-video

Kling 2.5 Turbo Standard: Top-tier image-to-video generation with unparalleled motion fluidity, cinematic visuals, and exceptional prompt precision.

stylizedtransform