EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 20 · 28 per page
video-to-video
KlingREVIEW REQUIRED

Kling O1 Edit Video [Standard]

fal-ai/kling-video/o1/standard/video-to-video/edit

Edit an existing video using natural-language instructions, transforming subjects, settings, and style while retaining the original motion structure.

image-to-image
falREVIEW REQUIRED

Bria Background Replace

fal-ai/bria/background/replace

Bria Background Replace allows for efficient swapping of backgrounds in images via text prompts or reference image, delivering realistic and polished results. Trained exclusively on licensed data for safe and risk-free commercial use

image editing
text-to-audio
falREVIEW REQUIRED

Stable Audio 3 Small Music Text to Audio

fal-ai/stable-audio-3/small/music/text-to-audio

Stable Audio 3 Small Music is a 459 million parameter latent diffusion model that generates full stereo music compositions up to 2 minutes from text prompts, lightweight enough for on-device deployment.

musicon-devicelightweight
image-to-image
IdeogramREVIEW REQUIRED

Ideogram Object Removal

fal-ai/ideogram/object-removal

Prompt-free object removal from an image and mask, erasing objects with their shadows and reflections and reconstructing the scene cleanly.

utilityediting
image-to-image
falREVIEW REQUIRED

Phota

fal-ai/phota/edit

Phota's model enables personalized photo editing, preserving identity while erasing distractions seamlessly.

editpersonalizationtypographyphota
image-to-video
MiniMaxREVIEW REQUIRED

MiniMax Hailuo 2.3 Fast [Pro] (Image to Video)

fal-ai/minimax/hailuo-2.3-fast/pro/image-to-video

MiniMax Hailuo-2.3-Fast Image To Video API (Pro, 1080p): Advanced fast image-to-video generation model with 1080p resolution

image-to-video
video-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 Extend Video

blackforestlabs/flux-3/extend-video

FLUX 3 is Black Forest Labs' frontier video model. This endpoint continues an existing clip beyond its final frame, generating additional footage that stays consistent with the original motion and scene.

stylizedtransformlipsync
image-to-image
falREVIEW REQUIRED

Florence-2 Large

fal-ai/florence-2-large/open-vocabulary-detection

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

multimodalvisiondetection
text-to-image
IdeogramREVIEW REQUIRED

Ideogram Transparent

fal-ai/ideogram/v3/generate-transparent

Generate images with transparent backgrounds using Ideogram Transparent model

stylizedtransformtypography
image-to-image
OpenAIREVIEW REQUIRED

GPT Image 1 Mini

fal-ai/gpt-image-1-mini/edit

GPT Image 1 mini combines OpenAI's advanced language capabilities, powered by GPT-5, with GPT Image 1 Mini for efficient image generation.

image-to-image
video-to-video
falREVIEW REQUIRED

Wan VACE 14B

fal-ai/wan-vace-14b/inpainting

VACE is a video generation model that uses a source image, mask, and video to create prompted videos with controllable sources.

image-to-videovideo-to-videotext-to-video
video-to-video
falREVIEW REQUIRED

Sam 3 1

fal-ai/sam-3-1/video

SAM 3.1 builds comes with Object Multiplex, a shared-memory approach for joint multi-object tracking that delivers faster speeds with larger number of objects tracked.

segmentationmaskreal-time
image-to-image
IdeogramREVIEW REQUIRED

Ideogram V4.0q Image to Image

ideogram/v4/image-to-image

Ideogram V4.0q Image-to-Image transforms an input image with a text prompt, restyling and reworking the composition while preserving its core structure for prompt-faithful, high-fidelity edits.

realismtypographystylized
video-to-video
GoogleREVIEW REQUIRED

Veo 3.1 Fast

fal-ai/veo3.1/fast/extend-video

Extend Veo-Created Videos up to 30 seconds

extend-video
text-to-speech
falREVIEW REQUIRED

Zonos2 Text to Speech

fal-ai/zonos2

Zonos2 is a text-to-speech model that clones a voice from a short sample and speaks naturally across many languages.

text-to-speechttsvoice cloning
text-to-image
kreaREVIEW REQUIRED

Krea 2 Medium Text to Image Turbo

krea/v2/medium/turbo/text-to-image

Generate high-fidelity images extremely fast from text with Krea 2 Medium Turbo, supporting aspect ratio, creativity, seed controls, and optional style references.

stylizedtransformtypography
image-to-image
microsoftREVIEW REQUIRED

MAI Image 2.5 Pro (Edit)

microsoft/mai-image-2.5-pro/edit

Apply precise, controllable edits to a reference image while preserving composition, typography, identity, and fine visual detail.

image-editingtypographyphotorealismcontrollable-editing
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V5 Image To Video

fal-ai/pixverse/v5/image-to-video

Generate high quality video clips from text and image prompts using PixVerse v5

stylizedtransform
text-to-video
falREVIEW REQUIRED

LTX Video 2.3 Pro

fal-ai/ltx-2.3/text-to-video

LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.

stylizedtransformlipsync
image-to-image
RecraftREVIEW REQUIRED

Recraft V3

fal-ai/recraft/v3/image-to-image

Recraft V3 is a text-to-image model with the ability to generate long texts, vector art, images in brand style, and much more. As of today, it is SOTA in image generation, proven by Hugging Face's industry-leading Text-to-Image Benchmark by Artificial Analysis.

vectortypographystyle
text-to-video
KlingREVIEW REQUIRED

Kling Video V3 Text to Video 4K

fal-ai/kling-video/v3/4k/text-to-video

Kling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling

stylizedtransformlipsync
text-to-video
AlibabaREVIEW REQUIRED

Wan v2.6 Text to Video

wan/v2.6/text-to-video

Wan 2.6 text-to-video model.

text-to-video
image-to-image
Luma AIREVIEW REQUIRED

Luma Uni-1 Edit

luma/agent/uni-1/v1/edit

Luma Uni-1 Edit reworks a source image from a text instruction, preserving the original composition while applying style changes and following optional reference images to steer the result.

stylizedtransform
image-to-3d
hitem3dREVIEW REQUIRED

Hi3D Image to 3D

hitem3d/hi3d/v3.0/image-to-3d

Generate 3D models from a single image with Hi3D V3.0.

image-to-3d3dmesh
image-to-image
falREVIEW REQUIRED

Sam 3 1

fal-ai/sam-3-1/image-rle

SAM 3.1 builds comes with Object Multiplex, a shared-memory approach for joint multi-object tracking that delivers faster speeds with larger number of objects tracked.

segmentationmaskreal-time
text-to-3d
falREVIEW REQUIRED

Hunyuan 3D Pro Text to 3D

fal-ai/hunyuan-3d/v3.1/pro/text-to-3d

Generate 3D models from text prompts with Hunyuan 3D Pro

3dhunyuantext-to-3d
text-to-video
lightricksREVIEW REQUIRED

Ltx 2.5 Text to Video Fast

lightricks/ltx-2.5/text-to-video/fast

LTX-2.5 is Lightricks' open-source audio-video model. This endpoint generates synchronized video and audio from a text prompt in a single pass, in a speed-optimized mode built for rapid iteration and previews.

stylizedtransformlipsync
text-to-audio
MiniMaxREVIEW REQUIRED

Minimax Music 2.5

fal-ai/minimax-music/v2.5

MiniMax Music 2.5 creates complete tracks with singing, backing music, and detailed arrangements from lyrics and a style description.

stylizedtransformlipsync