EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 12 · 28 per page
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX 2 Lora

fal-ai/flux-2/lora

Text-to-image generation with LoRA support for FLUX.2 [dev] from Black Forest Labs. Custom style adaptation and fine-tuned model variations.

video-to-video
MiniMaxREVIEW REQUIRED

H3 Max Extend Video

minimax/h3-max/extend-video

H3 Max Extend Video adds a text-guided continuation to an existing video. It supports prompt expansion, adjustable duration and aspect ratio, and output resolutions from 480p to 2K, returning either the full extended video or only the new footage.

extendvideocontinuation
audio-to-audio
falREVIEW REQUIRED

Sam Audio

fal-ai/sam-audio/separate

Audio separation with SAM Audio. Isolate any sound using natural language—professional-grade audio editing made simple for creators, researchers, and accessibility applications.

audio-to-audiosam-audio
text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Voice Design

fal-ai/minimax/voice-design

Design a personalized voice from a text description, and generate speech from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality text-to-speech.

speech
text-to-image
xAIREVIEW REQUIRED

Grok Imagine Image

xai/grok-imagine-image/quality/text-to-image

Grok Imagine Pro is an advanced AI model from xAI that creates high-quality visuals from text prompts and allows you to edit or analyze existing images.

stylizedtransformtypography
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Layered

fal-ai/qwen-image-layered

Qwen-Image-Layered is a model capable of decomposing an image into multiple RGBA layers.

qwenlayer
text-to-image
GoogleREVIEW REQUIRED

Gemini 3.1 Flash Image Preview

fal-ai/gemini-3.1-flash-image-preview

Gemini 3.1 Flash Image (a.k.a Nano Banana 2) is Google's new state-of-the-art fast image generation and editing model

vision
falREVIEW REQUIRED

Video Understanding

fal-ai/video-understanding

A video understanding model to analyze video content and answer questions about what's happening in the video based on user prompts.

utilityvision
text-to-audio
falREVIEW REQUIRED

ACE Step Prompt To Audio

fal-ai/ace-step/prompt-to-audio

Generate music from a simple prompt using ACE-Step

text-to-audiotext-to-music
image-to-video
PixVerseREVIEW REQUIRED

PixVerse Swap

fal-ai/pixverse/swap

Generate high quality video clips by swapping person, objects and background using Pixverse Swap.

text-to-speech
MiniMaxREVIEW REQUIRED

MiniMax Speech 2.6 [HD]

fal-ai/minimax/speech-2.6-hd

Generate speech from text prompts and different voices using the MiniMax Speech-2.6 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.

text-to-speech
video-to-video
falREVIEW REQUIRED

sync.so -- lipsync 1.9.0-beta

fal-ai/sync-lipsync

Generate realistic lipsync animations from audio using advanced algorithms for high-quality synchronization.

animationlip sync
text-to-audio
cassetteaiREVIEW REQUIRED

Sound Effects Generator

cassetteai/sound-effects-generator

Create stunningly realistic sound effects in seconds - CassetteAI's Sound Effects Model generates high-quality SFX up to 30 seconds long in just 1 second of processing time

soundsfxsound-effectscassetteai
image-to-video
lightricksREVIEW REQUIRED

LTX 2.5 Image to Video Fast

lightricks/ltx-2.5/image-to-video/fast

LTX-2.5 is Lightricks' open-source audio-video model. This endpoint animates a still image into video with synchronized audio in a single pass, in a speed-optimized mode for quick iteration.

stylizedtransformlip-sync
audio-to-audio
AlibabaREVIEW REQUIRED

Qwen 3 TTS - Clone Voice [1.7B]

fal-ai/qwen-3-tts/clone-voice/1.7b

Clone your voices using Qwen3-TTS Clone-Voice model with zero shot cloning capabilities and use it on text-to-speech models to create speeches of yours!

clone-voicevoice-clone
text-to-image
falREVIEW REQUIRED

Z Image Turbo Lora

fal-ai/z-image/turbo/lora

Text-to-Image endpoint with LoRA support for Z-Image Turbo, a super fast text-to-image model of 6B parameters developed by Tongyi-MAI.

z-imagelorafast
image-to-video
AlibabaREVIEW REQUIRED

Wan v2.6 Image to Video

wan/v2.6/image-to-video

Wan 2.6 image-to-video model.

image-to-video
text-to-speech
xAIREVIEW REQUIRED

xAI Text to Speech

xai/tts/v1

Generate speech with expressive and realistic voices from xAI

text-to-speech
falREVIEW REQUIRED

Chatterbox

fal-ai/chatterbox/text-to-speech

Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.

text-to-speech
text-to-image
AlibabaREVIEW REQUIRED

Qwen Image 2512

fal-ai/qwen-image-2512

Qwen Image 2512 is an improved version of Qwen Image with better text rendering, finer natural textures, and more realistic human generation.

qwen2512
image-to-image
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.7/edit

Transform and edit existing images with text-guided instructions using the WAN 2.7 model for creative image manipulation.

wanimage-to-imageimage-editing
3d-to-3d
MeshyREVIEW REQUIRED

Meshy Rigging Multi Animation

fal-ai/meshy/rigging/multi-animation

Meshy auto-rigs a humanoid 3D model fitting a skeleton and binding the mesh, then applies several motion presets from its animation library

stylizedtransform3D
image-to-image
Black Forest LabsREVIEW REQUIRED

Flux Pro Erase

fal-ai/flux-pro/v1/erase

Latest object erasing model from Black Forest Labs. Remove undesired objects, texts from images.

utilityediting
speech-to-speech
resemble-aiREVIEW REQUIRED

Chatterboxhd

resemble-ai/chatterboxhd/speech-to-speech

Transform voices using Resemble AI's Chatterbox. Convert audio to new voices or your own samples, with expressive results and built-in perceptual watermarking.

image-to-image
falREVIEW REQUIRED

CodeFormer

fal-ai/codeformer

Fix distorted or blurred photos of people with CodeFormer.

image-restorationfacesutility
text-to-video
KlingREVIEW REQUIRED

Kling O3 Text to Video [Pro]

fal-ai/kling-video/o3/pro/text-to-video

Generate realistic videos using Kling O3 from Kling Team!

text-to-video
image-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 First Last Frame to Video

blackforestlabs/flux-3/first-last-frame-to-video

FLUX 3 is Black Forest Labs' frontier video model. This endpoint generates the video between a defined start and end frame, interpolating a smooth, coherent transition from the first image to the last.

stylizedtransformlipsync
video-to-video
ByteDanceREVIEW REQUIRED

Bytedance Dreamactor V2

fal-ai/bytedance/dreamactor/v2

Transfer motion from a video to characters in an image using Dreamactor v2. Great performance for non-human and multiple characters

motion-controldreamactor