EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 4 · 28 per page
image-to-video
MiniMaxREVIEW REQUIRED

H3 Max Lip Sync Image to Video

minimax/h3-max/lip-sync/image-to-video

H3 Max Lip Sync generates a video from an image and supplied audio, synchronizing mouth movements to the soundtrack. It supports optional transcription guidance and output resolutions from 480p to 2K.

lipsyncanimationaudio
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [pro] Fill

fal-ai/flux-pro/v1/fill

FLUX.1 [pro] Fill is a high-performance endpoint for the FLUX.1 [pro] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.

editing
image-to-image
xAIREVIEW REQUIRED

Grok Imagine Image 2.0

xai/grok-imagine-image/v2.0/edit

Edit images with xAi's Grok Imagine 2.0 model.

xaigrokimage-editing
text-to-audio
falREVIEW REQUIRED

Stable Audio 2.5

fal-ai/stable-audio-25/text-to-audio

Generate high quality music and sound effects using Stable Audio 2.5 from StabilityAI

audio
text-to-speech
GoogleREVIEW REQUIRED

Gemini 3.8 Flash TTS

google/gemini-3.8-flash-tts

Generate expressive speech with Gemini 3.8 Flash TTS. Choose from 30 voices, direct delivery with style instructions, and create single-speaker narration or two-speaker dialogue.

text-to-speechaudiodialoguevoiceover
image-to-video
KlingREVIEW REQUIRED

Kling O3 Image to Video [Pro]

fal-ai/kling-video/o3/pro/image-to-video

Generate a video by taking a start frame and an end frame, animating the transition between them while following text-driven style and scene guidance.

image-to-video
image-to-video
xAIREVIEW REQUIRED

Grok Imagine Video

xai/grok-imagine-video/image-to-video

Generate videos from images with audio using xAI's Grok Imagine Video model.

grokxaiimage-to-videoi2v
image-to-3d
falREVIEW REQUIRED

Hunyuan 3D Pro Image to 3D

fal-ai/hunyuan-3d/v3.1/pro/image-to-3d

Generate 3D models from images with Hunyuan 3D Pro

3dhunyuanimage-to-3d
text-to-video
KlingREVIEW REQUIRED

Kling Video v3 Text to Video [Pro]

fal-ai/kling-video/v3/pro/text-to-video

Kling 3.0 Pro: Top-tier text-to-video with cinematic visuals, fluid motion, and native audio generation, with multi-shot support.

text-to-video
image-to-video
KlingREVIEW REQUIRED

Kling AI Avatar v2 Standard

fal-ai/kling-video/ai-avatar/v2/standard

Kling AI Avatar v2 Standard: Endpoint for creating avatar videos with realistic humans, animals, cartoons, or stylized characters

text-to-image
IdeogramREVIEW REQUIRED

Ideogram Text to Image

fal-ai/ideogram/v3

Generate high-quality images, posters, and logos with Ideogram V3. Features exceptional typography handling and realistic outputs optimized for commercial and creative use.

realismtypography
text-to-image
xAIREVIEW REQUIRED

Grok Imagine Image

xai/grok-imagine-image

Generate highly aesthetic images with xAI's Grok Imagine Image generation model.

xaigroktext-to-image
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.5 US Reference to Video

bytedance/seedance-2.5/us/reference-to-video

US-hosted ByteDance Seedance 2.5 generates video with native audio from up to 30 images, 10 videos, and 10 audio references. Supports reference-guided generation, video editing, and extension at 480p, 720p or 1080p.

stylizedtransformlipsync
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 9B

fal-ai/flux-2/klein/9b/edit

Image-to-image editing with FLUX.2 [klein] 9B from Black Forest Labs. Precise modifications using natural language descriptions and hex color control.

image-to-image
topazREVIEW REQUIRED

Topaz Upscale Image Precision

topaz/upscale/image/precision

Professional photo upscaling powered by Topaz Labs. Gigapixel precision models (Standard V2, High Fidelity, Low Resolution, CGI, Text Refine) enlarge images faithfully up to 4x. Best for photos that must stay true to the original.

upscaleimage
text-to-speech
ElevenLabsREVIEW REQUIRED

ElevenLabs TTS Turbo v2.5

fal-ai/elevenlabs/tts/turbo-v2.5

Generate high-speed text-to-speech audio using ElevenLabs TTS Turbo v2.5.

audio
image-to-image
GoogleREVIEW REQUIRED

Gemini 3 Pro Image Preview

fal-ai/gemini-3-pro-image-preview/edit

Gemini 3 Pro Image (a.k.a Nano Banana Pro) is Google's state-of-the-art high-fidelity image generation and editing model

realismtypography
image-to-video
xAIREVIEW REQUIRED

Grok Imagine Video 1.5

xai/grok-imagine-video/v1.5/image-to-video

Generate videos from images with audio using xAI's Grok Imagine 1.5 Video model.

stylizedtransformlipsync
image-to-image
GoogleREVIEW REQUIRED

Gemini 2.5 Flash Image

fal-ai/gemini-25-flash-image/edit

Google's famous original image generation and editing model, a.k.a Nano Banana

image-editing
text-to-image
kreaREVIEW REQUIRED

Krea 2 Large

krea/v2/large/text-to-image

Generate high-fidelity images from text with Krea 2 Large, supporting aspect ratio, creativity, seed controls, and optional style references.

text-to-imageimage-generationstyle-referencekrea
image-to-image
OpenAIREVIEW REQUIRED

GPT-Image 1.5

fal-ai/gpt-image-1.5/edit

GPT Image 1.5 generates high-fidelity images with strong prompt adherence, preserving composition, lighting, and fine-grained detail.

openaigpt-image
image-to-image
falREVIEW REQUIRED

Bria Expand Image

fal-ai/bria/expand

Bria Expand expands images beyond their borders in high quality. Trained exclusively on licensed data for safe and risk-free commercial use. Access the model's source code and weights: https://bria.ai/contact-us

outpainting
image-to-video
KlingREVIEW REQUIRED

Kling O3 Reference to Video [Pro]

fal-ai/kling-video/o3/pro/reference-to-video

Transform images, elements, and text into consistent, high-quality video scenes, ensuring stable character identity, object details, and environments.

reference-to-video
text-to-speech
GoogleREVIEW REQUIRED

Gemini 3.1 Flash Tts

fal-ai/gemini-3.1-flash-tts

Newest audio model from Google introduces granular audio tags that give you precise control to direct AI speech for expressive audio generation.

lipsyncavatar
video-to-video
falREVIEW REQUIRED

Sync Lipsync 2.0

fal-ai/sync-lipsync/v2

Generate realistic lipsync animations from audio using advanced algorithms for high-quality synchronization with Sync Lipsync 2.0 model

animationlip sync
image-to-image
RecraftREVIEW REQUIRED

Recraft Crisp Upscale

fal-ai/recraft/upscale/crisp

Enhances a given raster image using 'crisp upscale' tool, boosting resolution with a focus on refining small details and faces.

upscaling
text-to-image
metaREVIEW REQUIRED

Meta Muse Image Text to Image

meta/muse-image/text-to-image

Meta's Muse Image model has faithful instruction-following and exceptional visual fidelity, with fine details like text, plots, and QR codes rendered accurately.

realismtypographystylized
image-to-video
KlingREVIEW REQUIRED

Kling AI Avatar v2 Pro

fal-ai/kling-video/ai-avatar/v2/pro

Kling AI Avatar v2 Pro: The premium endpoint for creating avatar videos with realistic humans, animals, cartoons, or stylized characters