EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 14 · 28 per page
image-to-image
falREVIEW REQUIRED

SeedVR2

fal-ai/seedvr/upscale/image/seamless

Use SeedVR2 to upscale images, retaining seamless tiling

upscaleimage-to-imageseamlesstiling
text-to-video
xAIREVIEW REQUIRED

Grok Imagine Video 1.5 Text to Video

xai/grok-imagine-video/v1.5/text-to-video

Generate videos from prompts with audio using xAI's Grok Imagine 1.5 Video model.

stylizedtransformlipsync
image-to-image
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.7/pro/edit

Edit and transform images using text instructions with the WAN 2.7 Pro model for precise, professional-grade image modifications.

wanimage-editingpro
image-to-video
MiniMaxREVIEW REQUIRED

MiniMax Hailuo 2.3 [Pro] (Image to Video)

fal-ai/minimax/hailuo-2.3/pro/image-to-video

MiniMax Hailuo-2.3 Image To Video API (Pro, 1080p): Advanced image-to-video generation model with 1080p resolution

image-to-video
text-to-image
kreaREVIEW REQUIRED

Krea 2 Medium

krea/v2/medium/text-to-image

Generate high-quality images from text with Krea 2 Medium, supporting aspect ratio, creativity controls, seeds, and optional style references.

text-to-imageimage-generationstyle-referencekrea
image-to-image
falREVIEW REQUIRED

Image Outpaint

fal-ai/image-apps-v2/outpaint

Directional outpainting. Choose edges to expand. left, right, top, or center (uniform all sides). Only expanded areas are generated; an optional zoom-out pulls the frame back by the chosen amount.

outpainting
unknown
ByteDanceREVIEW REQUIRED

Seedance 2.5

bytedance/seedance-2.5/draft/complete

Draft completion endpoint for Seedance 2.5 - submit a draft id to regenerate the task at 1080p.

video-to-video
AlibabaREVIEW REQUIRED

Happy Horse Video Edit

alibaba/happy-horse/video-edit

HappyHorse video editing supports advanced video editing through natural language instructions. It allows for local or global editing of video elements using up to 5 reference images.

happy-horsevideo-editingvideo-to-video
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 US Reference to Video

bytedance/seedance-2.0/us/reference-to-video

US hosted version of ByteDance's most advanced reference-to-video model. Generate video from up to 9 images, 3 videos, and 3 audio clips with native audio and cinematic camera control.

stylizedtransformlipsync
text-to-speech
AlibabaREVIEW REQUIRED

Qwen 3 TTS - Voice Design [1.7B]

fal-ai/qwen-3-tts/voice-design/1.7b

Create custom voices using Qwen3-TTS Voice Design model and later use Clone Voice model to create your own voices!

text-to-speechvoice-design
video-to-video
GoogleREVIEW REQUIRED

Gemini Omni Flash

google/gemini-omni-flash/edit

Edits generated video across multiple conversational turns while preserving scene coherence. Applies iterative changes through natural-language instructions without regenerating the full sequence from scratch.

stylizedtransformlipsync
image-to-3d
falREVIEW REQUIRED

Hyper3D Rodin

fal-ai/hyper3d/rodin

Rodin by Hyper3D generates realistic and production ready 3D models from text or images.

stylized
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] with Controlnets and Loras

fal-ai/flux-general

A versatile endpoint for the FLUX.1 [dev] model that supports multiple AI extensions including LoRA, ControlNet conditioning, and IP-Adapter integration, enabling comprehensive control over image generation through various guidance methods.

loracontrolnetip-adapter
text-to-image
OpenAIREVIEW REQUIRED

gpt-image-1

fal-ai/gpt-image-1/text-to-image

OpenAI's latest image generation and editing model: gpt-1-image.

image-to-image
OpenAIREVIEW REQUIRED

gpt-image-1

fal-ai/gpt-image-1/edit-image

OpenAI's latest image generation and editing model: gpt-1-image.

text-to-video
KlingREVIEW REQUIRED

Kling O3 Text to Video [Standard]

fal-ai/kling-video/o3/standard/text-to-video

Generate realistic videos using Kling O3 from Kling Team!

text-to-video
video-to-video
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.7/edit-video

Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylizedtransformlipsync
text-to-speech
falREVIEW REQUIRED

Chatterbox

fal-ai/chatterbox/text-to-speech/multilingual

Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.

text-to-speechmultilingual
image-to-video
Luma AIREVIEW REQUIRED

Luma Ray 3.2 Image to Video

luma/agent/ray/v3.2/image-to-video

Luma Ray 3.2 animates a source image into cinematic motion guided by a text prompt, preserving the starting frame's look while controlling resolution, duration, and seamless looping.

stylizedtransformlipsync
text-to-image
RecraftREVIEW REQUIRED

Recraft V4 Pro

fal-ai/recraft/v4/pro/text-to-image

Recraft V4 was developed with designers to bring true visual taste to AI image generation. Built for brand systems and production-ready workflows, it goes beyond prompt accuracy — delivering stronger composition, refined lighting, realistic materials, and a cohesive aesthetic. The result is imagery shaped by professional design judgment, ready for immediate real-world use without additional post-processing.

text-to-image
image-to-3d
MeshyREVIEW REQUIRED

V7 Image to 3D

meshy/v7/image-to-3d

Turns a single image into a fully textured, PBR-ready 3D mesh with complete geometry, in game-ready Smart Topology at a target polygon count

stylizedtransform
video-to-video
MiniMaxREVIEW REQUIRED

H3 Max

minimax/h3-max/insert-video

Insert a new scene into an existing video with H3 Max. Guide the scene with a prompt, reference images, or reference videos, then return to the original footage.

video-to-videovideo-editingscene-insertion
image-to-video
falREVIEW REQUIRED

LTX 2.3 Video Pro

fal-ai/ltx-2.3/image-to-video

LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.

stylizedtransformlipsync
video-to-video
xAIREVIEW REQUIRED

Grok Imagine Video

xai/grok-imagine-video/edit-video

Edit videos using xAI's Grok Imagine

video-editv2vgrokxai
text-to-video
KlingREVIEW REQUIRED

Kling Video V3 Turbo Pro Text to Video

fal-ai/kling-video/v3/turbo/pro/text-to-video

Generate high quality 1080p videos using Kling's Turbo 3.0 model, with improved lipsync and multishot generation capabilities.

klingv31080pturbo
text-to-speech
GoogleREVIEW REQUIRED

Gemini 3.8 Flash Lite TTS

google/gemini-3.8-flash-lite-tts

Generate expressive speech with Gemini 3.8 Flash Lite TTS. Choose from 30 voices, guide delivery with style instructions, and create single-speaker narration or two-speaker dialogue.

text-to-speechaudiodialoguevoiceover
video-to-video
veedREVIEW REQUIRED

Video Background Removal

veed/video-background-removal/fast

Remove background from any video with people and objects. No green screen needed.

image-to-video
MiniMaxREVIEW REQUIRED

MiniMax Hailuo 02 [Pro] (Image to Video)

fal-ai/minimax/hailuo-02/pro/image-to-video

MiniMax Hailuo-02 Image To Video API (Pro, 1080p): Advanced image-to-video generation model with 1080p resolution