EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 8 · 28 per page
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 4B

fal-ai/flux-2/klein/4b

Text-to-image generation with FLUX.2 [klein] 4B from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.

text-to-video
GoogleREVIEW REQUIRED

Gemini Omni Flash 1.1 Text to Video

google/gemini-omni-flash/v1.1/text-to-video

Gemini Omni Flash 1.1 is Google's multimodal video model. This endpoint generates video with synchronized native audio from a text prompt, grounded in Gemini's real-world knowledge and physics understanding, with cinematic camera control expressed in natural language.

stylizedtransformlipsync
text-to-image
AlibabaREVIEW REQUIRED

Qwen Image 3 Text to Image

alibaba/qwen-image-3/text-to-image

Generates images from a text prompt at resolutions up to 2048×2048, with automatic prompt rewriting and prompt-guided resolution selection, building on Qwen's strength in complex text rendering and precise prompt adherence

stylizedtransformtypography
image-to-image
clarityaiREVIEW REQUIRED

Crystal Upscaler

clarityai/crystal-upscaler

An advanced image enhancement tool designed specifically for facial details and portrait photography, utilizing Clarity AI's upscaling technology.

image-to-image
text-to-speech
ElevenLabsREVIEW REQUIRED

Eleven v4 Turbo

elevenlabs/tts/eleven-v4-turbo

Generate speech with Eleven v4 Turbo from ElevenLabs. Choose a voice and control delivery with audio tags, stability, similarity settings, and IPA pronunciation.

audiotext-to-speechvoiceover
text-to-image
falREVIEW REQUIRED

Krea 2 Turbo

fal-ai/krea-2/turbo

Generate high-fidelity images from text in seconds with Krea 2 Turbo, the speed-optimized open-source version of Krea 2, preserving its aesthetic range for rapid ideation.

stylizedtransformtypography
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Kontext [dev]

fal-ai/flux-kontext/dev

Frontier image editing model.

text-to-audio
falREVIEW REQUIRED

Lyria 3 Pro

fal-ai/lyria3/pro

Lyria 3 Pro is the latest music model from Google

audiosfx
image-to-image
falREVIEW REQUIRED

Segment Anything Model 2

fal-ai/sam2/image

SAM 2 is a model for segmenting images and videos in real-time.

segmentationmaskreal-time
text-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Fast Text to Video

bytedance/seedance-2.0/fast/text-to-video

ByteDance's most advanced text-to-video model, fast tier. Lower latency and cost with cinematic output, native audio, multi-shot editing, and director-level camera control.

stylizedtransformlipsync
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 4B

fal-ai/flux-2/klein/4b/edit

Image-to-image editing with FLUX.2 [klein] 4B from Black Forest Labs. Precise modifications using natural language descriptions and hex color control.

video-to-video
GoogleREVIEW REQUIRED

Gemini Omni Flash 1.1 Edit

google/gemini-omni-flash/v1.1/edit

Gemini Omni Flash 1.1 is Google's multimodal video model. This endpoint edits video through natural-language instruction, applying the requested change while preserving the parts of the scene you want kept, and carrying character and scene consistency across successive edits.

utilityeditingtransform
image-to-video
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.7/image-to-video

Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylizedtransformlipsync
text-to-image
IdeogramREVIEW REQUIRED

Ideogram V4.0 Text to Image

ideogram/v4

Generate high-quality images, posters, and logos with Ideogram's latest V4.0q — producing crisp visuals with accurate text rendering, fine detail, and full creative control for polished, ready-to-use designs.

realismtypographystylized
text-to-video
GoogleREVIEW REQUIRED

Veo3.1 Lite Text to Video

fal-ai/veo3.1/lite

Veo 3.1 Lite balances practical utility with professional capabilities, supporting Text-to-Video and Image-to-Video

stylizedtransformlipsync
image-to-3d
falREVIEW REQUIRED

Hyper3D - Rodin V2.5 - Image to 3D

fal-ai/hyper3d/rodin/v2.5

Rodin V2.5 by Hyper3D generates realistic and production ready 3D models from text or images.

image-to-3d
text-to-audio
MiniMaxREVIEW REQUIRED

Minimax Music

fal-ai/minimax-music/v2

Generate music from text prompts using the MiniMax Music 2.0 model, which leverages advanced AI techniques to create high-quality, diverse musical compositions.

musicaudio
text-to-image
AlibabaREVIEW REQUIRED

Qwen Image

fal-ai/qwen-image

Qwen-Image is an image generation foundation model in the Qwen series that achieves significant advances in complex text rendering and precise image editing.

text-to-image
training
Black Forest LabsREVIEW REQUIRED

Train Flux LoRA

fal-ai/flux-lora-fast-training

Train styles, people and other subjects at blazing speeds.

lorapersonalization
image-to-image
ByteDanceREVIEW REQUIRED

Seedream 5.0 Pro Layerize

bytedance/seedream/v5/pro/layerize

Splits a finished image into independent, editable transparent-PNG layers — background plus separate elements, from a text description, returning 2 to 17 layers per call for non-destructive reuse in design tools.

utilityediting
text-to-image
ByteDanceREVIEW REQUIRED

Seedream

bytedance/seedream/v5/flash/text-to-image

Seedream 5.0 Flash is a fast image generation and editing model, built for workflows where speed and budget matter.

realismtypographystylized
text-to-image
GoogleREVIEW REQUIRED

Gemini 3 Pro Image Preview

fal-ai/gemini-3-pro-image-preview

Gemini 3 Pro Image (a.k.a Nano Banana Pro) is Google's state-of-the-art high-fidelity image generation and editing model

realismtypography
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Mini

bytedance/seedance-2.0/mini/reference-to-video

Seedance 2.0 Mini is a faster version of Seedance 2.0 that brings great performance and high generation speed at a lower cost.

stylizedtransformlipsync
image-to-image
xAIREVIEW REQUIRED

Grok Imagine Image Editing Quality

xai/grok-imagine-image/quality/edit

Grok Imagine Pro is an advanced AI model from xAI that creates high-quality visuals from text prompts and allows you to edit or analyze existing images.

stylizedtransformtypography
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 Mini Image to Video

bytedance/seedance-2.0/mini/image-to-video

Seedance 2.0 Mini is a faster version of Seedance 2.0 that brings great performance and high generation speed at a lower cost.

stylizedtransformlipsync
text-to-video
xAIREVIEW REQUIRED

Grok Imagine Video

xai/grok-imagine-video/text-to-video

Generate videos with audio from text using Grok Imagine Video.

xaigrokt2vtext-to-video
image-to-3d
MeshyREVIEW REQUIRED

Meshy 7.1 Image to 3D

meshy/v7.1/image-to-3d

Meshy 7.1 generates 3D models from a single image, with standard, low-poly, and Smart Topology modes, optional textures and PBR maps, and geometry resolution up to 4K.

3dimage-to-3dtexturespbr
vision
falREVIEW REQUIRED

NSFW Filter

fal-ai/imageutils/nsfw

Predict the probability of an image being NSFW.

filtersafetyutility