EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 15 · 28 per page
training
Black Forest LabsREVIEW REQUIRED

Train Flux LoRAs For Portraits

fal-ai/flux-lora-portrait-trainer

FLUX LoRA training optimized for portrait generation, with bright highlights, excellent prompt following and highly detailed results.

lorapersonalization
text-to-video
falREVIEW REQUIRED

Wan 2.5 Text to Video

fal-ai/wan-25-preview/text-to-video

Wan 2.5 text-to-video model.

image-to-video
falREVIEW REQUIRED

Wan-2.1 Image-to-Video

fal-ai/wan-i2v

Wan-2.1 is a image-to-video model that generates high-quality videos with high visual quality and motion diversity from images

image to videomotion
vision
falREVIEW REQUIRED

Moondream3 Preview [Detect]

fal-ai/moondream3-preview/detect

Moondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.

Vision
image-to-image
falREVIEW REQUIRED

Hunyuan Image

fal-ai/hunyuan-image/v3/instruct/edit

Image editing endpoint for Hunyuan Image 3.0 Instruct.

tencenthunyuan-imageinstructedit
text-to-image
KlingREVIEW REQUIRED

Kling Image

fal-ai/kling-image/o3/text-to-image

Kling Omni 3: Top-tier text-to-image with flawless consistency.

text-to-image
image-to-image
ByteDanceREVIEW REQUIRED

Seedream

bytedance/seedream/v5/flash/layerize

Seedream 5.0 Flash is a fast image generation and editing model, built for workflows where speed and budget matter.

utilityediting
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V6 Transition

fal-ai/pixverse/v6/transition

Pixverse's latest v6 Model.

image-to-videofirst-frame-last-frametransition
image-to-video
falREVIEW REQUIRED

Ffmpeg Api Images to Video

fal-ai/ffmpeg-api/images-to-video

A fal.ai endpoint that stitches an ordered list of images into an MP4 video by holding each image for a specified number of frames at a configurable frame rate

utilityediting
text-to-3d
tripo3dREVIEW REQUIRED

Tripo H3.1 Text to 3D

tripo3d/h3.1/text-to-3d

Generate 3D models from text descriptions using Tripo H3.1.

3dtext-to-3d3d-generationtripo
text-to-speech
falREVIEW REQUIRED

Index TTS 2.0

fal-ai/index-tts-2/text-to-speech

Generate natural, clear speeches using Index TTS 2.0 from IndexTeam

text-to-speech
image-to-image
IdeogramREVIEW REQUIRED

Ideogram V3 Character

fal-ai/ideogram/character

Generate consistent character appearances across multiple images. Maintain facial features, proportions, and distinctive traits for cohesive storytelling and branding

character-consistency
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit 2511

fal-ai/qwen-image-edit-2511/lora

Endpoint for Qwen's Image Editing 2511 model with LoRa support.

stylizedtransformlora
image-to-3d
falREVIEW REQUIRED

Sam 3

fal-ai/sam-3/3d-body

SAM 3D allows for accurate 3D reconstruction of human body shape and position from a single image.

3dhumanpose
text-to-audio
falREVIEW REQUIRED

Stable Audio 3 Small SFX Text to Audio

fal-ai/stable-audio-3/small/sfx/text-to-audio

Stable Audio 3 Small SFX is a 459 million parameter latent diffusion model that generates high-quality sound effects from text prompts, designed for on-device deployment on mobile phones and consumer laptops.

sfxsound-effectson-device
text-to-image
falREVIEW REQUIRED

Stable Diffusion XL Lightning

fal-ai/fast-lightning-sdxl

Run SDXL at the speed of light

diffusionlightningreal-time
vision
falREVIEW REQUIRED

Moondream 3 Preview [Query]

fal-ai/moondream3-preview/query

Moondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.

Vision
audio-to-audio
KlingREVIEW REQUIRED

Kling Video Create Voice

fal-ai/kling-video/create-voice

Create Voices to be used with Kling Models Voice Control

text-to-image
RecraftREVIEW REQUIRED

Recraft V4.1 Text to Image Pro

fal-ai/recraft/v4.1/pro/text-to-image

Recraft V4.1 Pro pushes the V4.1 model into high-resolution territory — up to 2048×2048 and ultra-wide formats. Made for hero imagery, campaign work, and print, it preserves the same design taste at sizes ready for the final deliverable.

stylizedtransformtypography
text-to-image
RecraftREVIEW REQUIRED

Recraft V4 (Vector)

fal-ai/recraft/v4/text-to-vector

Recraft V4 was developed with designers to bring true visual taste to AI image generation. Built for brand systems and production-ready workflows, it goes beyond prompt accuracy — delivering stronger composition, refined lighting, realistic materials, and a cohesive aesthetic. The result is imagery shaped by professional design judgment, ready for immediate real-world use without additional post-processing.

text-to-imagetext-to-vector
image-to-video
KlingREVIEW REQUIRED

Kling Video

fal-ai/kling-video/o3/4k/reference-to-video

Kling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling

stylizedtransformlipsync
image-to-video
pixelcutREVIEW REQUIRED

Pixelcut Looping Video

pixelcut/looping-video

Turn one product photo into a seamless 5 to 15 second video loop with a locked camera, subtle ambient motion or a full 360° spin.

image-to-videoloopinge-commerceproduct
text-to-video
MiniMaxREVIEW REQUIRED

MiniMax Hailuo 02 [Standard] (Text to Video)

fal-ai/minimax/hailuo-02/standard/text-to-video

MiniMax Hailuo-02 Text To Video API (Standard, 768p): Advanced video generation model with 768p resolution

image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit

fal-ai/qwen-image-edit/inpaint

Inpainting Endpoint for the Qwen Edit Image editing model.

image-to-imageinpaintingqwen-image
image-to-image
topazREVIEW REQUIRED

Topaz Upscale Image Transparent

topaz/upscale/image/transparent

Professional transparent-image upscaling powered by Topaz Labs. Preserves the alpha channel end to end with PNG output. Best for logos, stickers and assets with transparency.

upscaleimage
video-to-audio
soniloREVIEW REQUIRED

V1.1 Video to Sound Effects

sonilo/v1.1/video-to-sound-effects

Analyzes a video and generates synchronized, royalty-free sound effects timed to visible actions. Returns the generated sound-effects audio track for commercial use.

sfxaudioeffects
text-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 Text To Video Draft

blackforestlabs/flux-3/text-to-video/draft

FLUX.3 is Black Forest Labs' frontier audio/video model. Generate fast, low-cost draft previews from a text prompt, with a reusable draft cache for full-quality enhancement.

stylizedtransformlipsync
image-to-image
KlingREVIEW REQUIRED

Kling Image

fal-ai/kling-image/v3/image-to-image

Kling Image V3: Latest kling image model

image-to-image