EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 25 · 28 per page
text-to-video
AlibabaREVIEW REQUIRED

Happy Horse 1.1 Text to Video

alibaba/happy-horse/v1.1/text-to-video

Happy Horse 1.1 is Alibaba's #1-ranked video model. This text-to-video endpoint generates 1080p video with synchronized native audio and multilingual lip-sync from a text prompt alone.

happy-horsevideotext
vision
falREVIEW REQUIRED

Florence 2 Large OCR

fal-ai/florence-2-large/ocr

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

ocrmultimodalvision
image-to-video
falREVIEW REQUIRED

LTX-Video 13B 0.9.8 Distilled

fal-ai/ltxv-13b-098-distilled/image-to-video

Generate long videos from prompts and images using LTX Video-0.9.8 13B Distilled and custom LoRA

videoltx-videoimage-to-video
text-to-image
AlibabaREVIEW REQUIRED

Wan v2.6 Text to Image

wan/v2.6/text-to-image

Wan 2.6 text-to-image model.

text-to-image
video-to-video
falREVIEW REQUIRED

Wan VACE 14B

fal-ai/wan-vace-14b/depth

VACE is a video generation model that uses a source image, mask, and video to create prompted videos with controllable sources.

image-to-videovideo-to-videotext-to-video
video-to-video
soniloREVIEW REQUIRED

V1.1 Video to Video Music

sonilo/v1.1/video-to-video-music

Generates perfectly synced music for any video. Return a licensed music soundtrack ready for commercial use (optional preservation of the original speech in video)

musiceditingrestoration
video-to-video
PixVerseREVIEW REQUIRED

PixVerse V6 Extend

fal-ai/pixverse/v6/extend

Pixverse's latest v6 Model.

video-to-videoextend
text-to-video
Luma AIREVIEW REQUIRED

Luma Ray 3.2 Text to Video

luma/agent/ray/v3.2/text-to-video

Luma Ray 3.2 generates cinematic video from a text prompt, with control over resolution, duration, and seamless looping, plus reference images to lock in subject and style.

stylizedtransformlipsync
vision
falREVIEW REQUIRED

Moondream2

fal-ai/moondream2/object-detection

Moondream2 is a highly efficient open-source vision language model that combines powerful image understanding capabilities with a remarkably small footprint.

image-to-image
speech-to-text
falREVIEW REQUIRED

Cohere Transcribe

fal-ai/cohere-transcribe

Cohere Transcribe turns your business audio into accurate text, ready for search, analytics, and automation

speechtranscribestt
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V5 Transition

fal-ai/pixverse/v5/transition

Create seamless transition between images using PixVerse v5

stylizedtransform
video-to-video
falREVIEW REQUIRED

LTX Video 2.3 Pro

fal-ai/ltx-2.3/extend-video

LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.

stylizedtransformlipsync
text-to-3d
MeshyREVIEW REQUIRED

Meshy 7.1 Text to 3D

meshy/v7.1/text-to-3d

Meshy 7.1 generates 3D models from text prompts, with untextured preview and textured full modes, standard, low-poly, and Smart Topology options, and geometry resolution up to 4K.

3dtext-to-3dtexturespbr
image-to-image
falREVIEW REQUIRED

FILM

fal-ai/film

Interpolate images with FILM - Frame Interpolation for Large Motion

interpolation
video-to-video
falREVIEW REQUIRED

Heygen

fal-ai/heygen/v2/translate/speed

Heygen Translate Model with Extreme Speed

video-to-video
image-to-image
falREVIEW REQUIRED

DRCT-Super-Resolution

fal-ai/drct-super-resolution

Upscale your images with DRCT-Super-Resolution.

upscalinghigh-res
image-to-image
falREVIEW REQUIRED

Z Image Turbo Controlnet

fal-ai/z-image/turbo/controlnet

Generate images from text and edge, depth or pose images using Z-Image Turbo, Tongyi-MAI's super-fast 6B model.

training
falREVIEW REQUIRED

Z Image Trainer

fal-ai/z-image-trainer

Train LoRAs on Z-Image Turbo, a super fast text-to-image model of 6B parameters developed by Tongyi-MAI.

turboz-imagefasttrainer
image-to-image
falREVIEW REQUIRED

Face Retoucher

fal-ai/retoucher

Automatically retouches faces to smooth skin and remove blemishes.

editing
image-to-image
Black Forest LabsREVIEW REQUIRED

Flux 2 Lora Gallery

fal-ai/flux-2-lora-gallery/multiple-angles

Generates same object from different angles (azimuth/elevation)

stylizedtransform
audio-to-video
AlibabaREVIEW REQUIRED

Wan-2.2 Speech-to-Video 14B

fal-ai/wan/v2.2-14b/speech-to-video

Wan-S2V is a video model that generates high-quality videos from static images and audio, with realistic facial expressions, body movements, and professional camera work for film and television applications

audio-to-videotalking-head
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit

fal-ai/qwen-image-edit/image-to-image

Image to Image Endpoint for Qwen's Image Editing model. Has superior text editing capabilities.

stylizedtransform
image-to-video
AlibabaREVIEW REQUIRED

Wan v2.2 A14B Image-to-Video A14B with LoRAs

fal-ai/wan/v2.2-a14b/image-to-video/lora

Wan-2.2 image-to-video is a video model that generates high-quality videos with high visual quality and motion diversity from text prompts and images. This endpoint supports LoRAs made for Wan 2.2

image-to-videomotionlora
image-to-image
falREVIEW REQUIRED

Phota Enhance

fal-ai/phota/enhance

Enhance images while preserving identities with Phota

stylizedtransformtypographyphota
text-to-3d
tripo3dREVIEW REQUIRED

Tripo P2 Text to 3D

tripo3d/p2/text-to-3d

Tripo P2 generates 3D models from a text prompt, with optional PBR textures, adjustable face counts, and triangle or quad mesh topology.

3dtext-to-3d3d-generationlow-poly
text-to-video
falREVIEW REQUIRED

Hunyuan Video

fal-ai/hunyuan-video

Hunyuan Video is an Open video generation model with high visual quality, motion diversity, text-video alignment, and generation stability. This endpoint generates videos from text descriptions.

motion
text-to-image
RecraftREVIEW REQUIRED

Recraft V4 Styles Pro Text to Image

recraft/v4/style/pro/text-to-image

Generates raster images that hold a consistent style, from either a saved style ID or reference images attached directly.

stylizedtransformediting