EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 42 · 28 per page
image-to-image
falREVIEW REQUIRED

Bria

fal-ai/bria/reimagine

Structure Reference allows generating new images while preserving the structure of an input image, guided by text prompts. Perfect for transforming sketches, illustrations, or photos into new illustrations. Trained exclusively on licensed data for safe and risk-free commercial use.

image-to-image
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.2-a14b/image-to-image

Wan 2.2's 14B model edit high-resolution, photorealistic images with powerful prompt understanding and fine-grained visual detail

image-to-image
image-to-video
falREVIEW REQUIRED

LTX-2.3 22B

fal-ai/ltx-2.3-22b/image-to-video/lora

Generate video with audio from images using LTX-2.3 and custom LoRA

image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit 2509 Lora

fal-ai/qwen-image-edit-2509-lora

LoRA endpoint for the Qwen Image Edit 2509 model.

image-to-imageimage-editing
image-to-video
falREVIEW REQUIRED

Kandinsky 6.0 Lite

fal-ai/kandinsky6-lite/image-to-video

Kandinsky 6.0 Lite animates a single image into a short video clip, with fast, lightweight inference and prompt-guided motion.

image-to-videokandinskylightweight
text-to-image
falREVIEW REQUIRED

Bagel

fal-ai/bagel

Bagel is a 7B parameter from Bytedance-Seed multimodal model that can generate both text and images.

text-to-imagemultimodal
image-to-video
falREVIEW REQUIRED

Wan-2.1 Image-to-Video with LoRAs

fal-ai/wan-i2v-lora

Add custom LoRAs to Wan-2.1 is a image-to-video model that generates high-quality videos with high visual quality and motion diversity from images

image to videomotionlora
video-to-video
falREVIEW REQUIRED

Controlfoley

fal-ai/controlfoley

Foley Control is a video-to-audio model that automatically generates synchronized sound effects for videos, using text prompts to shape the type of sound while matching the timing and action on screen.

stylizedtransformlipsync
text-to-audio
falREVIEW REQUIRED

Kokoro TTS (Hindi)

fal-ai/kokoro/hindi

A fast and expressive Hindi text-to-speech model with clear pronunciation and accurate intonation.

speech
image-to-video
falREVIEW REQUIRED

Cosmos Predict 2.5 2B

fal-ai/cosmos-predict-2.5/image-to-video

Generate video from text and images using NVIDIA's 2B Cosmos Post-Trained Model

text-to-audio
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/text-to-audio

Text to Audio high-quality using LTX-2.3

text-to-audio
text-to-video
falREVIEW REQUIRED

Cosmos Predict 2.5 2B

fal-ai/cosmos-predict-2.5/text-to-video

Generate video from text using NVIDIA's 2B Cosmos Post-Trained Model

image-to-video
falREVIEW REQUIRED

Vidu Template to Video

fal-ai/vidu/template-to-video

Vidu Template to Video lets you create different effects by applying motion templates to your images.

motiontemplate
text-to-video
falREVIEW REQUIRED

Vidu Text to Video

fal-ai/vidu/q1/text-to-video

Vidu Q1 Text to Video generates high-quality 1080p videos with exceptional visual quality and motion diversity

stylizedtransform
video-to-video
falREVIEW REQUIRED

LTX-2.3 22B

fal-ai/ltx-2.3-22b/reference-video-to-video

Generate video with audio from reference video, text and images using LTX-2.3

text-to-image
falREVIEW REQUIRED

Z-Image Turbo Seamless Tiling Lora

fal-ai/z-image/turbo/tiling/lora

Generate seamlessly tiling photorealistic images from text using Z-Image Turbo and custom LoRA

z-imageturboseamlesstiling
text-to-video
falREVIEW REQUIRED

Wan-2.1 Pro Text-to-Video

fal-ai/wan-pro/text-to-video

Wan-2.1 Pro is a premium text-to-video model that generates high-quality 1080p videos at 30fps with up to 6 seconds duration, delivering exceptional visual quality and motion diversity from text prompts

text to videomotion
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V4 Effects

fal-ai/pixverse/v4/effects

Generate high quality video clips with different effects using PixVerse v4

image-to-video
image-to-image
falREVIEW REQUIRED

Stepx Edit2

fal-ai/stepx-edit2

Image-to-image editing with Step1X-Edit v2 from StepFun. Reasoning-enhanced modifications through a thinking–editing–reflection loop with MLLM world knowledge for abstract instruction comprehension.

3d-to-3d
hitem3dREVIEW REQUIRED

Hi3D Texture

hitem3d/hi3d/texture

Texture an existing geometry mesh using a reference image with Hi3D.

3d-to-3d
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V4.5 Image To Video Fast

fal-ai/pixverse/v4.5/image-to-video/fast

Generate fast high quality video clips from text and image prompts using PixVerse v4.5

stylizedtransform
image-to-video
falREVIEW REQUIRED

Kandinsky 6.0 Pro

fal-ai/kandinsky6-pro/image-to-video

Kandinsky 6.0 Pro turns a start frame into a high-resolution, cinematic video clip with controllable motion and strong visual fidelity.

image-to-videokandinskylightweight
image-to-image
briaREVIEW REQUIRED

Fibo Edit [Rewrite Text]

bria/fibo-edit/rewrite_text

Precisely rewrite text inside images while preserving typography, fonts, and layout. High-quality, brand-safe edits trained exclusively on licensed data for safe commercial use.

briafibo-edittext-rewritingimage-editing
vision
falREVIEW REQUIRED

Sa2VA 4B Image

fal-ai/sa2va/4b/image

Sa2VA is an MLLM capable of question answering, visual prompt understanding, and dense object segmentation at both image and video levels

multimodalvision
image-to-image
falREVIEW REQUIRED

Leffa Pose Transfer

fal-ai/leffa/pose-transfer

Leffa Pose Transfer is an endpoint for changing pose of an image with a reference image.

poseutility
video-to-video
falREVIEW REQUIRED

Krea Wan 14B

fal-ai/krea-wan-14b/video-to-video

Superfast video model based on Wan 2.1 14b by Krea, excelling at real-time video-editing.

text-to-video
PixVerseREVIEW REQUIRED

PixVerse V4.5 Text To Video Fast

fal-ai/pixverse/v4.5/text-to-video/fast

Generate high quality and fast video clips from text and image prompts using PixVerse v4.5 fast

stylizedtransform