EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 43 · 28 per page
image-to-video
falREVIEW REQUIRED

Vidu Reference to Video

fal-ai/vidu/reference-to-video

Vidu Reference to Video creates videos by using a reference images and combining them with a prompt.

motionreference
image-to-image
briaREVIEW REQUIRED

Bria Product Dimensions

bria/product-dimensions

Bria Product Dimensions turns one product photo and its measurements into a marketplace-ready dimension image with callout lines, labels, and weight or capacity readouts

stylizedtransformtypography
image-to-image
falREVIEW REQUIRED

Image Preprocessors

fal-ai/image-preprocessors/hed

Holistically-Nested Edge Detection (HED) preprocessor.

preprocessdetectionutilitycontrolnet
audio-to-audio
falREVIEW REQUIRED

Stable Audio 25

fal-ai/stable-audio-25/inpaint

Generate high quality music and sound effects using Stable Audio 2.5 from StabilityAI

audio
text-to-video
falREVIEW REQUIRED

Wan-2.1 Text-to-Video with LoRAs

fal-ai/wan-t2v-lora

Add custom LoRAs to Wan-2.1 is a text-to-video model that generates high-quality videos with high visual quality and motion diversity from images

"text to video""motion""lora"
image-to-3d
falREVIEW REQUIRED

Hunyuan3D

fal-ai/hunyuan3d/v2/multi-view/turbo

Generate 3D models from your images using Hunyuan 3D. A native 3D generative model enabling versatile and high-quality 3D asset creation.

stylized
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] with Controlnets and Loras

fal-ai/flux-general/differential-diffusion

A specialized FLUX endpoint combining differential diffusion control with LoRA, ControlNet, and IP-Adapter support, enabling precise, region-specific image transformations through customizable change maps.

loracontrolnetip-adapter
video-to-video
moonvalleyREVIEW REQUIRED

Marey Realism V1.5

moonvalley/marey/motion-transfer

Pull motion from a reference video and apply it to new subjects or scenes.

video-to-video
falREVIEW REQUIRED

AMT Interpolation

fal-ai/amt-interpolation

Interpolate between video frames

interpolationediting
image-to-video
falREVIEW REQUIRED

LongCat Video

fal-ai/longcat-video/image-to-video/480p

Generate long videos from images using LongCat Video

text-to-image
falREVIEW REQUIRED

Sensenova U1 Infographic

fal-ai/sensenova-u1-infographic

Generate Infographic Image with Sensenova U1

text-to-video
falREVIEW REQUIRED

Krea Wan 14b- Text to Video

fal-ai/krea-wan-14b/text-to-video

Fast Text-to-Video endpoint for Krea's Wan 14b model.

text to videofast
text-to-image
AlibabaREVIEW REQUIRED

Wan

fal-ai/wan/v2.2-5b/text-to-image

Wan 2.2's 5B model generates high-resolution, photorealistic images with powerful prompt understanding and fine-grained visual detail

text-to-video
KlingREVIEW REQUIRED

Kling LipSync Text-to-Video

fal-ai/kling-video/lipsync/text-to-video

Kling LipSync is a text-to-video model that generates realistic lip movements from text input.

text to videolipsync
text-to-audio
falREVIEW REQUIRED

Kokoro TTS (Mandarin Chinese)

fal-ai/kokoro/mandarin-chinese

A highly efficient Mandarin Chinese text-to-speech model that captures natural tones and prosody.

speech
text-to-video
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/text-to-video

Generate high-quality video with audio from text using LTX-2.3

text-to-videovideo
video-to-video
falREVIEW REQUIRED

LTX-2.3 22B

fal-ai/ltx-2.3-22b/video-to-video

Generate video with audio from videos using LTX-2.3

image-to-image
falREVIEW REQUIRED

Hidream I1 Full

fal-ai/hidream-i1-full/image-to-image

HiDream-I1 full is a new open-source image generative foundation model with 17B parameters that achieves state-of-the-art image generation quality within seconds.

image-to-imagehidream
text-to-image
briaREVIEW REQUIRED

Fibo Bbq Preview

bria/fibo-bbq-preview/generate

A preview to the next level of control of Text-to-Image models.

image-to-image
IdeogramREVIEW REQUIRED

Ideogram V4.0q Image to Image LoRA

ideogram/v4/image-to-image/lora

Ideogram V4.0q Image-to-Image LoRA applies a custom-trained LoRA on top of an input image, steering edits toward a specific style, subject, or brand identity while keeping the source composition intact.

stylizedtransformrealism
image-to-image
falREVIEW REQUIRED

Image Editing Wojak Style

fal-ai/image-editing/wojak-style

Transform your photos into wojak style while keeping the original characters likeness

stylizedtransform
audio-to-audio
falREVIEW REQUIRED

Stable Audio 3

fal-ai/stable-audio-3/small/sfx/audio-to-audio

Stable Audio 3 Small SFX audio-to-audio is a 459 million parameter latent diffusion model that transforms input audio into new sound-effect variations guided by text prompts.

sfxsound-effectsstyle-transfer
image-to-image
Luma AIREVIEW REQUIRED

Luma Photon

fal-ai/luma-photon/modify

Edit images from your prompts using Luma Photon. Photon is the most creative, personalizable, and intelligent visual models for creatives, bringing a step-function change in the cost of high-quality image generation.

image-to-image
video-to-video
mirelo-aiREVIEW REQUIRED

Mirelo SFX

mirelo-ai/sfx-v1/video-to-video

Generate synced sounds for any video, and return it with its new sound track (like MMAudio)

video-to-videosfx
text-to-image
falREVIEW REQUIRED

Latent Consistency Models (v1.5/XL)

fal-ai/fast-lcm-diffusion

Run SDXL at the speed of light

lcmdiffusionturboreal-time
text-to-video
PixVerseREVIEW REQUIRED

PixVerse V4 Text To Video

fal-ai/pixverse/v4/text-to-video

Generate high quality video clips from text and image prompts using PixVerse v4

text-to-json
briaREVIEW REQUIRED

Fibo

bria/fibo/generate/structured_prompt

Structured Prompt Generation endpoint for Fibo, Bria's SOTA Open source model.

briafibostructured-prompting
text-to-image
falREVIEW REQUIRED

Bitdance

fal-ai/bitdance

Image generation with BitDance. Fast, high-resolution photorealistic images using an autoregressive LLM— for efficient, high-quality results.

text-to-image