EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 28 · 28 per page
text-to-audio
MiniMaxREVIEW REQUIRED

MiniMax (Hailuo AI) Music v1.5

fal-ai/minimax-music/v1.5

Generate music from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality, diverse musical compositions.

music
image-to-image
falREVIEW REQUIRED

Glm Image

fal-ai/glm-image/image-to-image

Create high-quality images with accurate text rendering and rich knowledge details—supports editing, style transfer, and maintaining consistent characters across multiple images.

image-to-image
text-to-video
PixVerseREVIEW REQUIRED

PixVerse C1 Text To Video

fal-ai/pixverse/c1/text-to-video

Generate film-grade videos from text prompts with native audio, up to 1080p and 15 seconds, using PixVerse C1.

video-generationtext-to-videopixversecinematic
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 4B LoRA

fal-ai/flux-2/klein/4b/edit/lora

Image-to-image editing with FLUX.2 [klein] 4B from Black Forest Labs and custom LoRA. Precise modifications using natural language descriptions and hex color control.

text-to-image
RecraftREVIEW REQUIRED

Recraft V4 Styles Text to Vector

recraft/v4/style/text-to-vector

Generates vector images that hold a consistent style, from either a saved style ID or reference images attached directly.

stylizedtransformediting
image-to-image
falREVIEW REQUIRED

Image2Pixel

fal-ai/image2pixel

Turn images into pixel-perfect retro art

post-processingpixel-art
text-to-image
RecraftREVIEW REQUIRED

Recraft V4.1 Text to Image Utility

fal-ai/recraft/v4.1/utility/text-to-image

Recraft V4.1 Utility is a faster, lighter variant of V4.1 made for high-volume creative workflows. Ideal for ideation, A/B exploration, and content pipelines, it keeps Recraft's design sensibility while optimizing for throughput and cost.

stylizedtransformtypography
text-to-image
falREVIEW REQUIRED

Ernie Image

fal-ai/ernie-image

High-quality text-to-image model by Baidu. Supports English, Chinese, and Japanese prompts with built-in prompt expansion.

realismchinesemultilingualportrait
image-to-image
falREVIEW REQUIRED

Hair Change

fal-ai/image-apps-v2/hair-change

Change hairstyles and hair colors in photos realistically.

hair-editstyle-change
image-to-image
falREVIEW REQUIRED

Image Editing Hair Change

fal-ai/image-editing/hair-change

Experiment with different hairstyles, from bald to any style you can imagine, while maintaining natural lighting and realistic results.

stylizedtransform
image-to-image
briaREVIEW REQUIRED

Fibo Edit [Relight]

bria/fibo-edit/relight

Precise, controllable photo re-lighting with structured text inputs. Apply natural lighting styles, soften harsh shadows, and transform scene illumination - production-ready and trained exclusively on licensed data.

briafibo-editrelightingjson
text-to-image
nvidiaREVIEW REQUIRED

Cosmos 3 Super

nvidia/cosmos-3-super/text-to-image

Cosmos3 is a collection of Omnimodal world models capable of generating dynamic, high-quality video, image, audio, and action commands from combinations of text, image, video, and action trajectory inputs.

stylizedtransformrealism
text-to-image
falREVIEW REQUIRED

Sana Sprint

fal-ai/sana/sprint

Sana Sprint is a text-to-image model capable of generating 4K images with exceptional speed.

text to image4khigh-speed
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] Control LoRA Canny

fal-ai/flux-control-lora-canny

FLUX Control LoRA Canny is a high-performance endpoint that uses a control image to transfer structure to the generated image, using a Canny edge map.

lorastyle transfer
video-to-video
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/inpaint

Inpaint high-quality video using LTX-2.3

image-to-video
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/image-to-video

Generate high-quality video with audio from images using LTX-2.3

image-to-video
image-to-video
falREVIEW REQUIRED

Hunyuan Video Image-to-Video Inference

fal-ai/hunyuan-video-image-to-video

Image to Video for the high-quality Hunyuan Video I2V model.

motion
video-to-video
falREVIEW REQUIRED

Wan VACE 14B

fal-ai/wan-vace-14b

VACE is a video generation model that uses a source image, mask, and video to create prompted videos with controllable sources.

image-to-videovideo-to-videotext-to-video
text-to-speech
falREVIEW REQUIRED

Dia

fal-ai/dia-tts

Dia directly generates realistic dialogue from transcripts. Audio conditioning enables emotion control. Produces natural nonverbals like laughter and throat clearing.

text-to-speech
text-to-image
Luma AIREVIEW REQUIRED

Luma Photon

fal-ai/luma-photon

Generate images from your prompts using Luma Photon. Photon is the most creative, personalizable, and intelligent visual models for creatives, bringing a step-function change in the cost of high-quality image generation.

vision
falREVIEW REQUIRED

Florence 2 Large Detailed Caption

fal-ai/florence-2-large/detailed-caption

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

captioningmultimodalvision
text-to-image
falREVIEW REQUIRED

Playground v2.5

fal-ai/playground-v25

State-of-the-art open-source model in aesthetic quality

artisticstyle
vision
falREVIEW REQUIRED

Moondream2

fal-ai/moondream2

Moondream2 is a highly efficient open-source vision language model that combines powerful image understanding capabilities with a remarkably small footprint.

Vision
image-to-image
falREVIEW REQUIRED

PuLID

fal-ai/pulid

Tuning-free ID customization.

editingcustomizationpersonalization
audio-to-text
nvidiaREVIEW REQUIRED

Nemotron 3 Nano Omni

nvidia/nemotron-3-nano-omni/audio

Audio reasoning variant of NVIDIA's Nemotron 3 Nano Omni. 30B A3B hybrid Transformer-Mamba MoE - accepts audio plus a prompt and returns text.

nemotronnvidiaaudio-to-textaudio-understanding
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit Plus Lora

fal-ai/qwen-image-edit-plus-lora

LoRA endpoint for the Qwen Image Edit Plus model.

image-to-imageimage-editing
image-to-video
mirage-apiREVIEW REQUIRED

Mirage Avatar X

mirage-api/avatar-x/reference-to-video

The Avatar X API offers access to Mirage's most advanced generation model yet, delivering industry-leading identity preservation and expressivity in AI video

avatarlipsynctalking-head