EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 36 · 28 per page
video-to-video
AlibabaREVIEW REQUIRED

V2.6

wan/v2.6/reference-to-video/flash

Wan 2.6 reference-to-video flash model.

reference-to-video
speech-to-text
falREVIEW REQUIRED

Speech-to-Text

fal-ai/speech-to-text/turbo

Leverage the rapid processing capabilities of AI models to enable accurate and efficient real-time speech-to-text transcription.

image-to-image
hitem3dREVIEW REQUIRED

Hi3D Image to Relief

hitem3d/hi3d/image-to-relief

Generate a 3D relief depth map with Hi3D from a single image.

depthhi3drelief
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V5.6 Image To Video

fal-ai/pixverse/v5.6/image-to-video

Use the latest pixverse v5.6 model to turn your texts and images into amazing videos.

image-to-video
image-to-image
falREVIEW REQUIRED

Joyai Image Edit

fal-ai/joyai-image-edit

All-in-one image AI with JoyAI-Image. Understand, create, and edit images through natural language—the model's deep visual understanding powers more accurate generation and precise editing in a unified system.

image-to-imageimage-editing
video-to-video
briaREVIEW REQUIRED

Video

bria/video/background-removal/green-screen-despill

Remove background from videos filmed using chromakey, with automatic green spill suppression for clean, professional edges.

image-to-image
falREVIEW REQUIRED

Z Image Turbo Inpaint Lora

fal-ai/z-image/turbo/inpaint/lora

Generate images from text, an image, a mask and custom LoRA using Z-Image Turbo, Tongyi-MAI's super-fast 6B model.

inpainting
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] Control LoRA Canny

fal-ai/flux-control-lora-canny/image-to-image

FLUX Control LoRA Canny is a high-performance endpoint that uses a control image using a Canny edge map to transfer structure to the generated image and another initial image to guide color.

lorastyle transfer
text-to-video
falREVIEW REQUIRED

LTX-Video 13B 0.9.8 Distilled

fal-ai/ltxv-13b-098-distilled

Generate long videos from prompts using LTX Video-0.9.8 13B Distilled and custom LoRA

videoltx-videotext-to-video
image-to-image
falREVIEW REQUIRED

Stable Diffusion XL

fal-ai/fast-sdxl/inpainting

Run SDXL at the speed of light

diffusionhigh-resloraip-adapter
image-to-image
briaREVIEW REQUIRED

Fibo Edit [Erase by Text]

bria/fibo-edit/erase_by_text

Remove unwanted objects from images with a text prompt - fast, precise editing that seamlessly blends results. Built for production scale and trained on licensed data for safe commercial use.

briafibo-editprompt-eraser
image-to-video
falREVIEW REQUIRED

Vidu Start-End to Video

fal-ai/vidu/start-end-to-video

Vidu Start-End to Video generates smooth transition videos between specified start and end images.

motiontransition
image-to-image
falREVIEW REQUIRED

Z Image Turbo Controlnet Lora

fal-ai/z-image/turbo/controlnet/lora

Generate images from text and edge, depth or pose images using custom LoRA and Z-Image Turbo, Tongyi-MAI's super-fast 6B model.

turboz-imagefastlora
text-to-audio
falREVIEW REQUIRED

Stable Audio 3

fal-ai/stable-audio-3/small/music/base/text-to-audio

Stable Audio 3 Small Music Base is the foundational 459 million parameter checkpoint generating full music compositions up to 2 minutes from text prompts, intended as the unmodified base for fine-tuning.

musicon-devicelightweight
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 SRPO [dev]

fal-ai/flux/srpo

FLUX.1 SRPO [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.

training
MiniMaxREVIEW REQUIRED

MiniMax H3 Reference to Video LoRA Trainer

minimax/h3/ref2va/trainer

Train a MiniMax H3 LoRA with reference conditioning, so different modalities animate into video with audio; captions optional.

image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 SRPO [dev]

fal-ai/flux/srpo/image-to-image

FLUX.1 SRPO [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.

text-to-image
falREVIEW REQUIRED

Lumina Image 2

fal-ai/lumina-image/v2

Lumina-Image-2.0 is a 2 billion parameter flow-based diffusion transforer which features improved performance in image quality, typography, complex prompt understanding, and resource-efficiency.

diffusiontypographystyle
audio-to-audio
falREVIEW REQUIRED

Workflow Utilities Audio Compressor

fal-ai/workflow-utilities/audio-compressor

FFMPEG Utility for Audio Compression

image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit 2509 Lora Gallery

fal-ai/qwen-image-edit-2509-lora-gallery/multiple-angles

Precise camera position and angle control (rotation, zoom, vertical movement)

stylizedtransform
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V4 Image To Video

fal-ai/pixverse/v4/image-to-video

Generate high quality video clips from text and image prompts using PixVerse v4

image-to-video
falREVIEW REQUIRED

Wan-2.1 Pro Image-to-Video

fal-ai/wan-pro/image-to-video

Wan-2.1 Pro is a premium image-to-video model that generates high-quality 1080p videos at 30fps with up to 6 seconds duration, delivering exceptional visual quality and motion diversity from images

image to videomotion
image-to-video
falREVIEW REQUIRED

Vidu Start End to Video

fal-ai/vidu/q1/start-end-to-video

Vidu Q1 Start-End to Video generates smooth transition 1080p videos between specified start and end images.

stylizedtransform
text-to-image
falREVIEW REQUIRED

Bria Text-to-Image Base

fal-ai/bria/text-to-image/base

Bria's Text-to-Image model, trained exclusively on licensed data for safe and risk-free commercial use. Available also as source code and weights. For access to weights: https://bria.ai/contact-us

image generation
video-to-video
falREVIEW REQUIRED

Bernini-R Reference Edit Video

fal-ai/bernini-r/reference-edit-video

Edit a video guided by reference images with Bernini-R, bringing an object, material, background, style, or weather from a reference image into your video.

editreferencetransform
text-to-speech
MiniMaxREVIEW REQUIRED

Minimax

fal-ai/minimax/preview/speech-2.5-turbo

Generate fast speech from text prompts and different voices using the MiniMax Speech-02 Turbo model, which leverages advanced AI techniques to create high-quality text-to-speech.

text-to-speech
text-to-image
falREVIEW REQUIRED

CogView

fal-ai/cogview4

Generate high quality images from text prompts using CogView4. Longer text prompts will result in better quality images.

stylized
image-to-video
falREVIEW REQUIRED

Flashhead

fal-ai/flashhead

SoulX-FlashHead is a unified 1.3B-parameter framework designed for high-fidelity, infinite-length, and real-time streaming portrait video generation.

portraitvideostreamingreal-time