EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 37 · 28 per page
text-to-image
falREVIEW REQUIRED

SDXL ControlNet Union

fal-ai/sdxl-controlnet-union

An efficent SDXL multi-controlnet text-to-image model.

diffusioncontrolnetcomposition
image-to-video
falREVIEW REQUIRED

Framepack

fal-ai/framepack

Framepack is an efficient Image-to-video model that autoregressively generates videos.

image to videomotion
image-to-image
falREVIEW REQUIRED

Product Holding

fal-ai/image-apps-v2/product-holding

Place products naturally in a person’s hands for realistic marketing visuals.

productmarketing
text-to-video
falREVIEW REQUIRED

Kandinsky5 Pro

fal-ai/kandinsky5-pro/text-to-video

Kandinsky 5.0 Pro is a diffusion model for fast, high-quality text-to-video generation.

video-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 Action

fal-ai/flux-3-action/so101

FLUX 3 Action turns what the robot sees into what it does next. Give it the scene camera image, the wrist camera image, the current SO-101 joint state and a plain-language instruction.

roboticarm
training
Black Forest LabsREVIEW REQUIRED

Flux Kontext Trainer

fal-ai/flux-kontext-trainer

LoRA trainer for FLUX.1 Kontext [dev]

text-to-image
falREVIEW REQUIRED

Sana v1.5 1.6B

fal-ai/sana/v1.5/1.6b

Sana v1.5 1.6B is a lightweight text-to-image model that delivers 4K image generation with impressive efficiency.

text to image4klightweight
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Layered

fal-ai/qwen-image-layered/lora

Qwen-Image-Layered is a model capable of decomposing an image into multiple RGBA layers. Use loras to get your custom outputs.

qwenlora
video-to-video
topazREVIEW REQUIRED

Topaz Denoise Video

topaz/denoise/video

Professional video denoising powered by Topaz Labs. Nyx models remove noise at source resolution, with Nyx Fast as a lighter, cheaper pass. Best for low-light and high-ISO footage.

denoisevideo
image-to-image
briaREVIEW REQUIRED

Fibo Edit [Colorize]

bria/fibo-edit/colorize

Image colorization and color-grading model. Bring color to black-and-white photos or apply curated color treatments using simple style-based commands.

briafibo-editcolor
text-to-audio
falREVIEW REQUIRED

Zonos-Audio-Clone

fal-ai/zonos

Clone voice of any person and speak anything in their voice using zonos' voice cloning.

voice cloning
audio-to-audio
AlibabaREVIEW REQUIRED

Qwen 3 TTS - Clone Voice [0.6B]

fal-ai/qwen-3-tts/clone-voice/0.6b

Clone your voices using Qwen3-TTS Clone-Voice model with zero shot cloning capabilities and use it on text-to-speech models to create speeches of yours!

clone-voicevoice-clone
text-to-image
falREVIEW REQUIRED

DeepSeek Janus-Pro

fal-ai/janus

DeepSeek Janus-Pro is a novel text-to-image model that unifies multimodal understanding and generation through an autoregressive framework

stylized
text-to-image
falREVIEW REQUIRED

Fooocus

fal-ai/fooocus

Default parameters with automated optimizations and quality improvements.

stylized
image-to-image
falREVIEW REQUIRED

Bagel

fal-ai/bagel/edit

Bagel is a 7B parameter multimodal model from Bytedance-Seed that can generate both images and text.

image-to-imageimage-editing
image-to-image
falREVIEW REQUIRED

DreamOmni2

fal-ai/dreamomni2/edit

DreamOmni2 is a unified multimodal model for text and image guided image editing.

video-to-video
briaREVIEW REQUIRED

Bria Video Eraser

bria/bria_video_eraser/erase/prompt

A high-fidelity capability for erasing unwanted objects, people, or visual elements from videos while maintaining aesthetic quality and temporal consistency

briaerase
image-to-video
falREVIEW REQUIRED

High Quality Stable Video Diffusion

fal-ai/stable-video

Generate short video clips from your images using SVD v1.1

image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [schnell] Redux

fal-ai/flux/schnell/redux

FLUX.1 [schnell] Redux is a high-performance endpoint for the FLUX.1 [schnell] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.

style transfer
vision
falREVIEW REQUIRED

Marlin

fal-ai/marlin

Marlin is a 2B video VLM tuned for the two questions developers actually want to ask of their videos: what is happening, and when?

utilityediting
text-to-audio
falREVIEW REQUIRED

Stable Audio 3 Small SFX Base Text to Audio

fal-ai/stable-audio-3/small/sfx/base/text-to-audio

Stable Audio 3 Small SFX Base is the foundational 459 million parameter checkpoint generating sound effects from text prompts, intended as the unmodified base for fine-tuning.

sfxsound-effectson-device
text-to-audio
falREVIEW REQUIRED

Kokoro TTS (Italian)

fal-ai/kokoro/italian

A high-quality Italian text-to-speech model delivering smooth and expressive speech synthesis.

speech
image-to-video
falREVIEW REQUIRED

Vidu

fal-ai/vidu/q2/image-to-video/pro

Use the latest Vidu Q2 models which much more better quality and control on your videos.

image-to-video
video-to-video
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/extend-video

Extend high-quality video with audio from input video using LTX-2.3

extendlonger
video-to-video
falREVIEW REQUIRED

LTX Video-0.9.7 13B Distilled

fal-ai/ltx-video-13b-distilled/multiconditioning

Generate videos from prompts, images, and videos using LTX Video-0.9.7 13B Distilled and custom LoRA

videoltx-videovideo-to-videomulticondition-to-video
audio-to-audio
falREVIEW REQUIRED

ACE Step Audio Outpaint

fal-ai/ace-step/audio-outpaint

Extend the beginning or end of provided audio with lyrics and/or style using ACE-Step

audio-to-audioaudio-outpaintaudio-extend
image-to-image
Black Forest LabsREVIEW REQUIRED

Juggernaut Flux Pro

rundiffusion-fal/juggernaut-flux/pro/image-to-image

Juggernaut Pro Flux by RunDiffusion is the flagship Juggernaut model rivaling some of the most advanced image models available, often surpassing them in realism. It combines Juggernaut Base with RunDiffusion Photo and features enhancements like reduced background blurriness.

image generation
video-to-video
PixVerseREVIEW REQUIRED

PixVerse Extend

fal-ai/pixverse/extend

PixVerse Extend model is a video extending tool for your videos using with high-quality video extending techniques

utilityediting