EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 31 · 28 per page
text-to-image
falREVIEW REQUIRED

Fooocus Inpainting

fal-ai/fooocus/inpaint

Default parameters with automated optimizations and quality improvements.

stylizedediting
text-to-image
falREVIEW REQUIRED

Illusion Diffusion

fal-ai/illusion-diffusion

Create illusions conditioned on image.

compositionstylized
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V5 Effects

fal-ai/pixverse/v5/effects

Generate high quality video clips with different effects using PixVerse v5

image-to-video
video-to-video
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/outpaint

Outpaint high-quality video using LTX-2.3

outpaintoutpainting
text-to-video
MiniMaxREVIEW REQUIRED

MiniMax H3 Text to Video LoRA

minimax/h3/text-to-video/lora

Generate video with synchronized audio from a text prompt using MiniMax H3; load a trained LoRA at adjustable strength to lock in style, character, or motion.

utilityediting
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] Depth with LoRAs

fal-ai/flux-lora-depth

Generate high-quality images from depth maps using Flux.1 [dev] depth estimation model. The model produces accurate depth representations for scene understanding and 3D visualization.

depthlorautilitycomposition
video-to-video
falREVIEW REQUIRED

Wan 2.2 VACE Fun A14B

fal-ai/wan-22-vace-fun-a14b/depth

VACE Fun for Wan 2.2 A14B from Alibaba-PAI

video-to-video
KlingREVIEW REQUIRED

Kling O1 Reference Video to Video [Standard]

fal-ai/kling-video/o1/standard/video-to-video/reference

Kling O1 Omni generates new shots guided by an input reference video, preserving cinematic language such as motion, and camera style to produce seamless scene continuity.

image-to-image
IdeogramREVIEW REQUIRED

Ideogram Upscale

fal-ai/ideogram/upscale

Ideogram Upscale enhances the resolution of the reference image by up to 2X and might enhance the reference image too. Optionally refine outputs with a prompt for guided improvements.

upscalinghigh-res
image-to-3d
falREVIEW REQUIRED

Hunyuan World

fal-ai/hunyuan_world/image-to-world

Hunyuan World 1.0 turns a single image into a panorama or a 3D world. It creates realistic scenes from the image, allowing you to explore and view it from different angles.

image-to-image
falREVIEW REQUIRED

Moondream3 Preview [Segment]

fal-ai/moondream3-preview/segment

Moondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.

masksegmentation
text-to-image
falREVIEW REQUIRED

Hidream O1 Image

fal-ai/hidream-o1-image/dev

Unified image generation with HiDream-O1-Image. Create, edit, and personalize high-resolution images up to 2K—single native model handles text-to-image, editing, and custom subjects without external components.

llm
falREVIEW REQUIRED

Video Prompt Generator

fal-ai/video-prompt-generator

Generate video prompts using a variety of techniques including camera direction, style, pacing, special effects and more.

motiontransformationchatclaude
vision
falREVIEW REQUIRED

Florence 2 Large Caption

fal-ai/florence-2-large/caption

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

captioningmultimodalvision
text-to-image
RecraftREVIEW REQUIRED

Recraft V4 Styles Pro Text to Vector

recraft/v4/style/pro/text-to-vector

Generates vector images that hold a consistent style, from either a saved style ID or reference images attached directly.

stylizedtransformediting
text-to-video
MiniMaxREVIEW REQUIRED

H3 Max VHS

minimax/h3-max/styles/vhs

Generates 768p VHS-style video with audio from text prompts or an optional first-frame image. Supports 5–15 second clips and adjustable tape damage, from subtle analog noise to strong tracking distortion.

vhsanalogretrotape-damage
video-to-video
briaREVIEW REQUIRED

Bria Video Eraser Erase Mask

bria/bria_video_eraser/erase/mask

A high-fidelity capability for erasing unwanted objects, people, or visual elements from videos while maintaining aesthetic quality and temporal consistency.

briaerase
image-to-image
briaREVIEW REQUIRED

Genfill

bria/genfill/v2

The GenFill Route enables the generation of objects by prompt in a specific region of an image. You can define the area for object generation by using a mask that outlines the region where the object will be created. Our model is optimized to work seamlessly with blob-shaped masks.

training
Black Forest LabsREVIEW REQUIRED

FLUX 2 [klein] 9b Base Trainer

fal-ai/flux-2-klein-9b-base-trainer

Fine-tune FLUX.2 [klein] 9B from Black Forest Labs with custom datasets. Create specialized LoRA adaptations for specific editing tasks.

image-to-image
falREVIEW REQUIRED

Vidu

fal-ai/vidu/q2/reference-to-image

Vidu Reference-to-Image creates images by using a reference images and combining them with a prompt.

images-to-imagreference-to-image
image-to-image
falREVIEW REQUIRED

Stable Diffusion XL Lightning

fal-ai/fast-lightning-sdxl/image-to-image

Run SDXL at the speed of light

diffusionlightningediting
text-to-video
MiniMaxREVIEW REQUIRED

H3 Max 16-bit Pixel

minimax/h3-max/styles/16bit-pixel

Generates 768p video with audio in a 16-bit pixel-art style from text prompts or an optional first-frame image. Supports durations of 5–15 seconds.

pixel-art16-bitanimationstylized
video-to-video
falREVIEW REQUIRED

Wan VACE 14B

fal-ai/wan-vace-14b/pose

VACE is a video generation model that uses a source image, mask, and video to create prompted videos with controllable sources.

image-to-videovideo-to-videotext-to-video
text-to-image
RecraftREVIEW REQUIRED

Recraft V4.1 Utility Text to Image

fal-ai/recraft/v4.1/utility/pro/text-to-image

Recraft V4.1 Utility Pro pairs the high-resolution output of V4.1 Pro with a faster, cost-efficient runtime. Designed for studios shipping large-format work at scale, it makes premium-quality raster generation viable across full creative pipelines.

stylizedtransformtypography
image-to-image
IdeogramREVIEW REQUIRED

Ideogram V4.0q Tiling

ideogram/v4/tiling

Ideogram V4.0q Tiling generates seamless, edge-matching textures and patterns that repeat infinitely in any direction, ideal for backgrounds, surfaces, and wallpapers.

stylizedtransformrealism