EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 21 · 28 per page
image-to-image
falREVIEW REQUIRED

Workflow Utilities Extract Nth Frame

fal-ai/workflow-utilities/extract-nth-frame

FFMPEG Untility for Extracting nth Frame

text-to-image
falREVIEW REQUIRED

Z Image Base

fal-ai/z-image/base

Z-Image is the foundation model of the Z- Image family, engineered for good quality, robust generative diversity, broad stylistic coverage, and precise prompt adherence.

z-imagebase
video-to-video
falREVIEW REQUIRED

Sam 3 1

fal-ai/sam-3-1/video-rle

SAM 3.1 builds comes with Object Multiplex, a shared-memory approach for joint multi-object tracking that delivers faster speeds with larger number of objects tracked.

segmentationmaskreal-time
text-to-image
Luma AIREVIEW REQUIRED

Luma Uni-1 Text to Image

luma/agent/uni-1/v1/text-to-image

Luma Uni-1 turns a text prompt into a single high-fidelity image, with control over aspect ratio and visual style, plus optional web-sourced and reference-image guidance for sharper grounding.

realismtypographystylized
video-to-video
KlingREVIEW REQUIRED

Kling O1 Reference Video to Video [Pro]

fal-ai/kling-video/o1/video-to-video/reference

Kling O1 Omni generates new shots guided by an input reference video, preserving cinematic language such as motion, and camera style to produce seamless scene continuity.

image-to-image
topazREVIEW REQUIRED

Topaz Restore Image

topaz/restore/image

Professional image restoration powered by Topaz Labs. Recover 3 generatively rebuilds natural detail; Dust-Scratch V2 cleans film dust and scratches. Best for old, damaged or degraded photos.

restoreimage
image-to-text
nvidiaREVIEW REQUIRED

Nemotron 3 Nano Omni

nvidia/nemotron-3-nano-omni/vision

Vision reasoning variant of NVIDIA's Nemotron 3 Nano Omni. 30B A3B hybrid Transformer-Mamba MoE - accepts an image plus a prompt and returns text.

nemotronnvidiaimage-to-textvision-language
video-to-audio
mirelo-aiREVIEW REQUIRED

Mirelo SFX V1.5

mirelo-ai/sfx-v1.5/video-to-audio

Generate synced sounds for any video, and return the new sound track (like MMAudio)

video-to-audiosfx
image-to-image
GoogleREVIEW REQUIRED

Google Virtual Try On

google/virtual-try-on

Generate realistic virtual try-on images from a person image and a clothing product image.

virtualtryclothes
image-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 Image To Video Draft

blackforestlabs/flux-3/image-to-video/draft

FLUX.3 is Black Forest Labs' frontier audio/video model. Generate fast, low-cost draft previews that animate a still image, with a reusable draft cache for full-quality enhancement.

stylizedtransformlipsync
image-to-3d
hitem3dREVIEW REQUIRED

Hi3D Multiview to 3D

hitem3d/hi3d/v3.0/multi-view-to-3d

Generate 3D models from multiple view images using Hi3D V3.0.

image-to-3dmultiview-to-3d3d
image-to-image
Black Forest LabsREVIEW REQUIRED

Flux Kontext Lora

fal-ai/flux-kontext-lora/inpaint

Fast inpainting endpoint for the FLUX.1 Kontext [dev] model with LoRA support, enabling rapid and high-quality image inpainting with reference images, while using pre-trained LoRA adaptations for specific styles, brand identities, and product-specific outputs.

image-editingimage-inpaintingimage-to-image
image-to-image
falREVIEW REQUIRED

Marigold Depth Estimation

fal-ai/imageutils/marigold-depth

Create depth maps using Marigold depth estimation.

depthutility
text-to-image
falREVIEW REQUIRED

Stable Diffusion 3.5 Large

fal-ai/stable-diffusion-v35-large

Stable Diffusion 3.5 Large is a Multimodal Diffusion Transformer (MMDiT) text-to-image model that features improved performance in image quality, typography, complex prompt understanding, and resource-efficiency.

diffusiontypographystyle
image-to-3d
MeshyREVIEW REQUIRED

V7 Multi Image to 3D

meshy/v7/multi-image-to-3d

econstructs a high-fidelity textured 3D model from multiple angle views of one object, with game-ready topology and polygon control

stylizedtransform
text-to-audio
mirelo-aiREVIEW REQUIRED

Mirelo SFX1.6

mirelo-ai/sfx1.6/text-to-audio

Generate ambient sounds for any text prompt. Now you can turn any SFX into a natural loop for ambient soundscapes.

text-to-audiosfx
text-to-image
Luma AIREVIEW REQUIRED

Luma Uni-1 Text to Image Max

luma/agent/uni-1/v1/max

Luma Uni-1 Max generates a single image at the model's highest fidelity, delivering richer detail and stronger prompt adherence than the base tier for hero-quality stills.

realismtypographystylized
video-to-video
pixelcutREVIEW REQUIRED

Pixelcut Video Background Removal

pixelcut/video-background-removal

Pixelcut's Video Background Remover is an AI segmentation model that erases backgrounds frame by frame, with seamless temporal consistency.

transformutilityrembg
audio-to-video
falREVIEW REQUIRED

Flashtalk

fal-ai/flashtalk

Audio-driven talking avatar generation powered by the SoulX-FlashTalk 14B model.

avatartalking-headaudio-drivenlip-sync
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev]

fal-ai/flux-1/dev

FLUX.1 [dev] is a 12 billion parameter flow transformer that generates high-quality images from text. It is suitable for personal and commercial use.

image-to-3d
tripo3dREVIEW REQUIRED

Tripo3D

tripo3d/tripo/v2.5/multiview-to-3d

State of the art Multiview to 3D Object generation. Generate 3D models from multiple images!

stylizedmultiview
video-to-video
briaREVIEW REQUIRED

Video

bria/video/erase/mask

High-fidelity mask-based video object removal with strong temporal consistency. Erase unwanted objects, people, or elements while preserving aesthetic quality. Trained on licensed data for risk-free commercial use.

briavideoerase
video-to-video
GoogleREVIEW REQUIRED

Veo 3.1

fal-ai/veo3.1/extend-video

Extend Veo-Created Videos up to 30 seconds

extend-video
video-to-video
xAIREVIEW REQUIRED

Grok Imagine Extend Video

xai/grok-imagine-video/extend-video

Extend videos with xAI's Grok Imagine video model

video-editv2vgrokxai
text-to-image
falREVIEW REQUIRED

Stable Diffusion with LoRAs

fal-ai/lora

Run Any Stable Diffusion model with customizable LoRA weights.

diffusionloracustomization
video-to-video
mirelo-aiREVIEW REQUIRED

Mirelo SFX V1.5

mirelo-ai/sfx-v1.5/video-to-video

Generate synced sounds for any video, and return it with its new sound track (like MMAudio)

video-to-videosfx