EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 30 · 28 per page
text-to-3d
falREVIEW REQUIRED

Hunyuan 3d

fal-ai/hunyuan-3d/v3.1/rapid/text-to-3d

Create detailed, fully-textured 3D models with text

3d
3d-to-3d
falREVIEW REQUIRED

Hunyuan 3D Part Splitter

fal-ai/hunyuan-3d/v3.1/part

Split 3D models into parts with Hunyuan 3D

3dhunyuanmesh
image-to-image
falREVIEW REQUIRED

try-on

fal-ai/cat-vton

Image based high quality Virtual Try-On

try-onfashionclothing
image-to-image
falREVIEW REQUIRED

Hidream O1 Image

fal-ai/hidream-o1-image/dev/edit

Unified image generation with HiDream-O1-Image. Create, edit, and personalize high-resolution images up to 2K—single native model handles text-to-image, editing, and custom subjects without external components.

3d-to-3d
falREVIEW REQUIRED

Hunyuan 3D Smart Topology

fal-ai/hunyuan-3d/v3.1/smart-topology

Optimize 3D mesh topology with Hunyuan 3D Smart Topology.

3dhunyuantopology
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Krea [dev] with LoRAs

fal-ai/flux-krea-lora

Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

lorapersonalization
video-to-video
soniloREVIEW REQUIRED

V1.1 Video to Video Sound Effects

sonilo/v1.1/video-to-video-sound-effects

Adds synchronized, royalty-free, commercial-use-safe sound effects to a video. Returns the finished video with the generated audio mixed in.

sfxaudioeffects
image-to-video
falREVIEW REQUIRED

LongCat Video Distilled

fal-ai/longcat-video/distilled/image-to-video/480p

Generate long videos from images using LongCat Video Distilled

text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 4B Base

fal-ai/flux-2/klein/4b/base

Text-to-image generation with FLUX.2 [klein] 4B Base from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.

image-to-image
Black Forest LabsREVIEW REQUIRED

Flux Vision Upscaler

fal-ai/flux-vision-upscaler

Flux Vision Upscaler for magnify/upscaling images with high fidelity and creativity.

video-to-video
falREVIEW REQUIRED

Wan VACE Video Edit

fal-ai/wan-vace-apps/video-edit

Edit videos using plain language and Wan VACE

video-editwan-vace
image-to-image
falREVIEW REQUIRED

Image Editing Retouch

fal-ai/image-editing/retouch

Retouch photos of faces. Remove blemishes and improve the skin.

text-to-3d
MeshyREVIEW REQUIRED

V7 Text to 3D

meshy/v7/text-to-3d

Turns text into a fully textured, PBR-ready 3D mesh with complete geometry, in game-ready Smart Topology at a target polygon count

stylizedtransform
text-to-video
PixVerseREVIEW REQUIRED

PixVerse V5.5 Text To Video

fal-ai/pixverse/v5.5/text-to-video

Generate high quality video clips from text and image prompts using PixVerse v5.5

text-to-video
image-to-video
Black Forest LabsREVIEW REQUIRED

Flux 3 First Last Frame to Video Draft

blackforestlabs/flux-3/first-last-frame-to-video/draft

FLUX.3 is Black Forest Labs' frontier audio/video model. Generate fast, low-cost draft previews between a start and an end frame, with a reusable draft cache for full-quality enhancement.

stylizedtransformlipsync
text-to-speech
AlibabaREVIEW REQUIRED

Qwen 3 TTS - Text to Speech [0.6B]

fal-ai/qwen-3-tts/text-to-speech/0.6b

Bring speech to your texts using Qwen3-TTS Custom-Voice model with pre-trained voices or use your custom voice with Qwen3-TTS Clone Voice model

text-to-speech
text-to-audio
falREVIEW REQUIRED

Kokoro TTS (Spanish)

fal-ai/kokoro/spanish

A natural-sounding Spanish text-to-speech model optimized for Latin American and European Spanish.

speech
text-to-image
falREVIEW REQUIRED

Realistic Vision

fal-ai/realistic-vision

Generate realistic images.

realismdiffusion
audio-to-video
falREVIEW REQUIRED

LTX-2.3 22B

fal-ai/ltx-2.3-22b/audio-to-video

Generate video with audio from audio, text and images using LTX-2

video-to-video
falREVIEW REQUIRED

LTX Video 2.3 Pro

fal-ai/ltx-2.3/retake-video

LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.

stylizedtransformlipsync
image-to-image
falREVIEW REQUIRED

Firered Image Edit

fal-ai/firered-image-edit

FireRed Image Edit is FireRed's state of the art open source editing model, re-trained from Qwen Image Edit 2509.

image-editingfirered
text-to-image
falREVIEW REQUIRED

Stable Diffusion 3.5 Medium

fal-ai/stable-diffusion-v35-medium

Stable Diffusion 3.5 Medium is a Multimodal Diffusion Transformer (MMDiT) text-to-image model that features improved performance in image quality, typography, complex prompt understanding, and resource-efficiency.

diffusiontypographystyle
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Krea [dev]

fal-ai/flux/krea/image-to-image

FLUX.1 Krea [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.

image-to-image
briaREVIEW REQUIRED

Fibo Edit [Restore]

bria/fibo-edit/restore

Photo restoration model that automatically denoises, deblurs, and enhances old or damaged photos - removes imperfections while preserving original character.

image-restorationfibo-editbriajson
image-to-image
briaREVIEW REQUIRED

Bria Product Holding (FIBO-Edit-1.5)

bria/fibo-edit-1.5/product-holding

Bria Product Holding edits a person photo to show the subject holding or carrying a product, using one to three product reference images and optional text instructions. Built on FIBO-Edit-1.5, it preserves the source aspect ratio by default and supports a selectable output aspect ratio.

product photographyproduct placemente-commercemarketing visuals
image-to-image
falREVIEW REQUIRED

Post Processing

fal-ai/post-processing

Post Processing is an endpoint that can enhance images using a variety of techniques including grain, blur, sharpen, and more.

stylizedutility
video-to-video
briaREVIEW REQUIRED

Video

bria/video/erase/prompt

Erase unwanted objects, people, or elements from video with a text prompt. High-fidelity output with strong temporal consistency, trained on licensed data for safe commercial use.

briavideoerase