EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 35 · 28 per page
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Subject

fal-ai/flux-subject

Super fast endpoint for the FLUX.1 [schnell] model with subject input capabilities, enabling rapid and high-quality image generation for personalization, specific styles, brand identities, and product-specific outputs.

personalizationcustomization
video-to-video
falREVIEW REQUIRED

Wan 2.1 VACE Long Reframe

fal-ai/wan-vace-apps/long-reframe

Reframe entire videos scene-by-scene using Wan VACE 2.1

video-to-video
falREVIEW REQUIRED

ID-V2V Relight

fal-ai/id-v2v/relight

Change a video’s lighting using a relit reference frame while preserving the scene, subjects, and original performance. ID-V2V Relight propagates the new illumination across the video.

relightingeditingcinematic
video-to-audio
falREVIEW REQUIRED

Sam Audio

fal-ai/sam-audio/visual-separate

Audio separation with SAM Audio. Isolate any sound using natural language—professional-grade audio editing made simple for creators, researchers, and accessibility applications.

video-to-audiosam-audio
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit Plus Lora Gallery

fal-ai/qwen-image-edit-plus-lora-gallery/remove-element

Remove unwanted elements (objects, people, text) while maintaining image consistency

stylizedtransform
image-to-image
falREVIEW REQUIRED

Latent Consistency Models (v1.5/XL)

fal-ai/fast-lcm-diffusion/image-to-image

Run SDXL at the speed of light

lcmdiffusionturboreal-time
image-to-3d
falREVIEW REQUIRED

Trellis 2

fal-ai/trellis-2/retexture

Generate 3D models from your images using Trellis 2. A native 3D generative model enabling versatile and high-quality 3D asset creation.

image-to-3D
image-to-image
falREVIEW REQUIRED

SDXL ControlNet Union

fal-ai/sdxl-controlnet-union/image-to-image

An efficent SDXL multi-controlnet image-to-image model.

diffusioncontrolnetcomposition
image-to-image
falREVIEW REQUIRED

Style Transfer

fal-ai/image-apps-v2/style-transfer

Apply artistic styles like impressionism, cubism, or surrealism to your images.

style-transfer
image-to-3d
falREVIEW REQUIRED

ReconViaGen 0.5

fal-ai/reconviagen-0.5

Generate 3D models from one or more images using ReconViaGen 0.5

multi-view3d-reconstruction
image-to-video
nvidiaREVIEW REQUIRED

Cosmos 3 Super Image to Video

nvidia/cosmos-3-super/image-to-video

Cosmos3 is a collection of Omnimodal world models capable of generating dynamic, high-quality video, image, audio, and action commands from combinations of text, image, video, and action trajectory inputs.

stylizedtransformlipsync
image-to-video
falREVIEW REQUIRED

Kandinsky5 Pro

fal-ai/kandinsky5-pro/image-to-video

Kandinsky 5.0 Pro is a diffusion model for fast, high-quality image-to-video generation.

image-to-3d
falREVIEW REQUIRED

Hunyuan3D

fal-ai/hunyuan3d/v2/mini/turbo

Generate 3D models from your images using Hunyuan 3D. A native 3D generative model enabling versatile and high-quality 3D asset creation.

stylized
text-to-image
falREVIEW REQUIRED

OmniGen v1

fal-ai/omnigen-v1

OmniGen is a unified image generation model that can generate a wide range of images from multi-modal prompts. It can be used for various tasks such as Image Editing, Personalized Image Generation, Virtual Try-On, Multi Person Generation and more!

multimodaleditingtry-on
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev]

fal-ai/flux-1/dev/image-to-image

FLUX.1 [dev] is a 12 billion parameter flow transformer that generates high-quality images from text. It is suitable for personal and commercial use.

image-to-image
falREVIEW REQUIRED

PhotoMaker

fal-ai/photomaker

Customizing Realistic Human Photos via Stacked ID Embedding

editingcustomizationrealismpersonalization
image-to-image
falREVIEW REQUIRED

Image Editing Face Enhancement

fal-ai/image-editing/face-enhancement

Enhance facial features with professional retouching while maintaining a natural, realistic look

stylizedtransform
training
falREVIEW REQUIRED

Phota Create Profile

fal-ai/phota/create-profile

Generate profiles using 30-50 images of a subject with Phota.

stylizedtransformtypographyphota
image-to-image
falREVIEW REQUIRED

Image Preprocessors

fal-ai/image-preprocessors/zoe

ZoeDepth preprocessor.

depthpreprocessutilitycontrolnet
video-to-audio
mirelo-aiREVIEW REQUIRED

Mirelo SFX

mirelo-ai/sfx-v1/video-to-audio

Generate synced sounds for any video, and return the new sound track (like MMAudio)

sfx
image-to-video
falREVIEW REQUIRED

AI Avatar Multi

fal-ai/ai-avatar/multi

MultiTalk model generates a multi-person conversation video from an image and audio files. Creates a realistic scene where multiple people speak in sequence.

stylizedtransform
video-to-video
falREVIEW REQUIRED

Wan VACE 14B

fal-ai/wan-vace-14b/outpainting

VACE is a video generation model that uses a source image, mask, and video to create prompted videos with controllable sources.

image-to-videovideo-to-videotext-to-video
video-to-video
AlibabaREVIEW REQUIRED

Wan v2.6 Reference to Video

wan/v2.6/reference-to-video

Wan 2.6 reference-to-video model.

reference-to-video
video-to-video
falREVIEW REQUIRED

Lightx

fal-ai/lightx/relight

Use tlightx capabilities to relight and recamera your videos.

video-to-video
text-to-image
AlibabaREVIEW REQUIRED

Wan v2.2 A14B Text-to-Image A14B with LoRAs

fal-ai/wan/v2.2-a14b/text-to-image/lora

Wan 2.2's 14B model with LoRA support generates high-fidelity images with enhanced prompt alignment, style adaptability.

image-to-video
falREVIEW REQUIRED

Vidu Image to Video

fal-ai/vidu/q1/image-to-video

Vidu Q1 Image to Video generates high-quality 1080p videos with exceptional visual quality and motion diversity from a single image

stylizedtransform
image-to-image
falREVIEW REQUIRED

Chrono Edit

fal-ai/chrono-edit

NVIDIA's Logically Consistent and Physics-Aware Image Editing Model

image-editing
image-to-image
falREVIEW REQUIRED

Florence-2 Large

fal-ai/florence-2-large/referring-expression-segmentation

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

multimodalvisionsegmentation