EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 38 · 28 per page
text-to-image
AlibabaREVIEW REQUIRED

Wan v2.2 A14B Text-to-Image A14B with LoRAs

fal-ai/wan/v2.2-a14b/text-to-image/lora

Wan 2.2's 14B model with LoRA support generates high-fidelity images with enhanced prompt alignment, style adaptability.

text-to-video
falREVIEW REQUIRED

Infinitalk

fal-ai/infinitalk/single-text

Infinitalk model generates a talking avatar video from a text and audio file. The avatar lip-syncs to the provided audio with natural facial expressions.

text-to-image
falREVIEW REQUIRED

Bria Text-to-Image Base

fal-ai/bria/text-to-image/base

Bria's Text-to-Image model, trained exclusively on licensed data for safe and risk-free commercial use. Available also as source code and weights. For access to weights: https://bria.ai/contact-us

image generation
text-to-audio
falREVIEW REQUIRED

Kokoro TTS (Italian)

fal-ai/kokoro/italian

A high-quality Italian text-to-speech model delivering smooth and expressive speech synthesis.

speech
image-to-image
falREVIEW REQUIRED

Cartoonify

fal-ai/cartoonify

Transform images into 3D cartoon artwork using an AI model that applies cartoon stylization while preserving the original image's composition and details.

stylizedtransform
image-to-video
falREVIEW REQUIRED

High Quality Stable Video Diffusion

fal-ai/stable-video

Generate short video clips from your images using SVD v1.1

text-to-image
falREVIEW REQUIRED

Fooocus

fal-ai/fooocus

Default parameters with automated optimizations and quality improvements.

stylized
video-to-video
falREVIEW REQUIRED

Ltx 2.3 Quality

fal-ai/ltx-2.3-quality/extend-video

Extend high-quality video with audio from input video using LTX-2.3

extendlonger
text-to-image
falREVIEW REQUIRED

Sana v1.5 1.6B

fal-ai/sana/v1.5/1.6b

Sana v1.5 1.6B is a lightweight text-to-image model that delivers 4K image generation with impressive efficiency.

text to image4klightweight
image-to-video
falREVIEW REQUIRED

Vidu Image to Video

fal-ai/vidu/q1/image-to-video

Vidu Q1 Image to Video generates high-quality 1080p videos with exceptional visual quality and motion diversity from a single image

stylizedtransform
image-to-video
falREVIEW REQUIRED

Pika Scenes (v2.2)

fal-ai/pika/v2.2/pikascenes

Pika Scenes v2.2 creates videos from a images with high quality output.

editingeffectsanimation
text-to-video
falREVIEW REQUIRED

Kandinsky 6.0 Pro

fal-ai/kandinsky6-pro/text-to-video

Kandinsky 6.0 Pro is Kandinsky Lab's flagship text-to-video model, generating high-resolution clips with cinematic motion and precise prompt adherence.

text-to-videokandinskylightweight
image-to-image
falREVIEW REQUIRED

Image Preprocessors

fal-ai/image-preprocessors/midas

MiDaS depth estimation preprocessor.

depthpreprocessutilitycontrolnet
image-to-image
falREVIEW REQUIRED

Image Editing Color Correction

fal-ai/image-editing/color-correction

Perfect your photos with professional color grading, balanced tones, and vibrant yet natural colors

stylizedtransform
image-to-video
falREVIEW REQUIRED

Vidu

fal-ai/vidu/q2/image-to-video/pro

Use the latest Vidu Q2 models which much more better quality and control on your videos.

image-to-video
image-to-image
IdeogramREVIEW REQUIRED

Ideogram V2 Remix

fal-ai/ideogram/v2/remix

Reimagine existing images with Ideogram V2's remix feature. Create variations and adaptations while preserving core elements and adding new creative directions through prompt guidance.

realismtypography
text-to-video
falREVIEW REQUIRED

Kandinsky5

fal-ai/kandinsky5/text-to-video

Kandinsky 5.0 is a diffusion model for fast, high-quality text-to-video generation.

text-to-audio
falREVIEW REQUIRED

Stable Audio 3

fal-ai/stable-audio-3/small/music/base/text-to-audio

Stable Audio 3 Small Music Base is the foundational 459 million parameter checkpoint generating full music compositions up to 2 minutes from text prompts, intended as the unmodified base for fine-tuning.

musicon-devicelightweight
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image Edit Plus Lora Gallery

fal-ai/qwen-image-edit-plus-lora-gallery/integrate-product

Blend products into backgrounds with automatic perspective and lighting correction

stylizedtransform
training
RecraftREVIEW REQUIRED

Recraft V3 Create Style

fal-ai/recraft/v3/create-style

Recraft V3 Create Style is capable of creating unique styles for Recraft V3 based on your images.

stylevectorpersonalization
image-to-3d
hitem3dREVIEW REQUIRED

Hi3D Multiview to 3D

hitem3d/hi3d/multi-view-to-3d

Generate 3D models from multiple view images using Hi3D.

image-to-3dmultiview-to-3d3d
image-to-video
falREVIEW REQUIRED

LongCat Video

fal-ai/longcat-video/image-to-video/720p

Generate long videos in 720p/30fps from images using LongCat Video

text-to-video
falREVIEW REQUIRED

LTX Video-0.9.5

fal-ai/ltx-video-v095

Generate videos from prompts using LTX Video-0.9.5

videotext-video
image-to-image
falREVIEW REQUIRED

NAFNet-denoise

fal-ai/nafnet/denoise

Use NAFNet to fix issues like blurriness and noise in your images. This model specializes in image restoration and can help enhance the overall quality of your photography.

image-restorationdeblurdenoise
text-to-image
falREVIEW REQUIRED

ControlNet SDXL

fal-ai/fast-sdxl-controlnet-canny

Generate Images with ControlNet.

diffusioncontrolnetmanipulation
text-to-video
MiniMaxREVIEW REQUIRED

H3 Max Retro Toon 70s

minimax/h3-max/styles/retro-toon-70s

Generates 768p video with audio in a retro 1970s hand-painted animation style from text prompts or an optional first-frame image. Supports durations of 5–15 seconds.

retro1970shand-paintedanimation