EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 26 · 28 per page
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 9B Base LoRA

fal-ai/flux-2/klein/9b/base/lora

Text-to-image generation with LoRA support for FLUX.2 [klein] 9B Base from Black Forest Labs. Custom style adaptation and fine-tuned model variations.

text-to-image
falREVIEW REQUIRED

Hidream I1 Full

fal-ai/hidream-i1-full

HiDream-I1 full is a new open-source image generative foundation model with 17B parameters that achieves state-of-the-art image generation quality within seconds.

text-to-video
falREVIEW REQUIRED

Wan-2.1 Text-to-Video

fal-ai/wan-t2v

Wan-2.1 is a text-to-video model that generates high-quality videos with high visual quality and motion diversity from text prompts

text to videomotion
video-to-video
falREVIEW REQUIRED

Heygen

fal-ai/heygen/v2/translate/precision

Heygen Translate Model with Extreme Precision

video-to-video
image-to-image
falREVIEW REQUIRED

Wan 2.5 Image to Image

fal-ai/wan-25-preview/image-to-image

Wan 2.5 image-to-image model.

image-to-image
Black Forest LabsREVIEW REQUIRED

Flux 2 Lora Gallery

fal-ai/flux-2-lora-gallery/apartment-staging

Virtually furnishes an empty apartment

stylizedtransform
text-to-image
falREVIEW REQUIRED

Stable Diffusion V3

fal-ai/stable-diffusion-v3-medium

Stable Diffusion 3 Medium (Text to Image) is a Multimodal Diffusion Transformer (MMDiT) model that improves image quality, typography, prompt understanding, and efficiency.

diffusionstyle
image-to-video
falREVIEW REQUIRED

LTX-2.3 22B Distilled

fal-ai/ltx-2.3-22b/distilled/image-to-video

Generate video with audio from images using LTX-2.3 Distilled

image-to-image
briaREVIEW REQUIRED

Extract Object

bria/extract-object

Bria Extract Object uses text prompts to isolate a selected object from an image and return it as an RGBA PNG with a transparent background. Ideal for product, ecommerce, advertising, and creative editing workflows. Bria's Extract Object API leads in product shot extraction, outperforming SAM 3.1 where it counts most for commercial use.

text-to-image
falREVIEW REQUIRED

Krea 2 Text to Image Turbo Style

fal-ai/krea-2/turbo/style

Generate high-fidelity images from text with Krea 2 using a style reference image. Apply a reference image to guide the visual style into new generations, with aspect ratio, creativity, and seed controls.

stylizedstyle transferreference imagerealism
image-to-image
falREVIEW REQUIRED

Image Editing Object Removal

fal-ai/image-editing/object-removal

Remove unwanted objects or people from your photos while seamlessly blending the background.

stylizedtransform
image-to-image
falREVIEW REQUIRED

Hunyuan World

fal-ai/hunyuan_world

Hunyuan World 1.0 turns a single image into a panorama or a 3D world. It creates realistic scenes from the image, allowing you to explore and view it from different angles.

image-to-image
smoretalk-aiREVIEW REQUIRED

Rembg Enhance (Remove Background Enhance)

smoretalk-ai/rembg-enhance

Rembg-enhance is optimized for 2D vector images, 3D graphics, and photos by leveraging matting technology.

background removalimage editingutilitysegmentation
text-to-speech
KlingREVIEW REQUIRED

Kling TTS

fal-ai/kling-video/v1/tts

Generate speech from text prompts and different voices using the Kling TTS model, which leverages advanced AI techniques to create high-quality text-to-speech.

audio
audio-to-video
lightricksREVIEW REQUIRED

Ltx 2.5 Audio to Video Fast

lightricks/ltx-2.5/audio-to-video/fast

LTX-2.5 is Lightricks' open-source audio-video model. This endpoint generates video timed to a supplied audio clip in a speed-optimized mode — useful for music-driven content, dialogue-led shorts, and ads keyed to a track.

stylizedtransformlip-sync
image-to-image
briaREVIEW REQUIRED

Upscale

bria/upscale/creative

Professional-grade creative upscaler that doubles resolution up to 10MP, regenerating sharper textures, refined details, and cleaner faces. Trained exclusively on licensed data for risk-free commercial use.

briaaestheticsupscaler
image-to-image
topazREVIEW REQUIRED

Topaz Sharpen Image

topaz/sharpen/image

Professional photo sharpening powered by Topaz Labs. Models tuned per blur type (lens, motion, portrait, wildlife), plus Super Focus for generative recovery of severely blurred shots. Best for out-of-focus and motion-blurred photos.

sharpenimage
vision
falREVIEW REQUIRED

MoonDreamNext

fal-ai/moondream-next

MoonDreamNext is a multimodal vision-language model for captioning, gaze detection, bbox detection, point detection, and more.

multimodalvision
text-to-speech
MiniMaxREVIEW REQUIRED

Minimax

fal-ai/minimax/preview/speech-2.5-hd

Generate speech from text prompts and different voices using the MiniMax Speech-02 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.

speech
image-to-video
falREVIEW REQUIRED

LTX Video-0.9.7 13B Distilled

fal-ai/ltx-video-13b-distilled/image-to-video

Generate videos from prompts and images using LTX Video-0.9.7 13B Distilled and custom LoRA

videoltx-videoimage-to-video
image-to-video
falREVIEW REQUIRED

Wan-2.1 First-Last-Frame-to-Video

fal-ai/wan-flf2v

Wan-2.1 flf2v generates dynamic videos by intelligently bridging a given first frame to a desired end frame through smooth, coherent motion sequences.

image to videomotion
text-to-video
MiniMaxREVIEW REQUIRED

MiniMax (Hailuo AI) Video 01

fal-ai/minimax/video-01

Generate video clips from your prompts using MiniMax model

motiontransformation
text-to-speech
falREVIEW REQUIRED

Maya1

fal-ai/maya

Maya1 is a state-of-the-art speech model by Maya Research for expressive voice generation, built to capture real human emotion and precise voice design.

text-to-speechtts
image-to-image
falREVIEW REQUIRED

Age Modify

fal-ai/image-apps-v2/age-modify

Modify a face to look younger or older while keeping identity realistic.

age-transformationface-editing
image-to-image
AlibabaREVIEW REQUIRED

Qwen Image

fal-ai/qwen-image/image-to-image

Qwen-Image (Image-to-Image) transforms and edits input images with high fidelity, enabling precise style transfer, enhancement, and creative modification.

image-to-image
audio-to-audio
falREVIEW REQUIRED

Stable Audio 3 Medium Audio to Audio

fal-ai/stable-audio-3/medium/audio-to-audio

Stable Audio 3 Medium audio-to-audio is a 1.4 billion parameter latent diffusion model that transforms an input audio clip into new stereo variations up to 6 minutes guided by a text prompt.

musicstyle-transferremix
vision
falREVIEW REQUIRED

Moondream3 Preview [Point]

fal-ai/moondream3-preview/point

Moondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.

Vision
text-to-image
RecraftREVIEW REQUIRED

Recraft 20b

fal-ai/recraft-20b

Recraft 20b is a new and affordable text-to-image model.

image generationvector arttypographstyle