EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 32 · 28 per page
audio-to-audio
falREVIEW REQUIRED

Sam Audio

fal-ai/sam-audio/span-separate

Audio separation with SAM Audio. Isolate any sound using natural language—professional-grade audio editing made simple for creators, researchers, and accessibility applications.

audio-to-audiosam-audio
speech-to-speech
xAIREVIEW REQUIRED

Grok Voice

xai/grok-voice/realtime

Build real-time voice applications powered by Grok. Stream audio and text bidirectionally via WebSocket for voice assistants, phone agents, and interactive voice systems.

xaigrokvoiceagent
audio-to-video
PixVerseREVIEW REQUIRED

PixVerse VibeMV

pixverse/music-video/vibemv

PixVerse VibeMV generates music videos from audio, with optional character references and lyric subtitles. It supports visual style presets, custom style references, five aspect ratios, and output at 720p or 1080p.

audio-to-videomusic-videopixversevibemv
text-to-image
briaREVIEW REQUIRED

Fibo Gen 1.5 Text to Image

bria/fibo-gen-1.5/text-to-image

Text-to-image model with high-fidelity outputs, accurate typography, and style preset, strong in photorealism, textures, and beyond. JSON-structured prompts give enterprise and agentic workflows production-ready control. Trained on licensed data.

stylizedtransformrealism
image-to-video
falREVIEW REQUIRED

Ovi

fal-ai/ovi/image-to-video

Ovi can generate videos with audio from image and text inputs.

image-to-audio-videoimage-to-video
image-to-image
pixelcutREVIEW REQUIRED

Pixelcut Product Photo

pixelcut/product-photo

Pixelcut's Background Remover produces fast, high-quality cutouts built for e-commerce product imagery

utilityediting
vision
perceptronREVIEW REQUIRED

Isaac 0.1

perceptron/isaac-01

Isaac-01 is a multimodal vision-language model from Perceptron for various vision language tasks.

multimodalvision
image-to-video
MiniMaxREVIEW REQUIRED

MiniMax (Hailuo AI) Video 01

fal-ai/minimax/video-01-live/image-to-video

Generate video clips from your images using MiniMax Video model

motiontransformation
text-to-audio
falREVIEW REQUIRED

Kokoro TTS (French)

fal-ai/kokoro/french

An expressive and natural French text-to-speech model for both European and Canadian French.

speech
training
Black Forest LabsREVIEW REQUIRED

Turbo Flux Trainer

fal-ai/turbo-flux-trainer

A blazing fast FLUX dev LoRA trainer for subjects and styles.

text-to-audio
falREVIEW REQUIRED

Kokoro TTS (Brazilian Portuguese)

fal-ai/kokoro/brazilian-portuguese

A natural and expressive Brazilian Portuguese text-to-speech model optimized for clarity and fluency.

speech
text-to-image
falREVIEW REQUIRED

GLM Image

fal-ai/glm-image

Create high-quality images with accurate text rendering and rich knowledge details—supports editing, style transfer, and maintaining consistent characters across multiple images.

text-to-image
image-to-video
PixVerseREVIEW REQUIRED

PixVerse V4.5 Effects

fal-ai/pixverse/v4.5/effects

Generate high quality video clips with different effects using PixVerse v4.5

image-to-video
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Krea [dev] with LoRAs

fal-ai/flux-krea-lora/image-to-image

FLUX LoRA Image-to-Image is a high-performance endpoint that transforms existing images using FLUX models, leveraging LoRA adaptations to enable rapid and precise image style transfer, modifications, and artistic variations.

lorastyle transfer
workflow
falREVIEW REQUIRED

Workflow Utilities Pick Image By Index

fal-ai/workflow-utilities/pick-image-by-index

Choose the Nth image from an image URL list for workflows.

video-to-video
falREVIEW REQUIRED

Infinitalk

fal-ai/infinitalk/video-to-video

Infinitalk model generates a talking avatar video from an image and audio file. The avatar lip-syncs to the provided audio with natural facial expressions.

video-to-video
text-to-video
falREVIEW REQUIRED

Heygen v5 Digital Twin

fal-ai/heygen/avatar5/digital-twin

Create natural HeyGen Avatar V digital twin videos from text or audio, with lip-sync, optional backgrounds, captions, and MP4/WebM output.

avatardigital-twintalking-avatartext-to-video
image-to-3d
hitem3dREVIEW REQUIRED

Hi3D Image to 3D

hitem3d/hi3d/image-to-3d

Generate 3D models from a single image with Hi3D.

image-to-3d3dmesh
text-to-image
briaREVIEW REQUIRED

Fibo

bria/fibo/generate

SOTA open-source text-to-image model delivering high-fidelity outputs with accurate typography. JSON-structured prompts provide production-ready controllability for enterprise and agentic workflows. Trained exclusively on licensed data.

briafiboprompt-adherence
video-to-video
falREVIEW REQUIRED

Workflow Utilities Blend Video

fal-ai/workflow-utilities/blend-video

FFMPEG Utility for Blending Videos

text-to-image
falREVIEW REQUIRED

Longcat Image

fal-ai/longcat-image

LongCat image is a 6B parameter model excelling at multilingual text rendering, photorealism and deployment efficiency.

video-to-video
falREVIEW REQUIRED

Wan 2.2 VACE Fun A14B

fal-ai/wan-22-vace-fun-a14b/inpainting

VACE Fun for Wan 2.2 A14B from Alibaba-PAI

image-to-image
falREVIEW REQUIRED

Longcat Image

fal-ai/longcat-image/edit

LongCat image Edit is a 6B parameter image editing model excelling at multilingual text rendering, photorealism and deployment efficiency.

text-to-3d
MeshyREVIEW REQUIRED

Meshy 6 Preview

fal-ai/meshy/v6-preview/text-to-3d

Meshy-6-Preview is the latest model from Meshy. It generates realistic and production ready 3D models.

text-to-3d
text-to-video
MiniMaxREVIEW REQUIRED

MiniMax (Hailuo AI) Video 01 Live

fal-ai/minimax/video-01-live

Generate video clips from your prompts using MiniMax model

motiontransformation
image-to-image
falREVIEW REQUIRED

Image Editing Background Change

fal-ai/image-editing/background-change

Replace your photo's background with any scene you desire, from beach sunsets to urban landscapes, with perfect lighting and shadows

stylizedtransform