EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

All model endpoints

Page 16 · 28 per page
video-to-video
falREVIEW REQUIRED

Ltx 2.3

fal-ai/ltx-2.3/reframe

LTX-2.3 Reframe converts your videos to any aspect ratio without destructive cropping. It intelligently recenters the original footage and generatively fills the newly exposed areas with content that seamlessly matches the scene, so the result looks like it was shot natively in the target format. Turn landscape footage into vertical 9:16 for social, square 1:1 for feeds, or anything in between. Supports videos up to 60 seconds, with 720p and 1080p outputs across 1:1, 4:5, 5:4, 9:16 and 16:9.

reframesize
text-to-speech
resemble-aiREVIEW REQUIRED

Chatterboxhd

resemble-ai/chatterboxhd/text-to-speech

Generate expressive, natural speech with Resemble AI's Chatterbox. Features unique emotion control, instant voice cloning from short audio, and built-in watermarking.

json
falREVIEW REQUIRED

Ffmpeg Api

fal-ai/ffmpeg-api/loudnorm

Get EBU R128 loudness normalization from audio files using FFmpeg API.

ffmpeg
image-to-image
falREVIEW REQUIRED

Sam 3

fal-ai/sam-3/image-rle

SAM 3 is a unified foundation model for promptable segmentation in images and videos. It can detect, segment, and track objects using text or visual prompts such as points, boxes, and masks.

segmentationrlereal-time
image-to-image
AlibabaREVIEW REQUIRED

Wan v2.6 Image to Image

wan/v2.6/image-to-image

Wan 2.6 image-to-image model.

image-to-image
text-to-video
falREVIEW REQUIRED

LTX 2.3 Video Fast

fal-ai/ltx-2.3/text-to-video/fast

LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.

stylizedtransformlipsync
image-to-3d
falREVIEW REQUIRED

Hyper3d

fal-ai/hyper3d/rodin/v2

Rodin by Hyper3D generates realistic and production ready 3D models from text or images.

image-to-3dtext-to-3d
image-to-video
ByteDanceREVIEW REQUIRED

Seedance 2.0 US Image to Video

bytedance/seedance-2.0/us/image-to-video

US hosted version of ByteDance's most advanced image-to-video model. Animate still images into cinematic video with synchronized audio, start and end frame control, and motion prompts.

stylizedtransformlipsync
video-to-video
falREVIEW REQUIRED

RIFE

fal-ai/rife/video

Interpolate videos with RIFE - Real-Time Intermediate Flow Estimation

interpolation
image-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 9B Base LoRA

fal-ai/flux-2/klein/9b/base/edit/lora

Image-to-image editing with LoRA support for FLUX.2 [klein] 9B Base from Black Forest Labs. Specialized style transfer and domain-specific modifications.

text-to-image
KlingREVIEW REQUIRED

Kling Image

fal-ai/kling-image/v3/text-to-image

Kling V3: Latest Kling Image model

text-to-image
image-to-video
AlibabaREVIEW REQUIRED

Happy Horse 1.1 Image to Video

alibaba/happy-horse/v1.1/image-to-video

Happy Horse 1.1 is Alibaba's #1-ranked video model. This image-to-video endpoint animates a still image into 1080p video with synchronized native audio and multilingual lip-sync

happy-horsevideoimage
text-to-image
MiniMaxREVIEW REQUIRED

MiniMax (Hailuo AI) Text to Image

fal-ai/minimax/image-01

Generate high quality images from text prompts using MiniMax Image-01. Longer text prompts will result in better quality images.

stylizedrealism
image-to-image
falREVIEW REQUIRED

PATINA

fal-ai/patina

PATINA creates seamless high-resolution normal, roughness, basecolor (albedo), height (displacement) and metalness maps from images

pbrdisplacementmetalnessnormal
text-to-video
AlibabaREVIEW REQUIRED

Wan Text to Video

fal-ai/wan/v2.7/text-to-video

Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylizedtransformlipsync
text-to-image
microsoftREVIEW REQUIRED

Mai Image 2.5 Text to Image

microsoft/mai-image-2.5

MAI-Image-2.5 is Microsoft's photorealistic image generation and editing model that turns text prompts or uploaded images into high-quality, design-ready visuals with fine-grained, pixel-level control.

realismtypographystylized
image-to-video
AlibabaREVIEW REQUIRED

Wan v2.2 5B

fal-ai/wan/v2.2-5b/image-to-video

Wan 2.2's 5B model produces up to 5 seconds of video 720p at 24FPS with fluid motion and powerful prompt understanding

image-to-3d
falREVIEW REQUIRED

Hunyuan3D

fal-ai/hunyuan3d/v2

Generate 3D models from your images using Hunyuan 3D. A native 3D generative model enabling versatile and high-quality 3D asset creation.

stylized
video-to-video
Luma AIREVIEW REQUIRED

Luma Ray 3.2 Video to Video

luma/agent/ray/v3.2/video-to-video

Luma Ray 3.2 re-renders an existing video into new cinematic motion guided by a text prompt, preserving the source's look and movement while controlling resolution, duration, and HDR.

stylizedtransformlipsync
video-to-video
Luma AIREVIEW REQUIRED

Luma Ray 3.2 Reframe

luma/agent/ray/v3.2/reframe

Luma Ray 3.2 reframes an existing video into a new aspect ratio guided by a text prompt, preserving the original footage frame-for-frame while controlling resolution and outpainting the surrounding canvas.

stylizedtransformlipsync
image-to-image
microsoftREVIEW REQUIRED

Mai Image 2.5

microsoft/mai-image-2.5/edit

MAI-Image-2.5 is Microsoft's photorealistic image generation and editing model that turns text prompts or uploaded images into high-quality, design-ready visuals with fine-grained, pixel-level control.

realismtypographystylized
text-to-image
OpenAIREVIEW REQUIRED

GPT Image 1 Mini

fal-ai/gpt-image-1-mini

GPT Image 1 mini combines OpenAI's advanced language capabilities, powered by GPT-5, with GPT Image 1 Mini for efficient image generation.

text-to-image
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Krea [dev]

fal-ai/flux/krea

FLUX.1 Krea [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.

text-to-video
AlibabaREVIEW REQUIRED

Wan-2.2 Text-to-Video A14B

fal-ai/wan/v2.2-a14b/text-to-video

Wan-2.2 text-to-video is a video model that generates high-quality videos with high visual quality and motion diversity from text prompts.

text to videomotion
text-to-image
falREVIEW REQUIRED

Hunyuan Image

fal-ai/hunyuan-image/v3/text-to-image

Leverage the state-of-the-art capabilities of Hunyuan Image 3.0 to generate visual content that effectively conveys the messaging of your written material.

text-to-image
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] Inpainting with LoRAs

fal-ai/flux-lora/inpainting

Super fast endpoint for the FLUX.1 [dev] inpainting model with LoRA support, enabling rapid and high-quality image inpaingting using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

lorapersonalization