text-to-videoKling Video
fal-ai/kling-video/o3/4k/text-to-videoKling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
text-to-videofal-ai/kling-video/o3/4k/text-to-videoKling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling
text-to-videofal-ai/wan/v2.2-5b/text-to-video/fast-wanWan 2.2's 5B FastVideo model produces up to 5 seconds of video 720p at 24FPS with fluid motion and powerful prompt understanding
text-to-videofal-ai/pixverse/c1/text-to-videoGenerate film-grade videos from text prompts with native audio, up to 1080p and 15 seconds, using PixVerse C1.
text-to-videofal-ai/ltx-video-13b-distilledGenerate videos from prompts using LTX Video-0.9.7 13B Distilled and custom LoRA
text-to-videofal-ai/hunyuan-videoHunyuan Video is an Open video generation model with high visual quality, motion diversity, text-video alignment, and generation stability. This endpoint generates videos from text descriptions.
text-to-videofal-ai/minimax/hailuo-02/pro/text-to-videoMiniMax Hailuo-02 Text To Video API (Pro, 1080p): Advanced video generation model with 1080p resolution
text-to-videobytedance/seedance-2.0/us/text-to-videoUS hosted version of ByteDance's most advanced text-to-video model. Cinematic output with native audio, multi-shot editing, real-world physics, and director-level camera control.
text-to-videofal-ai/pixverse/v4.5/text-to-videoGenerate high quality video clips from text and image prompts using PixVerse v4.5
text-to-videominimax/h3-max/styles/vhsGenerates 768p VHS-style video with audio from text prompts or an optional first-frame image. Supports 5–15 second clips and adjustable tape damage, from subtle analog noise to strong tracking distortion.
text-to-videofal-ai/pixverse/v5.5/text-to-videoGenerate high quality video clips from text and image prompts using PixVerse v5.5
text-to-videominimax/h3/text-to-video/loraGenerate video with synchronized audio from a text prompt using MiniMax H3; load a trained LoRA at adjustable strength to lock in style, character, or motion.
text-to-videofal-ai/hunyuan-video-v1.5/text-to-videoHunyuan Video 1.5 is Tencent's latest and best video model
text-to-video
text-to-videofal-ai/minimax/video-01-liveGenerate video clips from your prompts using MiniMax model
text-to-videofal-ai/cogvideox-5bGenerate videos from prompts using CogVideoX-5B
text-to-videofal-ai/heygen/v3/video-agentGenerate videos with a single prompt. Describe what you want in plain text, and the agent handles avatar selection, scripting, scene composition - all in one.
text-to-videofal-ai/heygen/avatar5/digital-twinCreate natural HeyGen Avatar V digital twin videos from text or audio, with lip-sync, optional backgrounds, captions, and MP4/WebM output.
text-to-videofal-ai/ltxv-13b-098-distilledGenerate long videos from prompts using LTX Video-0.9.8 13B Distilled and custom LoRA
text-to-videofal-ai/kling-video/v1.6/standard/effectsGenerate video clips from your prompts using Kling 1.6 (std)
text-to-videofal-ai/kandinsky5-pro/text-to-videoKandinsky 5.0 Pro is a diffusion model for fast, high-quality text-to-video generation.
text-to-videominimax/h3-max/styles/retro-toon-70sGenerates 768p video with audio in a retro 1970s hand-painted animation style from text prompts or an optional first-frame image. Supports durations of 5–15 seconds.
fal-ai/ltx-2.3-22b/text-to-videoGenerate video with audio from text using LTX-2.3
text-to-videofal-ai/ltx-2.3-22b/distilled/text-to-videoGenerate video with audio from text using LTX-2.3 Distilled
text-to-videofal-ai/kling-video/v1.6/pro/effectsGenerate video clips from your prompts using Kling 1.6 (pro)
text-to-videofal-ai/longcat-video/text-to-video/720pGenerate long videos in 720p/30fps from text using LongCat Video
text-to-videofal-ai/kandinsky5/text-to-video/distillKandinsky 5.0 Distilled is a lightweight diffusion model for fast, high-quality text-to-video generation.
text-to-videominimax/h3-max/styles/16bit-pixelGenerates 768p video with audio in a 16-bit pixel-art style from text prompts or an optional first-frame image. Supports durations of 5–15 seconds.
text-to-videofal-ai/longcat-video/distilled/text-to-video/480pGenerate long videos from text using LongCat Video Distilled