text-to-videoH3 Max Low Poly
minimax/h3-max/styles/low-polyGenerates 768p video with audio in a retro low-poly 3D style from text prompts or an optional first-frame image. Supports durations of 5–15 seconds.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
text-to-videominimax/h3-max/styles/low-polyGenerates 768p video with audio in a retro low-poly 3D style from text prompts or an optional first-frame image. Supports durations of 5–15 seconds.
text-to-imagefal-ai/bria/text-to-image/hdBria's Text-to-Image model for HD images. Trained exclusively on licensed data for safe and risk-free commercial use. Available also as source code and weights. For access to weights: https://bria.ai/contact-us
video-to-videofal-ai/cosmos-predict-2.5/video-to-videoGenerate video from text and videos using NVIDIA's 2B Cosmos Post-Trained Model
image-to-imagefal-ai/unoAn AI model that transforms input images into new ones based on text prompts, blending reference visuals with your creative directions.
image-to-imagerundiffusion-fal/juggernaut-flux-lora/inpaintingJuggernaut Base Flux LoRA Inpainting by RunDiffusion is a drop-in replacement for Flux [Dev] inpainting that delivers sharper details, richer colors, and enhanced realism to all your LoRAs and LyCORIS with full compatibility.
text-to-videofal-ai/fast-svd-lcm/text-to-videoGenerate short video clips from your images using SVD v1.1 at Lightning Speed
image-to-imagefal-ai/pasdPixel-Aware Diffusion Model for Realistic Image Super-Resolution and Personalized Stylization
video-to-videofal-ai/ltx-2.3-quality/colorizationColorize high-quality video using LTX-2.3
image-to-imagefal-ai/qwen-image-edit-plus-lora-gallery/group-photoCreate group photos
text-to-imagefal-ai/ernie-image/lora/turboHigh-quality text-to-image model by Baidu. Supports English, Chinese, and Japanese prompts with built-in prompt expansion.
trainingfal-ai/qwen-image-edit-2509-trainerLoRA trainer for Qwen Image Edit 2509
image-to-imagefal-ai/image-editing/time-of-dayTransform your photos to any time of day, from golden hour to midnight, with appropriate lighting and atmosphere.
image-to-imagefal-ai/image-editing/broccoli-haircutTransform your character's hair into broccoli style while keeping the original characters likeness
video-to-videofal-ai/ltx-video-v095/multiconditioningGenerate videos from prompts,images, and videos using LTX Video-0.9.5
image-to-imagefal-ai/image-editing/scene-compositionPlace your subject in any scene you imagine, from enchanted forests to urban settings, with professional composition and lighting
image-to-imagefal-ai/qwen-image-edit-2509-lora-gallery/remove-elementRemove unwanted elements (objects, people, text) while maintaining image consistency
image-to-imagefal-ai/qwen-image-edit-plus-lora-gallery/shirt-designApply designs/graphics onto people's shirts
visionfal-ai/sa2va/8b/videoSa2VA is an MLLM capable of question answering, visual prompt understanding, and dense object segmentation at both image and video levels
video-to-videofal-ai/one-to-all-animation/1.3bOne-to-All Animation is a pose driven video model that animates characters from a single reference image, enabling flexible, alignment-free motion transfer across diverse styles and scenes
llmnvidia/nemotron-3-nano-omniOpen, efficient reasoning model from NVIDIA. 30B A3B hybrid Transformer-Mamba MoE, built for enterprise agentic workflows.
image-to-videofal-ai/ltx-2.3-quality/ingredientGenerate high-quality video with audio from reference, character sheet, storyboard using LTX-2.3
text-to-videomoonvalley/marey/t2vGenerate a video from a text prompt with Marey, a generative video model trained exclusively on fully licensed data.
image-to-imagefal-ai/chrono-edit-lora-gallery/upscalerUpscales and cleans up the image.
image-to-videofal-ai/amt-interpolation/frame-interpolationInterpolate between image frames
image-to-imagefal-ai/fast-sdxl-controlnet-canny/image-to-imageGenerate Images with ControlNet.
text-to-videofal-ai/wan/v2.2-5b/text-to-video/distillWan 2.2's 5B distill model produces up to 5 seconds of video 720p at 24FPS with fluid motion and powerful prompt understanding
image-to-imagefal-ai/ideogram/v2a/remixCreate variations of existing images with Ideogram V2A Remix while maintaining creative control through prompt guidance.
visionfal-ai/nemotron-diffusion-vlmNemotron-Labs-Diffusion-VLM-8B is the vision-language extension of the Nemotron-Labs-Diffusion family.