text-to-videoLTX-2.3 22B
fal-ai/ltx-2.3-22b/text-to-video/loraGenerate video with audio from text using LTX-2.3 and custom LoRA
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
text-to-videofal-ai/ltx-2.3-22b/text-to-video/loraGenerate video with audio from text using LTX-2.3 and custom LoRA
video-to-videofal-ai/ltx-2.3-22b/reference-video-to-video/loraGenerate video with audio from reference video, text and images using LTX-2.3 and custom LoRA
video-to-videomoonvalley/marey/pose-transferIdeal for matching human movement. Your input video determines human poses, gestures, and body movements that will appear in the generated video.
image-to-imagefal-ai/flux-general/rf-inversionA general purpose endpoint for the FLUX.1 [dev] model, implementing the RF-Inversion pipeline. This can be used to edit a reference image based on a prompt.
audio-to-audiofal-ai/stable-audio-3/small/music/base/audio-outpaintingStable Audio 3 Small Music Base audio outpainting is the foundational 459 million parameter checkpoint that extends music tracks via causal continuation guided by text prompts.
text-to-imagefal-ai/flux-krea-lora/streamSuper fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.
trainingfal-ai/ltx23-trainer-v2/t2vFine-tune LTX 2.3 on your own clips to teach it a new subject, character, object, or visual style, then generate full videos from a text prompt.
image-to-imageideogram/v4/tiling/loraIdeogram V4.0q Tiling LoRA produces seamless repeatable patterns guided by a custom-trained LoRA, locking a specific aesthetic or motif into tileable textures for cohesive, large-scale surface design.
text-to-videofal-ai/bernini-r/text-to-videoGenerate high-quality video from a text prompt with Bernini-R, ByteDance's unified video generation and editing model.
image-to-imagefal-ai/fast-lcm-diffusion/inpaintingRun SDXL at the speed of light
trainingminimax/h3/flf2v/trainerTrain a MiniMax H3 LoRA on first/last/both keyframe signatures, teaching it to generate video with audio that starts on one image and lands on another.
audio-to-videofal-ai/ltx-2.3-22b/distilled/audio-to-video/loraGenerate video with audio from audio, text and images using LTX-2.3 Distilled and custom LoRA
image-to-imagefal-ai/playground-v25/inpaintingState-of-the-art open-source model in aesthetic quality
trainingfal-ai/flux-2-trainer/editFine-tune FLUX.2 [dev] from Black Forest Labs with custom datasets. Create specialized LoRA adaptations for specific editing tasks.
visionfal-ai/florence-2-large/region-to-descriptionFlorence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
image-to-imagefal-ai/flux/krea/reduxFLUX.1 Krea [dev] Redux is a high-performance endpoint for the FLUX.1 Krea [dev] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.
text-to-imagefal-ai/fooocus/upscale-or-varyDefault parameters with automated optimizations and quality improvements.
image-to-imagefal-ai/qwen-image-edit-2509-lora-gallery/integrate-productBlend products into backgrounds with automatic perspective and lighting correction
text-to-videofal-ai/fast-svd/text-to-videoGenerate short video clips from your prompts using SVD v1.1
visionfal-ai/moondream-next/batchMoonDreamNext Batch is a multimodal vision-language model for batch captioning.
video-to-videofal-ai/pixverse/extend/fastPixVerse Extend model is a video extending tool for your videos using with high-quality video extending techniques
audio-to-audiofal-ai/stable-audio-3/small/sfx/base/audio-to-audioStable Audio 3 Small SFX Base audio-to-audio is the foundational 459 million parameter checkpoint that transforms input audio into new sound-effect variations guided by text prompts.
text-to-videofal-ai/pixverse/v4/text-to-video/fastGenerate high quality and fast video clips from text and image prompts using PixVerse v4 fast
trainingideogram/v4/trainerTrain custom LoRAs for personalization, styles or other use cases on top of Ideogram V4.
image-to-videofal-ai/bernini-r/reference-to-videoTurn up to five reference images into one continuous, consistent video with Bernini-R, with smooth, stable camera motion and no scene cuts.
image-to-imagefal-ai/ideogram/v2/editTransform existing images with Ideogram V2's editing capabilities. Modify, adjust, and refine images while maintaining high fidelity and realistic outputs with precise prompt control.
visionfal-ai/scene-finderSearch any video with a text prompt - Scene Finder locates the matching moments and returns their time segments and extracted frames.
video-to-videofal-ai/ltx-video-v095/extendGenerate videos from prompts and videos using LTX Video-0.9.5