text-to-imageWan v2.2 A14B Text-to-Image A14B with LoRAs
fal-ai/wan/v2.2-a14b/text-to-image/loraWan 2.2's 14B model with LoRA support generates high-fidelity images with enhanced prompt alignment, style adaptability.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
text-to-imagefal-ai/wan/v2.2-a14b/text-to-image/loraWan 2.2's 14B model with LoRA support generates high-fidelity images with enhanced prompt alignment, style adaptability.
text-to-videofal-ai/infinitalk/single-textInfinitalk model generates a talking avatar video from a text and audio file. The avatar lip-syncs to the provided audio with natural facial expressions.
text-to-imagefal-ai/bria/text-to-image/baseBria's Text-to-Image model, trained exclusively on licensed data for safe and risk-free commercial use. Available also as source code and weights. For access to weights: https://bria.ai/contact-us
text-to-audiofal-ai/kokoro/italianA high-quality Italian text-to-speech model delivering smooth and expressive speech synthesis.
image-to-imagefal-ai/cartoonifyTransform images into 3D cartoon artwork using an AI model that applies cartoon stylization while preserving the original image's composition and details.
image-to-videofal-ai/stable-videoGenerate short video clips from your images using SVD v1.1
text-to-imagefal-ai/fooocusDefault parameters with automated optimizations and quality improvements.
video-to-videofal-ai/ltx-2.3-quality/extend-videoExtend high-quality video with audio from input video using LTX-2.3
text-to-imagefal-ai/sana/v1.5/1.6bSana v1.5 1.6B is a lightweight text-to-image model that delivers 4K image generation with impressive efficiency.
image-to-videofal-ai/vidu/q1/image-to-videoVidu Q1 Image to Video generates high-quality 1080p videos with exceptional visual quality and motion diversity from a single image
image-to-videofal-ai/pika/v2.2/pikascenesPika Scenes v2.2 creates videos from a images with high quality output.
text-to-videofal-ai/kandinsky6-pro/text-to-videoKandinsky 6.0 Pro is Kandinsky Lab's flagship text-to-video model, generating high-resolution clips with cinematic motion and precise prompt adherence.
image-to-imagefal-ai/image-preprocessors/midasMiDaS depth estimation preprocessor.
image-to-imagefal-ai/image-editing/color-correctionPerfect your photos with professional color grading, balanced tones, and vibrant yet natural colors
image-to-videofal-ai/vidu/q2/image-to-video/proUse the latest Vidu Q2 models which much more better quality and control on your videos.
image-to-imagefal-ai/ideogram/v2/remixReimagine existing images with Ideogram V2's remix feature. Create variations and adaptations while preserving core elements and adding new creative directions through prompt guidance.
text-to-videofal-ai/kandinsky5/text-to-videoKandinsky 5.0 is a diffusion model for fast, high-quality text-to-video generation.
text-to-image
text-to-video
text-to-audiofal-ai/stable-audio-3/small/music/base/text-to-audioStable Audio 3 Small Music Base is the foundational 459 million parameter checkpoint generating full music compositions up to 2 minutes from text prompts, intended as the unmodified base for fine-tuning.
image-to-imagefal-ai/qwen-image-edit-plus-lora-gallery/integrate-productBlend products into backgrounds with automatic perspective and lighting correction
trainingfal-ai/recraft/v3/create-styleRecraft V3 Create Style is capable of creating unique styles for Recraft V3 based on your images.
image-to-3dhitem3d/hi3d/multi-view-to-3dGenerate 3D models from multiple view images using Hi3D.
image-to-videofal-ai/longcat-video/image-to-video/720pGenerate long videos in 720p/30fps from images using LongCat Video
text-to-videofal-ai/ltx-video-v095Generate videos from prompts using LTX Video-0.9.5
image-to-imagefal-ai/nafnet/denoiseUse NAFNet to fix issues like blurriness and noise in your images. This model specializes in image restoration and can help enhance the overall quality of your photography.
text-to-imagefal-ai/fast-sdxl-controlnet-cannyGenerate Images with ControlNet.
text-to-videominimax/h3-max/styles/retro-toon-70sGenerates 768p video with audio in a retro 1970s hand-painted animation style from text prompts or an optional first-frame image. Supports durations of 5–15 seconds.