image-to-videoVidu Reference to Video
fal-ai/vidu/reference-to-videoVidu Reference to Video creates videos by using a reference images and combining them with a prompt.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
image-to-videofal-ai/vidu/reference-to-videoVidu Reference to Video creates videos by using a reference images and combining them with a prompt.
image-to-imagebria/product-dimensionsBria Product Dimensions turns one product photo and its measurements into a marketplace-ready dimension image with callout lines, labels, and weight or capacity readouts
image-to-imagefal-ai/image-preprocessors/hedHolistically-Nested Edge Detection (HED) preprocessor.
audio-to-audiofal-ai/stable-audio-25/inpaintGenerate high quality music and sound effects using Stable Audio 2.5 from StabilityAI
text-to-videofal-ai/wan-t2v-loraAdd custom LoRAs to Wan-2.1 is a text-to-video model that generates high-quality videos with high visual quality and motion diversity from images
image-to-3dfal-ai/hunyuan3d/v2/multi-view/turboGenerate 3D models from your images using Hunyuan 3D. A native 3D generative model enabling versatile and high-quality 3D asset creation.
image-to-imagefal-ai/flux-general/differential-diffusionA specialized FLUX endpoint combining differential diffusion control with LoRA, ControlNet, and IP-Adapter support, enabling precise, region-specific image transformations through customizable change maps.
video-to-videomoonvalley/marey/motion-transferPull motion from a reference video and apply it to new subjects or scenes.
video-to-video
image-to-videofal-ai/longcat-video/image-to-video/480pGenerate long videos from images using LongCat Video
text-to-imagefal-ai/sensenova-u1-infographicGenerate Infographic Image with Sensenova U1
text-to-videofal-ai/krea-wan-14b/text-to-videoFast Text-to-Video endpoint for Krea's Wan 14b model.
text-to-imagefal-ai/wan/v2.2-5b/text-to-imageWan 2.2's 5B model generates high-resolution, photorealistic images with powerful prompt understanding and fine-grained visual detail
text-to-videofal-ai/kling-video/lipsync/text-to-videoKling LipSync is a text-to-video model that generates realistic lip movements from text input.
text-to-audiofal-ai/kokoro/mandarin-chineseA highly efficient Mandarin Chinese text-to-speech model that captures natural tones and prosody.
text-to-videofal-ai/ltx-2.3-quality/text-to-videoGenerate high-quality video with audio from text using LTX-2.3
video-to-videofal-ai/ltx-2.3-22b/video-to-videoGenerate video with audio from videos using LTX-2.3
image-to-imagefal-ai/hidream-i1-full/image-to-imageHiDream-I1 full is a new open-source image generative foundation model with 17B parameters that achieves state-of-the-art image generation quality within seconds.
text-to-imagebria/fibo-bbq-preview/generateA preview to the next level of control of Text-to-Image models.
image-to-imageideogram/v4/image-to-image/loraIdeogram V4.0q Image-to-Image LoRA applies a custom-trained LoRA on top of an input image, steering edits toward a specific style, subject, or brand identity while keeping the source composition intact.
image-to-imagefal-ai/image-editing/wojak-styleTransform your photos into wojak style while keeping the original characters likeness
audio-to-audiofal-ai/stable-audio-3/small/sfx/audio-to-audioStable Audio 3 Small SFX audio-to-audio is a 459 million parameter latent diffusion model that transforms input audio into new sound-effect variations guided by text prompts.
image-to-imagefal-ai/luma-photon/modifyEdit images from your prompts using Luma Photon. Photon is the most creative, personalizable, and intelligent visual models for creatives, bringing a step-function change in the cost of high-quality image generation.
video-to-videomirelo-ai/sfx-v1/video-to-videoGenerate synced sounds for any video, and return it with its new sound track (like MMAudio)
text-to-imagefal-ai/fast-lcm-diffusionRun SDXL at the speed of light
text-to-videofal-ai/pixverse/v4/text-to-videoGenerate high quality video clips from text and image prompts using PixVerse v4
text-to-jsonbria/fibo/generate/structured_promptStructured Prompt Generation endpoint for Fibo, Bria's SOTA Open source model.
text-to-imagefal-ai/bitdanceImage generation with BitDance. Fast, high-resolution photorealistic images using an autoregressive LLM— for efficient, high-quality results.