image-to-imageSeedVR2
fal-ai/seedvr/upscale/image/seamlessUse SeedVR2 to upscale images, retaining seamless tiling
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
image-to-imagefal-ai/seedvr/upscale/image/seamlessUse SeedVR2 to upscale images, retaining seamless tiling
text-to-videoxai/grok-imagine-video/v1.5/text-to-videoGenerate videos from prompts with audio using xAI's Grok Imagine 1.5 Video model.
image-to-imagefal-ai/wan/v2.7/pro/editEdit and transform images using text instructions with the WAN 2.7 Pro model for precise, professional-grade image modifications.
image-to-videofal-ai/minimax/hailuo-2.3/pro/image-to-videoMiniMax Hailuo-2.3 Image To Video API (Pro, 1080p): Advanced image-to-video generation model with 1080p resolution
text-to-imagekrea/v2/medium/text-to-imageGenerate high-quality images from text with Krea 2 Medium, supporting aspect ratio, creativity controls, seeds, and optional style references.
image-to-imagefal-ai/image-apps-v2/outpaintDirectional outpainting. Choose edges to expand. left, right, top, or center (uniform all sides). Only expanded areas are generated; an optional zoom-out pulls the frame back by the chosen amount.
unknownbytedance/seedance-2.5/draft/completeDraft completion endpoint for Seedance 2.5 - submit a draft id to regenerate the task at 1080p.
video-to-videoalibaba/happy-horse/video-editHappyHorse video editing supports advanced video editing through natural language instructions. It allows for local or global editing of video elements using up to 5 reference images.
image-to-videobytedance/seedance-2.0/us/reference-to-videoUS hosted version of ByteDance's most advanced reference-to-video model. Generate video from up to 9 images, 3 videos, and 3 audio clips with native audio and cinematic camera control.
text-to-speechfal-ai/qwen-3-tts/voice-design/1.7bCreate custom voices using Qwen3-TTS Voice Design model and later use Clone Voice model to create your own voices!
video-to-videogoogle/gemini-omni-flash/editEdits generated video across multiple conversational turns while preserving scene coherence. Applies iterative changes through natural-language instructions without regenerating the full sequence from scratch.
image-to-3dfal-ai/hyper3d/rodinRodin by Hyper3D generates realistic and production ready 3D models from text or images.
text-to-imagefal-ai/flux-generalA versatile endpoint for the FLUX.1 [dev] model that supports multiple AI extensions including LoRA, ControlNet conditioning, and IP-Adapter integration, enabling comprehensive control over image generation through various guidance methods.
text-to-imagefal-ai/gpt-image-1/text-to-imageOpenAI's latest image generation and editing model: gpt-1-image.
image-to-imagefal-ai/gpt-image-1/edit-imageOpenAI's latest image generation and editing model: gpt-1-image.
text-to-videofal-ai/kling-video/o3/standard/text-to-videoGenerate realistic videos using Kling O3 from Kling Team!
video-to-videofal-ai/wan/v2.7/edit-videoWan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.
text-to-speechfal-ai/chatterbox/text-to-speech/multilingualWhether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.
image-to-videoluma/agent/ray/v3.2/image-to-videoLuma Ray 3.2 animates a source image into cinematic motion guided by a text prompt, preserving the starting frame's look while controlling resolution, duration, and seamless looping.
text-to-imagefal-ai/recraft/v4/pro/text-to-imageRecraft V4 was developed with designers to bring true visual taste to AI image generation. Built for brand systems and production-ready workflows, it goes beyond prompt accuracy — delivering stronger composition, refined lighting, realistic materials, and a cohesive aesthetic. The result is imagery shaped by professional design judgment, ready for immediate real-world use without additional post-processing.
image-to-3dmeshy/v7/image-to-3dTurns a single image into a fully textured, PBR-ready 3D mesh with complete geometry, in game-ready Smart Topology at a target polygon count
video-to-videominimax/h3-max/insert-videoInsert a new scene into an existing video with H3 Max. Guide the scene with a prompt, reference images, or reference videos, then return to the original footage.
image-to-videofal-ai/ltx-2.3/image-to-videoLTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.
video-to-videoxai/grok-imagine-video/edit-videoEdit videos using xAI's Grok Imagine
text-to-videofal-ai/kling-video/v3/turbo/pro/text-to-videoGenerate high quality 1080p videos using Kling's Turbo 3.0 model, with improved lipsync and multishot generation capabilities.
text-to-speechgoogle/gemini-3.8-flash-lite-ttsGenerate expressive speech with Gemini 3.8 Flash Lite TTS. Choose from 30 voices, guide delivery with style instructions, and create single-speaker narration or two-speaker dialogue.
video-to-videoveed/video-background-removal/fastRemove background from any video with people and objects. No green screen needed.
image-to-videofal-ai/minimax/hailuo-02/pro/image-to-videoMiniMax Hailuo-02 Image To Video API (Pro, 1080p): Advanced image-to-video generation model with 1080p resolution