image-to-videoHappy Horse
alibaba/happy-horse/reference-to-videoGenerate 1080p video with synchronized native audio from a text prompt and references. Aspect ratios: 16:9, 9:16, 1:1, 4:3, 3:4. Duration: 3–15s.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
image-to-videoalibaba/happy-horse/reference-to-videoGenerate 1080p video with synchronized native audio from a text prompt and references. Aspect ratios: 16:9, 9:16, 1:1, 4:3, 3:4. Duration: 3–15s.
image-to-videofal-ai/minimax/hailuo-02-fast/image-to-videoCreate blazing fast and economical videos with MiniMax Hailuo-02 Image To Video API at 512p resolution
image-to-imagefal-ai/image-apps-v2/virtual-try-onTry on clothes virtually by combining person and clothing images.
audio-to-audiofal-ai/deepfilternet3Enhance speech audio by removing background noise and upsampling to 48KHz
video-to-videofal-ai/flashvsr/upscale/videoUpscale your videos using FlashVSR with the fastest speeds!
video-to-videofal-ai/void-video-inpaintingVOID removes objects from videos along with all interactions they induce on the scene
fal-ai/video-upscalerThe video upscaler endpoint uses RealESRGAN on each frame of the input video to upscale the video to a higher resolution.
text-to-imageideogram/v4/instantGenerate high-quality images, posters, and logos with Ideogram's latest V4.0q — producing crisp visuals with accurate text rendering, fine detail, and full creative control for polished, ready-to-use designs FRACTION OF A SECOND.
image-to-3dfal-ai/meshy/v6-preview/image-to-3dMeshy-6-Preview is the latest model from Meshy. It generates realistic and production ready 3D models.
text-to-3dfal-ai/hunyuan-motionGenerate 3D human motions via text-to-generation interface of Hunyuan Motion!
image-to-imagefal-ai/kling-image/o1Perform precise image edits using strong reference control, transforming subjects, styles, and local details while preserving visual consistency.
image-to-videofal-ai/bytedance/omnihumanOmniHuman generates video using an image of a human figure paired with an audio file. It produces vivid, high-quality videos where the character’s emotions and movements maintain a strong correlation with the audio.
image-to-imagefal-ai/ideogram/v3/reframeExtend existing images with Ideogram V3's reframe feature. Create expanded versions and adaptations while preserving main image and adding new creative directions through prompt guidance.
video-to-videobria/video/background-removalAutomatically remove backgrounds from videos -perfect for creating clean, professional content without a green screen.
image-to-imagefal-ai/flux-lora-fillFLUX.1 [dev] Fill is a high-performance endpoint for the FLUX.1 [pro] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.
image-to-imagefal-ai/iclight-v2An endpoint for re-lighting photos and changing their backgrounds per a given description
image-to-videofal-ai/pixverse/c1/image-to-videoAnimate images into cinematic videos with PixVerse C1, supporting 1080p resolution and native audio generation.
text-to-imagefal-ai/ideogram/v2Generate high-quality images, posters, and logos with Ideogram V2. Features exceptional typography handling and realistic outputs optimized for commercial and creative use.
image-to-imagefal-ai/object-removal/maskRemoves mask-selected objects and their visual effects, seamlessly reconstructing the scene with contextually appropriate content.
image-to-imagefal-ai/ideogram/v3/editTransform existing images with Ideogram V3's editing capabilities. Modify, adjust, and refine images while maintaining high fidelity and realistic outputs with precise prompt control.
image-to-videoblackforestlabs/flux-3/keyframes-to-videoFLUX 3 is Black Forest Labs' frontier video model. This endpoint builds video from a sequence of keyframes, generating the motion between each anchor point for precise control over how a shot progresses.
text-to-imageideogram/v4/fastGenerate high-quality images, posters, and logos with Ideogram's latest V4.0q — producing crisp visuals with accurate text rendering, fine detail, and full creative control for polished, ready-to-use designs IN A SECOND.
image-to-imagefal-ai/florence-2-large/object-detectionFlorence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
text-to-videofal-ai/bytedance/seedance/v1/pro/fast/text-to-videoText to Video endpoint for Seedance 1.0 Pro Fast, a next-generation video model designed to deliver maximum performance at minimal cost
image-to-imagebria/fibo-edit-1.5/virtual-try-onBria Virtual Try-On edits a person photo to show the subject wearing garments or accessories from one to three reference images, guided by optional text instructions. Built on FIBO-Edit-1.5, it supports multi-garment changes and preserves the source aspect ratio by default.
text-to-imagefal-ai/z-image/turbo/tilingGenerate seamlessly tiling photorealistic images from text using Z-Image Turbo
image-to-videofal-ai/pixverse/v5.5/image-to-videoGenerate high quality video clips from text and image prompts using PixVerse v5.5
video-to-videofal-ai/kling-video/o3/4k/video-to-video/editKling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling