video-to-videoHeygen Lipsync - Precision
fal-ai/heygen/v3/lipsync/precisionReplace or dub audio on an existing video with high-accuracy avatar-inference lip-sync.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
video-to-videofal-ai/heygen/v3/lipsync/precisionReplace or dub audio on an existing video with high-accuracy avatar-inference lip-sync.
video-to-videofal-ai/kling-video/o3/4k/video-to-video/referenceKling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling
video-to-videofal-ai/kling-video/o3/4k/video-to-video/editKling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling
fal-ai/video-upscalerThe video upscaler endpoint uses RealESRGAN on each frame of the input video to upscale the video to a higher resolution.
video-to-videotopaz/upscale/video/creativeProfessional creative video upscaling powered by Topaz Labs. Astra 2 reimagines fine detail and typically delivers 4K output. Best for cinematic shots that need maximum visual impact.
video-to-videofal-ai/void-video-inpaintingVOID removes objects from videos along with all interactions they induce on the scene
video-to-videobria/video/background-removalAutomatically remove backgrounds from videos -perfect for creating clean, professional content without a green screen.
video-to-videofal-ai/flashvsr/upscale/videoUpscale your videos using FlashVSR with the fastest speeds!
video-to-videofal-ai/kling-video/o1/standard/video-to-video/editEdit an existing video using natural-language instructions, transforming subjects, settings, and style while retaining the original motion structure.
video-to-videofal-ai/infinitalkInfinitalk model generates a talking avatar video from an image and audio file. The avatar lip-syncs to the provided audio with natural facial expressions.
video-to-videoblackforestlabs/flux-3/extend-videoFLUX 3 is Black Forest Labs' frontier video model. This endpoint continues an existing clip beyond its final frame, generating additional footage that stays consistent with the original motion and scene.
video-to-videofal-ai/wan-vace-14b/inpaintingVACE is a video generation model that uses a source image, mask, and video to create prompted videos with controllable sources.
video-to-videofal-ai/sam-3/video-rleSAM 3 is a unified foundation model for promptable segmentation in images and videos. It can detect, segment, and track objects using text or visual prompts such as points, boxes, and masks.
video-to-videofal-ai/sam-3-1/video-rleSAM 3.1 builds comes with Object Multiplex, a shared-memory approach for joint multi-object tracking that delivers faster speeds with larger number of objects tracked.
video-to-videofal-ai/veo3.1/fast/extend-videoExtend Veo-Created Videos up to 30 seconds
video-to-videopixelcut/video-background-removalPixelcut's Video Background Remover is an AI segmentation model that erases backgrounds frame by frame, with seamless temporal consistency.
video-to-videofal-ai/kling-video/o1/video-to-video/referenceKling O1 Omni generates new shots guided by an input reference video, preserving cinematic language such as motion, and camera style to produce seamless scene continuity.
video-to-videofal-ai/sam-3-1/videoSAM 3.1 builds comes with Object Multiplex, a shared-memory approach for joint multi-object tracking that delivers faster speeds with larger number of objects tracked.
video-to-videoxai/grok-imagine-video/extend-videoExtend videos with xAI's Grok Imagine video model
video-to-videofal-ai/ben/v2/videoA model for high quality and smooth background removal for videos.
video-to-video
video-to-videofal-ai/hunyuan-video-foleyUse the capabilities of the hunyuan foley model to bring life to your videos by adding sound effect to them.
video-to-videobria/video/erase/maskHigh-fidelity mask-based video object removal with strong temporal consistency. Erase unwanted objects, people, or elements while preserving aesthetic quality. Trained on licensed data for risk-free commercial use.
video-to-videomirelo-ai/sfx-v1.5/video-to-videoGenerate synced sounds for any video, and return it with its new sound track (like MMAudio)
video-to-videominimax/h3/reference-to-video/loraReferences into video with synchronized audio using MiniMax H3
video-to-videomirelo-ai/sfx1.6/video-to-videoGenerate synced sounds for any video, and return it with its new sound track (like MMAudio). Now up to 60 seconds!
video-to-videotopaz/interpolate/videoProfessional frame interpolation powered by Topaz Labs. Apollo, Chronos and Aion retime footage up to 120 fps, from smooth motion to extreme slow motion. Best for fluid 60fps output and slow-motion effects.
video-to-videofal-ai/wan-motionWan Motion is a streamlined character animation model that transfers motion from a driving video onto a reference character image. Based on Wan-Animate which preserves the original character's proportions, Simple uses pose retargeting to adapt the driving video's skeleton to match the reference character's body shape, producing more natural results when the two have different builds. It outputs at 720p with optimized defaults for fast, high-quality generation — just provide a video, an image, and an optional prompt.