audio-to-audioACE Step Audio Inpaint
fal-ai/ace-step/audio-inpaintModify a portion of provided audio with lyrics and/or style using ACE-Step
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
audio-to-audiofal-ai/ace-step/audio-inpaintModify a portion of provided audio with lyrics and/or style using ACE-Step
video-to-videofal-ai/wan/v2.2-a14b/video-to-videoWan-2.2 video-to-video is a video model that generates high-quality videos with high visual quality and motion diversity from text prompts and source videos.
text-to-imagefal-ai/luma-photon/flashGenerate images from your prompts using Luma Photon Flash. Photon Flash is the most creative, personalizable, and intelligent visual models for creatives, bringing a step-function change in the cost of high-quality image generation.
image-to-videofal-ai/hunyuan-video-v1.5/image-to-videoHunyuan Video 1.5 is Tencent's latest and best video model
image-to-imagefal-ai/nafnet/deblurUse NAFNet to fix issues like blurriness and noise in your images. This model specializes in image restoration and can help enhance the overall quality of your photography.
image-to-videofal-ai/pixverse/v4.5/transitionCreate seamless transition between images using PixVerse v4.5
audio-to-videofal-ai/ltx-2.3-quality/audio-to-videoGenerate high-quality video with audio from audio, text and images using LTX-2.3
video-to-videofal-ai/scail-2SCAIL-2 is an end-to-end character animation model that drives a reference character from a source video without relying on intermediate pose representations like skeleton maps.
speech-to-textfal-ai/speech-to-textLeverage the rapid processing capabilities of AI models to enable accurate and efficient real-time speech-to-text transcription.
text-to-imagerundiffusion-fal/juggernaut-flux/proJuggernaut Pro Flux by RunDiffusion is the flagship Juggernaut model rivaling some of the most advanced image models available, often surpassing them in realism. It combines Juggernaut Base with RunDiffusion Photo and features enhancements like reduced background blurriness.
text-to-imagefal-ai/sana/v1.5/4.8bSana v1.5 4.8B is a powerful text-to-image model that generates ultra-high quality 4K images with remarkable detail.
image-to-imagefal-ai/moondream-next/detectionMoonDreamNext Detection is a multimodal vision-language model for gaze detection, bbox detection, point detection, and more.
text-to-imagerundiffusion-fal/juggernaut-flux-loraJuggernaut Base Flux LoRA by RunDiffusion is a drop-in replacement for Flux [Dev] that delivers sharper details, richer colors, and enhanced realism to all your LoRAs and LyCORIS with full compatibility.
text-to-3dfal-ai/hunyuan3d-v3/text-to-3dTurn simple sketches into detailed, fully-textured 3D models. Instantly convert your concept designs into formats ready for Unity, Unreal, and Blender.
fal-ai/ltx-2.3-22b/text-to-videoGenerate video with audio from text using LTX-2.3
video-to-videotopaz/sdr-to-hdr/videoProfessional SDR-to-HDR conversion powered by Topaz Labs. Hyperion 2.5 redistributes luminance and color while preserving detail in text, faces and motion. Best for giving flat SDR footage a true HDR look.
text-to-videofal-ai/kling-video/v1.6/standard/effectsGenerate video clips from your prompts using Kling 1.6 (std)
image-to-imagefal-ai/luma-photon/reframeExtend and reframe images with Luma Photon Reframe. This advanced tool intelligently expands your visuals, seamlessly blending new content to enhance creativity and adaptability, offering unmatched personalization and quality for creators at a fraction of the cost.
visionfal-ai/got-ocr/v2GOT-OCR2 works on a wide range of tasks, including plain document OCR, scene text OCR, formatted document OCR, and even OCR for tables, charts, mathematical formulas, geometric shapes, molecular formulas and sheet music.
3d-to-3dtripo3d/tripo/remeshConverts triangle meshes into clean quad topology at a target polygon count, animation-ready with no manual retopology.
image-to-imagefal-ai/flux-krea-lora/inpaintingSuper fast endpoint for the FLUX.1 [dev] inpainting model with LoRA support, enabling rapid and high-quality image inpaingting using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.
image-to-imagefal-ai/image-apps-v2/makeup-applicationApply realistic makeup styles with adjustable intensity.
image-to-videofal-ai/fast-svd-lcmGenerate short video clips from your images using SVD v1.1 at Lightning Speed
image-to-imagefal-ai/post-processing/sharpenApply sharpening effects with three modes: basic unsharp mask, smart sharpening with edge preservation, and Contrast Adaptive Sharpening (CAS).
video-to-videofal-ai/workflow-utilities/reverse-videoFFMPEG Utility to Reverse Videos
image-to-imagefal-ai/telestyle-v2Restyle any image with TeleStyle v2 — provide an original image and a styling reference, and the model re-renders the original in the reference's visual style while preserving its content and composition.
text-to-imagefal-ai/hunyuan-image/v2.1/text-to-imageUse the amazing capabilities of hunyuan image 2.1 to generate images that express the feelings of your text.
video-to-videofal-ai/pixverse/sound-effectsAdd immersive sound effects and background music to your videos using PixVerse sound effects generation