text-to-imageRecraft V4 Styles Text to Image
recraft/v4/style/text-to-imageGenerates raster images that hold a consistent style, from either a saved style ID or reference images attached directly.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
text-to-imagerecraft/v4/style/text-to-imageGenerates raster images that hold a consistent style, from either a saved style ID or reference images attached directly.
text-to-speechfal-ai/minimax/speech-2.6-turboGenerate speech from text prompts and different voices using the MiniMax Speech-2.6 HD model, which leverages advanced AI techniques to create high-quality text-to-speech.
image-to-imagebria/fibo-edit/editHigh-fidelity image editing model with state-of-the-art controllability. Combines JSON + Mask + Image for precise, fine-grained edits ideal for production and enterprise workflows. Trained on licensed data - safe for commercial use.
video-to-videofal-ai/wan-motionWan Motion is a streamlined character animation model that transfers motion from a driving video onto a reference character image. Based on Wan-Animate which preserves the original character's proportions, Simple uses pose retargeting to adapt the driving video's skeleton to match the reference character's body shape, producing more natural results when the two have different builds. It outputs at 720p with optimized defaults for fast, high-quality generation — just provide a video, an image, and an optional prompt.
text-to-imagefal-ai/wan/v2.7/pro/text-to-imageGenerate premium-quality images from text prompts using the enhanced WAN 2.7 Pro model with superior detail and composition.
image-to-imagefal-ai/image-editing/text-removalRemove all text and writing from images while preserving the background and natural appearance.
trainingfal-ai/z-image-turbo-trainer-v2Fast LoRA trainer for Z-Image-Turbo, a super fast text-to-image model of 6B parameters developed by Tongyi-MAI.
image-to-imagefal-ai/image-editing/reframeThe reframe endpoint intelligently adjusts an image's aspect ratio while preserving the main subject's position, composition, pose, and perspective
text-to-video
video-to-videofal-ai/sam2/videoSAM 2 is a model for segmenting images and videos in real-time.
text-to-imagefal-ai/flux-2/klein/9b/baseText-to-image generation with FLUX.2 [klein] 9B Base from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.
video-to-videofal-ai/film/videoInterpolate videos with FILM - Frame Interpolation for Large Motion
image-to-imagefal-ai/ideogram/character/editModify consistent characters while preserving their core identity. Edit poses, expressions, or clothing without losing recognizable character features
text-to-3dfal-ai/meshy/v6/text-to-3dMeshy-6 is the latest model from Meshy. It generates realistic and production ready 3D models.
text-to-videofal-ai/wan/v2.2-5b/text-to-videoWan 2.2's 5B model produces up to 5 seconds of video 720p at 24FPS with fluid motion and powerful prompt understanding
image-to-imagefal-ai/instant-characterInstantCharacter creates high-quality, consistent characters from text prompts, supporting diverse poses, styles, and appearances with strong identity control.
image-to-3dtripo3d/p1/image-to-3dGenerate 3D models from a single image using Tripo P1.
text-to-imagefal-ai/hidream-o1-imageUnified image generation with HiDream-O1-Image. Create, edit, and personalize high-resolution images up to 2K—single native model handles text-to-image, editing, and custom subjects without external components.
image-to-videofal-ai/pixverse/v4.5/image-to-videoGenerate high quality video clips from text and image prompts using PixVerse v4.5
image-to-3dfal-ai/pixal3dPixal3D turns a single image into a high-fidelity 3D model with detailed geometry and realistic textures.
image-to-videofal-ai/pixverse/c1/reference-to-videoGenerate character-consistent videos from reference images using PixVerse C1, with subject and background references.
image-to-imagefal-ai/qwen-image-edit-plus-lora-gallery/multiple-anglesPrecise camera position and angle control (rotation, zoom, vertical movement)
image-to-videofal-ai/sadtalkerLearning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation
video-to-videofal-ai/hunyuan-video-foleyUse the capabilities of the hunyuan foley model to bring life to your videos by adding sound effect to them.
text-to-imagefal-ai/flux-control-lora-depthFLUX Control LoRA Depth is a high-performance endpoint that uses a control image to transfer structure to the generated image, using a depth map.
image-to-videofal-ai/minimax/video-01/image-to-videoGenerate video clips from your images using MiniMax Video model
image-to-videofal-ai/live-portraitTransfer expression from a video to a portrait.
text-to-videolightricks/ltx-2.5/text-to-video/proLTX-2.5 is Lightricks' open-source audio-video model. This endpoint generates synchronized video and audio from a text prompt in a single pass, in a quality-optimized mode for final, high-fidelity output.