trainingTrain Flux LoRAs For Portraits
fal-ai/flux-lora-portrait-trainerFLUX LoRA training optimized for portrait generation, with bright highlights, excellent prompt following and highly detailed results.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
trainingfal-ai/flux-lora-portrait-trainerFLUX LoRA training optimized for portrait generation, with bright highlights, excellent prompt following and highly detailed results.
text-to-videofal-ai/wan-25-preview/text-to-videoWan 2.5 text-to-video model.
image-to-videofal-ai/wan-i2vWan-2.1 is a image-to-video model that generates high-quality videos with high visual quality and motion diversity from images
visionfal-ai/moondream3-preview/detectMoondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.
image-to-imagefal-ai/hunyuan-image/v3/instruct/editImage editing endpoint for Hunyuan Image 3.0 Instruct.
text-to-imagefal-ai/kling-image/o3/text-to-imageKling Omni 3: Top-tier text-to-image with flawless consistency.
image-to-imagebytedance/seedream/v5/flash/layerizeSeedream 5.0 Flash is a fast image generation and editing model, built for workflows where speed and budget matter.
image-to-videofal-ai/pixverse/v6/transitionPixverse's latest v6 Model.
image-to-videofal-ai/ffmpeg-api/images-to-videoA fal.ai endpoint that stitches an ordered list of images into an MP4 video by holding each image for a specified number of frames at a configurable frame rate
text-to-3dtripo3d/h3.1/text-to-3dGenerate 3D models from text descriptions using Tripo H3.1.
text-to-speechfal-ai/index-tts-2/text-to-speechGenerate natural, clear speeches using Index TTS 2.0 from IndexTeam
image-to-imagefal-ai/ideogram/characterGenerate consistent character appearances across multiple images. Maintain facial features, proportions, and distinctive traits for cohesive storytelling and branding
image-to-imagefal-ai/qwen-image-edit-2511/loraEndpoint for Qwen's Image Editing 2511 model with LoRa support.
image-to-3dfal-ai/sam-3/3d-bodySAM 3D allows for accurate 3D reconstruction of human body shape and position from a single image.
text-to-audiofal-ai/stable-audio-3/small/sfx/text-to-audioStable Audio 3 Small SFX is a 459 million parameter latent diffusion model that generates high-quality sound effects from text prompts, designed for on-device deployment on mobile phones and consumer laptops.
text-to-imagefal-ai/fast-lightning-sdxlRun SDXL at the speed of light
visionfal-ai/moondream3-preview/queryMoondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.
audio-to-audiofal-ai/kling-video/create-voiceCreate Voices to be used with Kling Models Voice Control
text-to-imagefal-ai/recraft/v4.1/pro/text-to-imageRecraft V4.1 Pro pushes the V4.1 model into high-resolution territory — up to 2048×2048 and ultra-wide formats. Made for hero imagery, campaign work, and print, it preserves the same design taste at sizes ready for the final deliverable.
text-to-imagefal-ai/recraft/v4/text-to-vectorRecraft V4 was developed with designers to bring true visual taste to AI image generation. Built for brand systems and production-ready workflows, it goes beyond prompt accuracy — delivering stronger composition, refined lighting, realistic materials, and a cohesive aesthetic. The result is imagery shaped by professional design judgment, ready for immediate real-world use without additional post-processing.
image-to-videofal-ai/kling-video/o3/4k/reference-to-videoKling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling
image-to-videopixelcut/looping-videoTurn one product photo into a seamless 5 to 15 second video loop with a locked camera, subtle ambient motion or a full 360° spin.
text-to-videofal-ai/minimax/hailuo-02/standard/text-to-videoMiniMax Hailuo-02 Text To Video API (Standard, 768p): Advanced video generation model with 768p resolution
image-to-imagefal-ai/qwen-image-edit/inpaintInpainting Endpoint for the Qwen Edit Image editing model.
image-to-imagetopaz/upscale/image/transparentProfessional transparent-image upscaling powered by Topaz Labs. Preserves the alpha channel end to end with PNG output. Best for logos, stickers and assets with transparency.
video-to-audiosonilo/v1.1/video-to-sound-effectsAnalyzes a video and generates synchronized, royalty-free sound effects timed to visible actions. Returns the generated sound-effects audio track for commercial use.
text-to-videoblackforestlabs/flux-3/text-to-video/draftFLUX.3 is Black Forest Labs' frontier audio/video model. Generate fast, low-cost draft previews from a text prompt, with a reusable draft cache for full-quality enhancement.
image-to-imagefal-ai/kling-image/v3/image-to-imageKling Image V3: Latest kling image model