image-to-imageImage Preprocessors
fal-ai/image-preprocessors/scribbleScribble preprocessor.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
image-to-imagefal-ai/image-preprocessors/scribbleScribble preprocessor.
image-to-imagefal-ai/docresEnhance low-resolution, blur, shadowed documents with the superior quality of docres for sharper, clearer results.
trainingfal-ai/ernie-image-trainerLoRA trainer for ERNIE-Image, Baidu's powerful 8B-parameter text-to-image model.
text-to-audiofal-ai/ltx-2.3-quality/text-to-audio/loraText to Audio high-quality using LTX-2.3 with Lora
video-to-videobria/video/erase/keypointsHigh-fidelity keypoint-driven video object removal - minimal input, strong temporal consistency. Trained on licensed data for risk-free commercial video editing.
audio-to-videofal-ai/ltx-2.3-22b/audio-to-video/loraGenerate video with audio from audio, text and images using LTX-2.3 and custom LoRA
visionfal-ai/sa2va/4b/videoSa2VA is an MLLM capable of question answering, visual prompt understanding, and dense object segmentation at both image and video levels
trainingfal-ai/ltx23-v2v-trainerTrain LTX-2.3 22B for video transformation or video-conditioned generation.
visionfal-ai/florence-2-large/region-to-categoryFlorence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
jsonfal-ai/omnilottie/image-to-lottieConvert your assets into lottie using Omnilottie.
audio-to-videofal-ai/ltx-2.3-quality/audio-to-video/loraGenerate high-quality video with audio from audio, text and images using LTX-2.3 and custom LoRA
audio-to-audiomirelo-ai/sfx1.6/inpaint-audioErase and replace any moment in your audio with AI-driven precision.
image-to-videofal-ai/ltx-2.3-quality/image-to-video/loraGenerate high-quality video with audio from images using LTX-2.3 and custom LoRA
visionperceptron/isaac-01/openai/v1/chat/completionsOpenAI spec compatible endpoint of Isaac-01 which is a multimodal vision-language model from Perceptron for various vision language tasks.
jsonfal-ai/omnilottie/video-to-lottieConvert your assets into lottie using Omnilottie.
video-to-videobria/bria_video_eraser/erase/keypointsA high-fidelity capability for erasing unwanted objects, people, or visual elements from videos while maintaining aesthetic quality and temporal consistency.
video-to-videobria/video/background-removal/realtimeRemove video backgrounds in real time with Bria’s VRMBG 3.0 model. Built for live streaming, real-time video apps, content creation, and low-latency workflows that need fast, accurate background removal.
trainingfal-ai/wan-trainer/t2v-14bTrain custom LoRAs for Wan-2.1 T2V 14B
speech-to-textfal-ai/speech-to-text/turbo/streamLeverage the rapid processing capabilities of AI models to enable accurate and efficient real-time speech-to-text transcription.
text-to-speechfal-ai/maya/streamMaya1 is a state-of-the-art speech model by Maya Research for expressive voice generation, built to capture real human emotion and precise voice design.
image-to-imagefal-ai/florence-2-large/region-to-segmentationFlorence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
trainingfal-ai/ltx23-trainer-v2/ic-lora/v2vTrain an IC-LoRA that learns a video-to-video transformation from paired before/after clips, conditioned at inference on a reference (control) video.
text-to-video
text-to-videofal-ai/infinity-star/text-to-videoInfinityStar’s unified 8B spacetime autoregressive engine to turn any text prompt into crisp 720p videos - 10× faster than diffusion models.
audio-to-audiofal-ai/workflow-utilities/impulse-responseFFMPEG Utility for Impulse Response
image-to-imagefal-ai/qwen-image-edit-2509-lora-gallery/face-to-full-portraitGenerate full portrait from a cropped face photo
image-to-imagefal-ai/docres/dewarpEnhance wraped, folded documents with the superior quality of docres for sharper, clearer results.
image-to-imagefal-ai/image-preprocessors/pidiPIDI (Pidinet) preprocessor.