text-to-imageErnie Image
fal-ai/ernie-imageHigh-quality text-to-image model by Baidu. Supports English, Chinese, and Japanese prompts with built-in prompt expansion.
Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.
text-to-imagefal-ai/ernie-imageHigh-quality text-to-image model by Baidu. Supports English, Chinese, and Japanese prompts with built-in prompt expansion.
text-to-imagerecraft/v4/style/text-to-vectorGenerates vector images that hold a consistent style, from either a saved style ID or reference images attached directly.
text-to-imagefal-ai/recraft/v4.1/utility/text-to-imageRecraft V4.1 Utility is a faster, lighter variant of V4.1 made for high-volume creative workflows. Ideal for ideation, A/B exploration, and content pipelines, it keeps Recraft's design sensibility while optimizing for throughput and cost.
text-to-imagerecraft/v4/style/pro/text-to-vectorGenerates vector images that hold a consistent style, from either a saved style ID or reference images attached directly.
text-to-imagenvidia/cosmos-3-super/text-to-imageCosmos3 is a collection of Omnimodal world models capable of generating dynamic, high-quality video, image, audio, and action commands from combinations of text, image, video, and action trajectory inputs.
text-to-imagefal-ai/flux-control-lora-cannyFLUX Control LoRA Canny is a high-performance endpoint that uses a control image to transfer structure to the generated image, using a Canny edge map.
text-to-imagefal-ai/flux-krea-loraSuper fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.
text-to-imagefal-ai/nucleus-imageNucleus-Image is a text-to-image generation model built on a sparse mixture-of-experts (MoE) diffusion transformer architecture.
text-to-imagefal-ai/sana/sprintSana Sprint is a text-to-image model capable of generating 4K images with exceptional speed.
text-to-imagefal-ai/qwen-image-max/text-to-imageText-to-Image endpoint for Qwen-Image-Max. Qwen Image Max improves upon the Qwen Image Plus series by enhancing the realism and naturalness of images.
text-to-imagefal-ai/flux-2/klein/4b/baseText-to-image generation with FLUX.2 [klein] 4B Base from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.
text-to-image
text-to-imagefal-ai/recraft/v4.1/utility/pro/text-to-imageRecraft V4.1 Utility Pro pairs the high-resolution output of V4.1 Pro with a faster, cost-efficient runtime. Designed for studios shipping large-format work at scale, it makes premium-quality raster generation viable across full creative pipelines.
text-to-image
text-to-imagefal-ai/fooocus/inpaintDefault parameters with automated optimizations and quality improvements.
text-to-imagefal-ai/playground-v25State-of-the-art open-source model in aesthetic quality
text-to-imagefal-ai/illusion-diffusionCreate illusions conditioned on image.
text-to-imagerundiffusion-fal/juggernaut-flux/lightningJuggernaut Lightning Flux by RunDiffusion provides blazing-fast, high-quality images rendered at five times the speed of Flux. Perfect for mood boards and mass ideation, this model excels in both realism and prompt adherence.
text-to-imagebria/fibo-gen-1.5/text-to-imageText-to-image model with high-fidelity outputs, accurate typography, and style preset, strong in photorealism, textures, and beyond. JSON-structured prompts give enterprise and agentic workflows production-ready control. Trained on licensed data.
text-to-imagefal-ai/hidream-o1-image/devUnified image generation with HiDream-O1-Image. Create, edit, and personalize high-resolution images up to 2K—single native model handles text-to-image, editing, and custom subjects without external components.
text-to-imagefal-ai/stable-diffusion-v35-mediumStable Diffusion 3.5 Medium is a Multimodal Diffusion Transformer (MMDiT) text-to-image model that features improved performance in image quality, typography, complex prompt understanding, and resource-efficiency.
text-to-image
text-to-image
text-to-imagebria/fibo/generateSOTA open-source text-to-image model delivering high-fidelity outputs with accurate typography. JSON-structured prompts provide production-ready controllability for enterprise and agentic workflows. Trained exclusively on licensed data.
text-to-imagefal-ai/glm-imageCreate high-quality images with accurate text rendering and rich knowledge details—supports editing, style transfer, and maintaining consistent characters across multiple images.
text-to-imagefal-ai/pony-v7Pony V7 is a finetuned text to image for superior aesthetics and prompt following.
text-to-imagefal-ai/hidream-i1-devHiDream-I1 dev is a new open-source image generative foundation model with 17B parameters that achieves state-of-the-art image generation quality within seconds.
text-to-imagefal-ai/aura-flowAuraFlow v0.3 is an open-source flow-based text-to-image generation model that achieves state-of-the-art results on GenEval. The model is currently in beta.