EXPLORE / MODEL DIRECTORY

Explore models.Know what runs.

Browse the public fal model directory. Models become callable through Kaista Cloud only after supply, pricing, and compatibility review.

3Kaista-ready endpointsPublic catalog connected
PUBLIC REFERENCE CATALOG

text to image

Page 5 · 28 per page
text-to-image
falREVIEW REQUIRED

Ernie Image

fal-ai/ernie-image

High-quality text-to-image model by Baidu. Supports English, Chinese, and Japanese prompts with built-in prompt expansion.

realismchinesemultilingualportrait
text-to-image
RecraftREVIEW REQUIRED

Recraft V4 Styles Text to Vector

recraft/v4/style/text-to-vector

Generates vector images that hold a consistent style, from either a saved style ID or reference images attached directly.

stylizedtransformediting
text-to-image
RecraftREVIEW REQUIRED

Recraft V4.1 Text to Image Utility

fal-ai/recraft/v4.1/utility/text-to-image

Recraft V4.1 Utility is a faster, lighter variant of V4.1 made for high-volume creative workflows. Ideal for ideation, A/B exploration, and content pipelines, it keeps Recraft's design sensibility while optimizing for throughput and cost.

stylizedtransformtypography
text-to-image
RecraftREVIEW REQUIRED

Recraft V4 Styles Pro Text to Vector

recraft/v4/style/pro/text-to-vector

Generates vector images that hold a consistent style, from either a saved style ID or reference images attached directly.

stylizedtransformediting
text-to-image
nvidiaREVIEW REQUIRED

Cosmos 3 Super

nvidia/cosmos-3-super/text-to-image

Cosmos3 is a collection of Omnimodal world models capable of generating dynamic, high-quality video, image, audio, and action commands from combinations of text, image, video, and action trajectory inputs.

stylizedtransformrealism
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 [dev] Control LoRA Canny

fal-ai/flux-control-lora-canny

FLUX Control LoRA Canny is a high-performance endpoint that uses a control image to transfer structure to the generated image, using a Canny edge map.

lorastyle transfer
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.1 Krea [dev] with LoRAs

fal-ai/flux-krea-lora

Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

lorapersonalization
text-to-image
falREVIEW REQUIRED

Nucleus Image

fal-ai/nucleus-image

Nucleus-Image is a text-to-image generation model built on a sparse mixture-of-experts (MoE) diffusion transformer architecture.

stylizedtransformtypography
text-to-image
falREVIEW REQUIRED

Sana Sprint

fal-ai/sana/sprint

Sana Sprint is a text-to-image model capable of generating 4K images with exceptional speed.

text to image4khigh-speed
text-to-image
AlibabaREVIEW REQUIRED

Qwen Image Max

fal-ai/qwen-image-max/text-to-image

Text-to-Image endpoint for Qwen-Image-Max. Qwen Image Max improves upon the Qwen Image Plus series by enhancing the realism and naturalness of images.

qwen-imagemax
text-to-image
Black Forest LabsREVIEW REQUIRED

FLUX.2 [klein] 4B Base

fal-ai/flux-2/klein/4b/base

Text-to-image generation with FLUX.2 [klein] 4B Base from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.

text-to-image
RecraftREVIEW REQUIRED

Recraft V4.1 Utility Text to Image

fal-ai/recraft/v4.1/utility/pro/text-to-image

Recraft V4.1 Utility Pro pairs the high-resolution output of V4.1 Pro with a faster, cost-efficient runtime. Designed for studios shipping large-format work at scale, it makes premium-quality raster generation viable across full creative pipelines.

stylizedtransformtypography
text-to-image
falREVIEW REQUIRED

Realistic Vision

fal-ai/realistic-vision

Generate realistic images.

realismdiffusion
text-to-image
falREVIEW REQUIRED

Fooocus Inpainting

fal-ai/fooocus/inpaint

Default parameters with automated optimizations and quality improvements.

stylizedediting
text-to-image
falREVIEW REQUIRED

Playground v2.5

fal-ai/playground-v25

State-of-the-art open-source model in aesthetic quality

artisticstyle
text-to-image
falREVIEW REQUIRED

Illusion Diffusion

fal-ai/illusion-diffusion

Create illusions conditioned on image.

compositionstylized
text-to-image
Black Forest LabsREVIEW REQUIRED

Juggernaut Flux Lightning

rundiffusion-fal/juggernaut-flux/lightning

Juggernaut Lightning Flux by RunDiffusion provides blazing-fast, high-quality images rendered at five times the speed of Flux. Perfect for mood boards and mass ideation, this model excels in both realism and prompt adherence.

image generation
text-to-image
briaREVIEW REQUIRED

Fibo Gen 1.5 Text to Image

bria/fibo-gen-1.5/text-to-image

Text-to-image model with high-fidelity outputs, accurate typography, and style preset, strong in photorealism, textures, and beyond. JSON-structured prompts give enterprise and agentic workflows production-ready control. Trained on licensed data.

stylizedtransformrealism
text-to-image
falREVIEW REQUIRED

Hidream O1 Image

fal-ai/hidream-o1-image/dev

Unified image generation with HiDream-O1-Image. Create, edit, and personalize high-resolution images up to 2K—single native model handles text-to-image, editing, and custom subjects without external components.

text-to-image
falREVIEW REQUIRED

Stable Diffusion 3.5 Medium

fal-ai/stable-diffusion-v35-medium

Stable Diffusion 3.5 Medium is a Multimodal Diffusion Transformer (MMDiT) text-to-image model that features improved performance in image quality, typography, complex prompt understanding, and resource-efficiency.

diffusiontypographystyle
text-to-image
briaREVIEW REQUIRED

Fibo

bria/fibo/generate

SOTA open-source text-to-image model delivering high-fidelity outputs with accurate typography. JSON-structured prompts provide production-ready controllability for enterprise and agentic workflows. Trained exclusively on licensed data.

briafiboprompt-adherence
text-to-image
falREVIEW REQUIRED

GLM Image

fal-ai/glm-image

Create high-quality images with accurate text rendering and rich knowledge details—supports editing, style transfer, and maintaining consistent characters across multiple images.

text-to-image
text-to-image
falREVIEW REQUIRED

Pony V7

fal-ai/pony-v7

Pony V7 is a finetuned text to image for superior aesthetics and prompt following.

diffusionstyle
text-to-image
falREVIEW REQUIRED

Hidream I1 Dev

fal-ai/hidream-i1-dev

HiDream-I1 dev is a new open-source image generative foundation model with 17B parameters that achieves state-of-the-art image generation quality within seconds.

text-to-image
falREVIEW REQUIRED

AuraFlow

fal-ai/aura-flow

AuraFlow v0.3 is an open-source flow-based text-to-image generation model that achieves state-of-the-art results on GenEval. The model is currently in beta.

typographystyle