
Stable Cascade: Image generation on a smaller & cheaper latent space.

Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Default parameters with automated optimizations and quality improvements.

Bria's Text-to-Image model for HD images. Trained exclusively on licensed data for safe and risk-free commercial use. Available also as source code and weights. For access to weights: https://bria.ai/contact-us

Train Ideogram on your photos, your style, your subject, your look, from a small set of reference images to images that feel consistently yours

Fooocus extreme speed mode as a standalone app.

Anime finetune of Würstchen V3.
![Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.](https://refinery.fal.media/url/https%3A%2F%2Fv3b.fal.media%2Ffiles%2Fb%2F0a9f9a61%2FT4z71gOSeWv0wALDdi2-b_qVoN8eec.png/tr:w-1920,q-80/T4z71gOSeWv0wALDdi2-b_qVoN8eec.webp)
Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.
![Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.](https://refinery.fal.media/url/https%3A%2F%2Fstorage.googleapis.com%2Ffal_cdn%2Ffal%2FUpscale-5.jpeg/tr:w-1920,q-80/Upscale-5.webp)
Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

Qwen-Image-2.0 is a next-generation foundational unified generation-and-editing model

Qwen-Image-2.0 is a next-generation foundational unified generation-and-editing model

Use the faster speed of piflow to generate images with same quality to that of slower models.

Dreamina showcases superior picture effects, with significant improvements in picture aesthetics, precise and diverse styles, and rich details.

Google’s highest quality image generation model

Seedream 3.0 is a bilingual (Chinese and English) text-to-image model that excels at text-to-image generation.

Google’s highest quality image generation model

Google’s highest quality image generation model

DreamO is an image customization framework designed to support a wide range of tasks while facilitating seamless integration of multiple conditions.

F Lite is a 10B parameter diffusion model created by Fal and Freepik, trained exclusively on copyright-safe and SFW content.

F Lite is a 10B parameter diffusion model created by Fal and Freepik, trained exclusively on copyright-safe and SFW content. This is a high texture density variant of the model.

Imagen3 is a high-quality text-to-image model that generates realistic images from text prompts.

Imagen3 Fast is a high-quality text-to-image model that generates realistic images from text prompts.

Switti is a scale-wise transformer for fast text-to-image generation that outperforms existing T2I AR models and competes with state-of-the-art T2I diffusion models while being faster than distilled diffusion models.

Switti is a scale-wise transformer for fast text-to-image generation that outperforms existing T2I AR models and competes with state-of-the-art T2I diffusion models while being faster than distilled diffusion models.
Every image model on fal shares the same SDK pattern. Swapping Nano Banana 2 for FLUX.2 [dev] or GPT Image 2 is a one-line endpoint change, with the auth flow and queue behavior unchanged.
bashnpm install --save @fal-ai/client
bashexport FAL_KEY="YOUR_API_KEY"
jsimport { fal } from "@fal-ai/client"; const result = await fal.subscribe("fal-ai/nano-banana-2", { input: { prompt: "A photorealistic Tokyo cafe at golden hour" } });
GPT Image 2, Nano Banana 2, Ideogram V3, and Recraft V3 and V4 all treat typography as a primary capability.
The FLUX family covers most photoreal production work on fal, with a few partner models filling specific niches.
For high-volume work on fal, the Turbo and distilled models generate in roughly 1,2 seconds.
You can generate low-resolution drafts on a Turbo model to scout the prompt space, then run only the keepers through a higher-fidelity model at full resolution.
Image generation on fal is priced per output. Some models charge per megapixel of image area, others charge per image.
| Model | Price |
|---|---|
| Seedream V4.5 | $0.04 / image |
| Nano Banana 2 | $0.08 / image at 1K (1.5x at 2K, 2x at 4K) |
| Nano Banana Pro | $0.15 / image |
As a worked example, a team generating 1,000 marketing images per month at 1K resolution pays roughly: