![FLUX.1 [schnell] is a 12 billion parameter flow transformer that generates high-quality images from text in 1 to 4 steps, suitable for personal and commercial use.](https://refinery.fal.media/url/https%3A%2F%2Fv3b.fal.media%2Ffiles%2Fb%2F0a9af64d%2FGxvCUPd3gO-MSYcy06g0x_1641cfe028c2429b8e12e4fc320eb0a8.jpg/tr:w-1920,q-80/GxvCUPd3gO-MSYcy06g0x_1641cfe028c2429b8e12e4fc320eb0a8.webp)
FLUX.1 [schnell] is a 12 billion parameter flow transformer that generates high-quality images from text in 1 to 4 steps, suitable for personal and commercial use.

Nano Banana Pro is Google's new state-of-the-art image generation and editing model

Nano Banana 2 is Google's new state-of-the-art fast image generation and editing model

GPT Image 2, OpenAI's latest image model, is capable of creating extremely detailed images with fine typography.

OpenAI's default image model for most applications. Fast, high-quality generation with natural lighting, rich textures, and support for complex layouts including transparent backgrounds.
![FLUX.1 [dev] is a 12 billion parameter flow transformer that generates high-quality images from text. It is suitable for personal and commercial use.](https://refinery.fal.media/url/https%3A%2F%2Fstorage.googleapis.com%2Ffal_cdn%2Ffal%2FUpscale-1.jpeg/tr:w-1920,q-80/Upscale-1.webp)
FLUX.1 [dev] is a 12 billion parameter flow transformer that generates high-quality images from text. It is suitable for personal and commercial use.

OpenAI's precision-focused image model, built for premium visual work, extra fidelity on intricate detail, in exchange for longer generation times.
![Image editing with FLUX.2 [pro] from Black Forest Labs. Ideal for high-quality image manipulation, style transfer, and sequential editing workflows](https://refinery.fal.media/url/https%3A%2F%2Fv3b.fal.media%2Ffiles%2Fb%2Fpenguin%2FUfryXXm9my6IM8HsoP9FL_054c2c2953dc491996904114c6e04836.jpg/tr:w-1920,q-80/UfryXXm9my6IM8HsoP9FL_054c2c2953dc491996904114c6e04836.webp)
Image editing with FLUX.2 [pro] from Black Forest Labs. Ideal for high-quality image manipulation, style transfer, and sequential editing workflows

Google's famous original image generation and editing model
FLUX1.1 [pro] is an enhanced version of FLUX.1 [pro], improved image generation capabilities, delivering superior composition, detail, and artistic fidelity compared to its predecessor.

ByteDance's Seedream 5.0 Pro is flagship text-to-image model, with deep-thinking prompt understanding, native text in 14 languages, and precise control over dense layouts and structured designs.
![FLUX1.1 [pro] ultra is the newest version of FLUX1.1 [pro], maintaining professional-grade image quality while delivering up to 2K resolution with improved photo realism.](https://refinery.fal.media/url/https%3A%2F%2Fstorage.googleapis.com%2Ffalserverless%2Fgallery%2Fflux-pro-v1-1-ultra.webp/tr:w-1920,q-80/flux-pro-v1-1-ultra.webp)
FLUX1.1 [pro] ultra is the newest version of FLUX1.1 [pro], maintaining professional-grade image quality while delivering up to 2K resolution with improved photo realism.

Z-Image Turbo is a super fast text-to-image model of 6B parameters developed by Tongyi-MAI.

A new-generation image creation model ByteDance, Seedream 4.5 integrates image generation and image editing capabilities into a single, unified architecture.
![Text-to-image generation with FLUX.2 [klein] 9B from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.](https://refinery.fal.media/url/https%3A%2F%2Fv3b.fal.media%2Ffiles%2Fb%2F0a8a7f3c%2F90FKDpwtSCZTqOu0jUI-V_64c1a6ec0f9343908d9efa61b7f2444b.jpg/tr:w-1920,q-80/90FKDpwtSCZTqOu0jUI-V_64c1a6ec0f9343908d9efa61b7f2444b.webp)
Text-to-image generation with FLUX.2 [klein] 9B from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.
![Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.](https://refinery.fal.media/url/https%3A%2F%2Fv3b.fal.media%2Ffiles%2Fb%2F0a9f91af%2FUMrx6t6mc59z33d20WIQP_iBIZYKKq.png/tr:w-1920,q-80/UMrx6t6mc59z33d20WIQP_iBIZYKKq.webp)
Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

A new-generation image creation model ByteDance, Seedream 4.0 integrates image generation and image editing capabilities into a single, unified architecture.
![Text-to-image generation with FLUX.2 [dev] from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.](https://refinery.fal.media/url/https%3A%2F%2Fv3b.fal.media%2Ffiles%2Fb%2Fpenguin%2FzSBCJtPpeIQwR5AC_IamX_b1e1137961754e4d851907c21f8c20cd.jpg/tr:w-1920,q-80/zSBCJtPpeIQwR5AC_IamX_b1e1137961754e4d851907c21f8c20cd.webp)
Text-to-image generation with FLUX.2 [dev] from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.

FLUX 3 Image is Black Forest Labs' newest image model. Generate detailed, well-composed images in native 2K and 4K with precise layout control and improved text rendering.

Nano banana lite is the efficiency-focused model in the image generation family. Sub-2 second latency with cost-effective generation and editing, fast multi-turn local edits, and 14 supported aspect ratios.

Generate highly aesthetic images with xAI's Grok Imagine Image generation model.

Generate high-quality images, posters, and logos with Ideogram V3. Features exceptional typography handling and realistic outputs optimized for commercial and creative use.

Meta's Muse Image model has faithful instruction-following and exceptional visual fidelity, with fine details like text, plots, and QR codes rendered accurately.

Generate high-fidelity images from text with Krea 2 Large, supporting aspect ratio, creativity, seed controls, and optional style references.
Text-to-image endpoints on fal return an array of image URLs. The call shape stays the same across all models on this page.
jsimport { fal } from "@fal-ai/client"; const result = await fal.subscribe("fal-ai/nano-banana-2", { input: { prompt: "A photorealistic Tokyo cafe at golden hour" } }); console.log(result.data.images[0].url);
You install @fal-ai/client, export FAL_KEY, and call any endpoint by its model string. Swapping to openai/gpt-image-2 or bytedance/seedream/v4.5/text-to-image is then only a one-line change.
In-image typography is a real differentiator for a few models on fal.
Use these models when you need readable text inside the image itself, such as posters, ads, product mockups, UI screens, packaging, logos, or multilingual creative assets.
The FLUX 1.x and FLUX 2 families anchor most photoreal workflows on fal.
Pick these models when realism, lighting, surface detail, camera feel, and material accuracy matter more than raw generation speed.
Several models on fal are tuned for fast, low-cost text-to-image generation.
A common workflow is to use faster models for prompt exploration and high-volume drafts, then send only the best candidates to higher-fidelity models for final output.
Pricing on text-to-image models is per output, with rates set by the model and resolution. Some models charge per megapixel, while others charge per image based on quality or size.
| Model | Price |
|---|---|
| FLUX.1 [schnell] | $0.003 / megapixel |
| Nano Banana 2 | $0.08 / image at 1K |
| Nano Banana Pro | $0.15 / image |
Nano Banana 2 supports higher-resolution pricing, with 2K and 4K outputs scaling at 1.5x and 2x the 1K rate.
You should also include add-ons in your calculations. Nano Banana 2 charges an extra $0.015 per generation for web search grounding and $0.002 for high-tier thinking, while GPT Image 2 uses quality-tier billing across low, medium, and high settings.