
Ballpoint pen sketch drawing style

HDR surrealistic effect with intense colors
![Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.](https://refinery.fal.media/url/https%3A%2F%2Fv3b.fal.media%2Ffiles%2Fb%2F0a9f9a61%2FT4z71gOSeWv0wALDdi2-b_qVoN8eec.png/tr:w-1920,q-80/T4z71gOSeWv0wALDdi2-b_qVoN8eec.webp)
Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

Default parameters with automated optimizations and quality improvements.
![Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.](https://refinery.fal.media/url/https%3A%2F%2Fstorage.googleapis.com%2Ffal_cdn%2Ffal%2FUpscale-5.jpeg/tr:w-1920,q-80/Upscale-5.webp)
Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

Train Ideogram on your photos, your style, your subject, your look, from a small set of reference images to images that feel consistently yours

Fooocus extreme speed mode as a standalone app.

Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Stable Cascade: Image generation on a smaller & cheaper latent space.

Anime finetune of Würstchen V3.

Nano Banana 2.1 by Google generates images from text prompts, with output resolutions up to 4K, adjustable aspect ratios, and optional web search grounding. This API is for integration purposes only and cannot be used.

Qwen-Image-2.0 is a next-generation foundational unified generation-and-editing model

Qwen-Image-2.0 is a next-generation foundational unified generation-and-editing model

Use the faster speed of piflow to generate images with same quality to that of slower models.

Dreamina showcases superior picture effects, with significant improvements in picture aesthetics, precise and diverse styles, and rich details.

Google’s highest quality image generation model

Seedream 3.0 is a bilingual (Chinese and English) text-to-image model that excels at text-to-image generation.

Google’s highest quality image generation model

Google’s highest quality image generation model

DreamO is an image customization framework designed to support a wide range of tasks while facilitating seamless integration of multiple conditions.

F Lite is a 10B parameter diffusion model created by Fal and Freepik, trained exclusively on copyright-safe and SFW content.

F Lite is a 10B parameter diffusion model created by Fal and Freepik, trained exclusively on copyright-safe and SFW content. This is a high texture density variant of the model.

Imagen3 is a high-quality text-to-image model that generates realistic images from text prompts.

Imagen3 Fast is a high-quality text-to-image model that generates realistic images from text prompts.
Text-to-image endpoints on fal return an array of image URLs. The call shape stays the same across all models on this page.
jsimport { fal } from "@fal-ai/client"; const result = await fal.subscribe("fal-ai/nano-banana-2", { input: { prompt: "A photorealistic Tokyo cafe at golden hour" } }); console.log(result.data.images[0].url);
You install @fal-ai/client, export FAL_KEY, and call any endpoint by its model string. Swapping to openai/gpt-image-2 or bytedance/seedream/v4.5/text-to-image is then only a one-line change.
In-image typography is a real differentiator for a few models on fal.
Use these models when you need readable text inside the image itself, such as posters, ads, product mockups, UI screens, packaging, logos, or multilingual creative assets.
The FLUX 1.x and FLUX 2 families anchor most photoreal workflows on fal.
Pick these models when realism, lighting, surface detail, camera feel, and material accuracy matter more than raw generation speed.
Several models on fal are tuned for fast, low-cost text-to-image generation.
A common workflow is to use faster models for prompt exploration and high-volume drafts, then send only the best candidates to higher-fidelity models for final output.
Pricing on text-to-image models is per output, with rates set by the model and resolution. Some models charge per megapixel, while others charge per image based on quality or size.
| Model | Price |
|---|---|
| FLUX.1 [schnell] | $0.003 / megapixel |
| Nano Banana 2 | $0.08 / image at 1K |
| Nano Banana Pro | $0.15 / image |
Nano Banana 2 supports higher-resolution pricing, with 2K and 4K outputs scaling at 1.5x and 2x the 1K rate.
You should also include add-ons in your calculations. Nano Banana 2 charges an extra $0.015 per generation for web search grounding and $0.002 for high-tier thinking, while GPT Image 2 uses quality-tier billing across low, medium, and high settings.