
MiDaS depth estimation preprocessor.

M-LSD line segment detection preprocessor.

PIDI (Pidinet) preprocessor.

Segment Anything Model (SAM) preprocessor.

Scribble preprocessor.

TEED (Temporal Edge Enhancement Detection) preprocessor.

ZoeDepth preprocessor.

Depth Anything v2 preprocessor.

Generate short video clips from your images using SVD v1.1

Animate a reference image with a driving video using ControlNeXt.
![A versatile endpoint for the FLUX.1 [dev] model that supports multiple AI extensions including LoRA, ControlNet conditioning, and IP-Adapter integration, enabling comprehensive control over image generation through various guidance methods.](https://refinery.fal.media/url/https%3A%2F%2Fv3b.fal.media%2Ffiles%2Fb%2F0a9f91b2%2Fbbpu6j6ryq35Wu6-WHkpA_2mmO6mnS.png/tr:w-1920,q-80/bbpu6j6ryq35Wu6-WHkpA_2mmO6mnS.webp)
A versatile endpoint for the FLUX.1 [dev] model that supports multiple AI extensions including LoRA, ControlNet conditioning, and IP-Adapter integration, enabling comprehensive control over image generation through various guidance methods.

SAM.

Stable Diffusion 3 Medium (Text to Image) is a Multimodal Diffusion Transformer (MMDiT) model that improves image quality, typography, prompt understanding, and efficiency.

SAM 2 is a model for segmenting images and videos in real-time.

SAM 2 is a model for segmenting images and videos in real-time.

FLUX General Image-to-Image is a versatile endpoint that transforms existing images with support for LoRA, ControlNet, and IP-Adapter extensions, enabling precise control over style transfer, modifications, and artistic variations through multiple guidance methods.

FLUX General Inpainting is a versatile endpoint that enables precise image editing and completion, supporting multiple AI extensions including LoRA, ControlNet, and IP-Adapter for enhanced control over inpainting results and sophisticated image modifications.

FLUX LoRA Image-to-Image is a high-performance endpoint that transforms existing images using FLUX models, leveraging LoRA adaptations to enable rapid and precise image style transfer, modifications, and artistic variations.

A specialized FLUX endpoint combining differential diffusion control with LoRA, ControlNet, and IP-Adapter support, enabling precise, region-specific image transformations through customizable change maps.

Default parameters with automated optimizations and quality improvements.
![Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.](https://refinery.fal.media/url/https%3A%2F%2Fv3b.fal.media%2Ffiles%2Fb%2F0a9f91af%2FUMrx6t6mc59z33d20WIQP_iBIZYKKq.png/tr:w-1920,q-80/UMrx6t6mc59z33d20WIQP_iBIZYKKq.webp)
Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.
![Super fast endpoint for the FLUX.1 [schnell] model with subject input capabilities, enabling rapid and high-quality image generation for personalization, specific styles, brand identities, and product-specific outputs.](https://refinery.fal.media/url/https%3A%2F%2Fstorage.googleapis.com%2Ffalserverless%2Fgallery%2Fflux-subject.webp/tr:w-1920,q-80/flux-subject.webp)
Super fast endpoint for the FLUX.1 [schnell] model with subject input capabilities, enabling rapid and high-quality image generation for personalization, specific styles, brand identities, and product-specific outputs.

Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

Sana can synthesize high-resolution, high-quality images with strong text-image alignment at a remarkably fast speed, with the ability to generate 4K images in less than a second.