
Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation

A natural and expressive Brazilian Portuguese text-to-speech model optimized for clarity and fluency.

Create seamless cinematic transitions between two images with PixVerse C1, with native audio and up to 1080p.

Wan-2.2 text-to-video is a video model that generates high-quality videos with high visual quality and motion diversity from text prompts. This endpoint supports LoRAs made for Wan 2.2.

Turn up to five reference images into one continuous, consistent video with Bernini-R, with smooth, stable camera motion and no scene cuts.
FLUX.1 [schnell] Redux is a high-performance endpoint for the FLUX.1 [schnell] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.

Vector font generation with VecGlypher. Create custom glyphs from text descriptions or reference images—outputs clean SVG paths directly without raster-to-vector conversion.

Split 3D models into parts with Hunyuan 3D

Generate 3D models from text descriptions using Tripo P1.

Interpolate images with RIFE - Real-Time Intermediate Flow Estimation
![Super fast text-to-image endpoint for the FLUX.1 Kontext [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.](https://refinery.fal.media/url/https%3A%2F%2Fstorage.googleapis.com%2Ffal_cdn%2Ffal%2FTraining-4.jpg/tr:w-1920,q-80/Training-4.webp)
Super fast text-to-image endpoint for the FLUX.1 Kontext [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

Generate high quality video clips with different effects using PixVerse v5
![FLUX1.1 [pro] ultra fine-tuned is the newest version of FLUX1.1 [pro] with a fine-tuned LoRA, maintaining professional-grade image quality while delivering up to 2K resolution with improved photo realism.](https://refinery.fal.media/url/https%3A%2F%2Fstorage.googleapis.com%2Ffalserverless%2Fgallery%2Fflux-pro-11-ultra.webp/tr:w-1920,q-80/flux-pro-11-ultra.webp)
FLUX1.1 [pro] ultra fine-tuned is the newest version of FLUX1.1 [pro] with a fine-tuned LoRA, maintaining professional-grade image quality while delivering up to 2K resolution with improved photo realism.

Blend products into backgrounds with automatic perspective and lighting correction

SCAIL-2 is an end-to-end character animation model that drives a reference character from a source video without relying on intermediate pose representations like skeleton maps.

Generate professional headshot photos with customizable backgrounds.

An efficent SDXL multi-controlnet image-to-image model.

Add a background to images with white/clean background

Accelerated image generation with Ideogram V2A Turbo. Create high-quality visuals, posters, and logos with enhanced speed while maintaining Ideogram's signature quality.

Generate long videos in 720p/30fps from text using LongCat Video Distilled

Reframe entire videos scene-by-scene using Wan VACE 2.1

Infinitalk model generates a talking avatar video from an image and audio file. The avatar lip-syncs to the provided audio with natural facial expressions.

Kandinsky 5.0 Pro is a diffusion model for fast, high-quality text-to-video generation.

Professional motion deblur powered by Topaz Labs. Themis 2 restores clarity to fast-moving, motion-blurred footage at source resolution. Best for sports and action footage.