
LongCat-Video-Avatar is an audio-driven video generation model that can generates super-realistic, lip-synchronized long video generation with natural dynamics and consistent identity.

Generates vector images that hold a consistent style, from either a saved style ID or reference images attached directly.

Photo restoration model that automatically denoises, deblurs, and enhances old or damaged photos - removes imperfections while preserving original character.

Turn images into pixel-perfect retro art

Use NAFNet to fix issues like blurriness and noise in your images. This model specializes in image restoration and can help enhance the overall quality of your photography.
![LoRA trainer for FLUX.1 Kontext [dev]](https://refinery.fal.media/url/https%3A%2F%2Fv3.fal.media%2Ffiles%2Fmonkey%2FpYXiffttc2Skv36wflufu_dec4efe0d27e4527b64acfbc0e91536a.jpg/tr:w-1920,q-80/pYXiffttc2Skv36wflufu_dec4efe0d27e4527b64acfbc0e91536a.webp)
LoRA trainer for FLUX.1 Kontext [dev]

Photorealistic Text-to-Image

Generate short video clips from your images using SVD v1.1

Create natural HeyGen Avatar V digital twin videos from text or audio, with lip-sync, optional backgrounds, captions, and MP4/WebM output.

Unified image generation with HiDream-O1-Image. Create, edit, and personalize high-resolution images up to 2K—single native model handles text-to-image, editing, and custom subjects without external components.
Generate video with audio from text using LTX-2.3

Generate realistic images.

Extend the beginning or end of provided audio with lyrics and/or style using ACE-Step

Remove background from videos filmed using chromakey, with automatic green spill suppression for clean, professional edges.

Replace your photo's background with any scene you desire, from beach sunsets to urban landscapes, with perfect lighting and shadows
![ControlLight is a LoRA fine-tune of FLUX.2 [klein] 9B that enhances low-light images while preserving scene structure and fine details, with a single alpha parameter that gives continuous control over enhancement strength from subtle to full brightening.](https://refinery.fal.media/url/https%3A%2F%2Fv3b.fal.media%2Ffiles%2Fb%2F0a9be8bb%2F8dTehLaCr78bz4vpkcZhb_a5bf209765d449658613b94c03976271.jpg/tr:w-1920,q-80/8dTehLaCr78bz4vpkcZhb_a5bf209765d449658613b94c03976271.webp)
ControlLight is a LoRA fine-tune of FLUX.2 [klein] 9B that enhances low-light images while preserving scene structure and fine details, with a single alpha parameter that gives continuous control over enhancement strength from subtle to full brightening.

Generate long videos from images using LongCat Video Distilled

Generate 3D models from your images using Hunyuan 3D. A native 3D generative model enabling versatile and high-quality 3D asset creation.

Stable Audio 3 Small Music Base is the foundational 459 million parameter checkpoint generating full music compositions up to 2 minutes from text prompts, intended as the unmodified base for fine-tuning.

Discover ultimate control with Pikaframes key frame interpolation, a stunning image-to-video feature that allows you to upload up to 5 keyframes, customize their transition length and prompt, and see their images come to life as seamless videos.

Vision

Image to Video for the high-quality Hunyuan Video I2V model.

State of the art Image to 3D Object generation
![Generate high-quality images from depth maps using Flux.1 [dev] depth estimation model. The model produces accurate depth representations for scene understanding and 3D visualization.](https://refinery.fal.media/url/https%3A%2F%2Fstorage.googleapis.com%2Ffalserverless%2Fgallery%2Fflux_lora.jpg/tr:w-1920,q-80/flux_lora.webp)
Generate high-quality images from depth maps using Flux.1 [dev] depth estimation model. The model produces accurate depth representations for scene understanding and 3D visualization.