
Restore old or damaged photos by fixing colors, scratches, and resolution.

Professional image restoration powered by Topaz Labs. Recover 3 generatively rebuilds natural detail; Dust-Scratch V2 cleans film dust and scratches. Best for old, damaged or degraded photos.

Generate video clips from your images using MiniMax Video model

Luma Ray 3.2 reframes an existing video into a new aspect ratio guided by a text prompt, preserving the original footage frame-for-frame while controlling resolution and outpainting the surrounding canvas.

Extend Veo-Created Videos up to 30 seconds

Luma Uni-1 Edit reworks a source image from a text instruction, preserving the original composition while applying style changes and following optional reference images to steer the result.

Generate high quality video clips from text and image prompts using PixVerse v5

Meshy-5 retexture applies new, high-quality textures to existing 3D models using either text prompts or reference images. It supports PBR material generation for realistic, production-ready results.
![Text-to-image generation with FLUX.2 [klein] 9B from Black Forest Labs and custom LoRA.](https://refinery.fal.media/url/https%3A%2F%2Fv3b.fal.media%2Ffiles%2Fb%2F0a928e3b%2FsIC-Ne9BMwZZtBvR3FwKN_9a724704a550471a9df59999e9e1017f.jpg/tr:w-1920,q-80/sIC-Ne9BMwZZtBvR3FwKN_9a724704a550471a9df59999e9e1017f.webp)
Text-to-image generation with FLUX.2 [klein] 9B from Black Forest Labs and custom LoRA.

Run any LLM (Large Language Model) with fal, powered by OpenRouter.

Wan-2.1 is a text-to-video model that generates high-quality videos with high visual quality and motion diversity from text prompts

VEED Fabric 1.0 is an image-to-video API that turns any image into a talking video

Generate character-consistent videos from reference images using PixVerse C1, with subject and background references.

Prompt-free object removal from an image and mask, erasing objects with their shadows and reflections and reconstructing the scene cleanly.

Extend Veo-Created Videos up to 30 seconds

Generate high-fidelity, design-ready images with precise typography, strong prompt alignment, and rich visual detail using Microsoft's flagship MAI Image 2.5 Pro.

FLUX Control LoRA Depth is a high-performance endpoint that uses a control image to transfer structure to the generated image, using a depth map.

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

Pixal3D turns a single image into a high-fidelity 3D model with detailed geometry and realistic textures.

Enhances a given raster image using the 'creative upscale' tool, increasing image resolution, making the image sharper and cleaner.

MuseTalk is a real-time high quality audio-driven lip-syncing model. Use MuseTalk to animate a face with your own audio.

Precise camera position and angle control (rotation, zoom, vertical movement)

Generate 3D models from a single image using Tripo P1.

Audio-driven talking avatar generation powered by the SoulX-FlashTalk 14B model.