
Train a LoRA that generates audio from a text prompt — the audio counterpart of text-to-video — learning a sound or style from your clips.

Outpaint high-quality video using LTX-2.3 with Lora

Generate video with audio from images using LTX-2 Distilled and custom LoRA

Generate video with audio from videos using LTX-2.3 and custom LoRA

Enhance wraped, folded documents with the superior quality of docres for sharper, clearer results.

Stable Audio 3 LoRA Trainer fine-tunes Stable Audio 3 base models on paired audio-caption datasets, producing compact LoRA weights that adapt generation toward a custom music style, sound palette, or domain.

Decompression / Denoise high-quality video using LTX-2.3

Create group photos

Generate video clips from your prompts using Kling 1.5 (pro)

Extend video with audio using LTX-2.3 and custom LoRA

Heygen Avatar V3 Model for Digital Twin

Train LTX-2.3 22B for video transformation or video-conditioned generation.

Train custom LoRAs for Wan-2.1 T2V 1.3B

Convert your assets into lottie using Omnilottie.

Add a realistic scene behind the object with white background

Generate video with audio from reference video, text and images using LTX-2.3 and custom LoRA

Train custom LoRAs for Wan-2.1 T2V 14B

Heygen Avatar 4 Digital Twin Model

Remove video backgrounds in real time with Bria’s VRMBG 3.0 model. Built for live streaming, real-time video apps, content creation, and low-latency workflows that need fast, accurate background removal.

Create cinematic transitions and scene progressions (camera movements, framing changes)

SAM 3D enables full scene reconstructions, placing objects and humans in a shared context together.

An open source, community-driven and native audio turn detection model by Pipecat AI.

Apply designs/graphics onto people's shirts

Deblur high-quality video using LTX-2.3

Sa2VA is an MLLM capable of question answering, visual prompt understanding, and dense object segmentation at both image and video levels

MultiTalk model generates a multi-person conversation video from an image and text inputs. Converts text to speech for each person, generating a realistic conversation scene.
![FLUX.1 [dev] Redux is a high-performance endpoint for the FLUX.1 [dev] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.](https://refinery.fal.media/url/https%3A%2F%2Fstorage.googleapis.com%2Ffal_cdn%2Ffal%2Ffor%2520videos-1.jpg/tr:w-1920,q-80/for%20videos-1.webp)
FLUX.1 [dev] Redux is a high-performance endpoint for the FLUX.1 [dev] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.

Marlin is a 2B video VLM tuned for the two questions developers actually want to ask of their videos: what is happening, and when?