Train a LoRA that generates audio from a text prompt — the audio counterpart of text-to-video — learning a sound or style from your clips.
LTX logo
ltx23-trainer-v2/t2a

Train a LoRA that generates audio from a text prompt — the audio counterpart of text-to-video — learning a sound or style from your clips.

training
Outpaint high-quality video using LTX-2.3 with Lora
LTX logo
ltx-2.3-quality/outpaint/lora

Outpaint high-quality video using LTX-2.3 with Lora

outpaint
outpainting
video-to-video
Generate video with audio from images using LTX-2 Distilled and custom LoRA
LTX logo
ltx-2-19b/distilled/image-to-video/lora

Generate video with audio from images using LTX-2 Distilled and custom LoRA

image-to-video
Generate video with audio from videos using LTX-2.3 and custom LoRA
LTX logo
ltx-2.3-22b/video-to-video/lora

Generate video with audio from videos using LTX-2.3 and custom LoRA

video-to-video
Enhance wraped, folded documents with the superior quality of docres for sharper, clearer results.
docres/dewarp

Enhance wraped, folded documents with the superior quality of docres for sharper, clearer results.

image-enhancement
image-to-image
Stable Audio 3 LoRA Trainer fine-tunes Stable Audio 3 base models on paired audio-caption datasets, producing compact LoRA weights that adapt generation toward a custom music style, sound palette, or domain.
stable-audio-3-trainer

Stable Audio 3 LoRA Trainer fine-tunes Stable Audio 3 base models on paired audio-caption datasets, producing compact LoRA weights that adapt generation toward a custom music style, sound palette, or domain.

music
audio
sfx
training
Decompression / Denoise high-quality video using LTX-2.3
LTX logo
ltx-2.3-quality/decompression

Decompression / Denoise high-quality video using LTX-2.3

decompression
denoise
video-to-video
Create group photos
Alibaba logo
qwen-image-edit-2509-lora-gallery/group-photo

Create group photos

stylized
transform
image-to-image
Generate video clips from your prompts using Kling 1.5 (pro)
Kling logo
kling-video/v1.5/pro/effects

Generate video clips from your prompts using Kling 1.5 (pro)

text-to-video
Extend video with audio using LTX-2.3 and custom LoRA
LTX logo
ltx-2.3-22b/extend-video/lora

Extend video with audio using LTX-2.3 and custom LoRA

video-to-video
Heygen Avatar V3 Model for Digital Twin
Heygen logo
heygen/avatar3/digital-twin

Heygen Avatar V3 Model for Digital Twin

text-to-video
Train LTX-2.3 22B for video transformation or video-conditioned generation.
LTX logo
ltx23-v2v-trainer

Train LTX-2.3 22B for video transformation or video-conditioned generation.

ltx2-video
fine-tuning
video-to-video
training
Train custom LoRAs for Wan-2.1 T2V 1.3B
Alibaba logo
wan-trainer/t2v

Train custom LoRAs for Wan-2.1 T2V 1.3B

lora
training
Convert your assets into lottie using Omnilottie.
omnilottie

Convert your assets into lottie using Omnilottie.

lottie
json
Add a realistic scene behind the object with white background
Alibaba logo
qwen-image-edit-2509-lora-gallery/add-background

Add a realistic scene behind the object with white background

stylized
transform
image-to-image
Generate video with audio from reference video, text and images using LTX-2.3 and custom LoRA
LTX logo
ltx-2.3-22b/reference-video-to-video/lora

Generate video with audio from reference video, text and images using LTX-2.3 and custom LoRA

video-to-video
Train custom LoRAs for Wan-2.1 T2V 14B
Alibaba logo
wan-trainer/t2v-14b

Train custom LoRAs for Wan-2.1 T2V 14B

lora
training
Heygen Avatar 4 Digital Twin Model
Heygen logo
heygen/avatar4/digital-twin

Heygen Avatar 4 Digital Twin Model

text-to-video
Remove video backgrounds in real time with Bria’s VRMBG 3.0 model. Built for live streaming, real-time video apps, content creation, and low-latency workflows that need fast, accurate background removal.
Bria AI logo
bria/video/background-removal/realtime

Remove video backgrounds in real time with Bria’s VRMBG 3.0 model. Built for live streaming, real-time video apps, content creation, and low-latency workflows that need fast, accurate background removal.

bria
video
background-removal
video-to-video
Create cinematic transitions and scene progressions (camera movements, framing changes)
Alibaba logo
qwen-image-edit-2509-lora-gallery/next-scene

Create cinematic transitions and scene progressions (camera movements, framing changes)

stylized
transform
image-to-image
SAM 3D enables full scene reconstructions, placing objects and humans in a shared context together.
sam-3/3d-align

SAM 3D enables full scene reconstructions, placing objects and humans in a shared context together.

align
3d
3d-to-3d
An open source, community-driven and native audio turn detection model by Pipecat AI.
smart-turn

An open source, community-driven and native audio turn detection model by Pipecat AI.

speech-to-text
Apply designs/graphics onto people's shirts
Alibaba logo
qwen-image-edit-2509-lora-gallery/shirt-design

Apply designs/graphics onto people's shirts

stylized
transform
image-to-image
Deblur high-quality video using LTX-2.3
LTX logo
ltx-2.3-quality/deblur

Deblur high-quality video using LTX-2.3

deblur
denoise
video-to-video
Sa2VA is an MLLM capable of question answering, visual prompt understanding, and dense object segmentation at both image and video levels
sa2va/4b/video

Sa2VA is an MLLM capable of question answering, visual prompt understanding, and dense object segmentation at both image and video levels

multimodal
vision
MultiTalk model generates a multi-person conversation video from an image and text inputs. Converts text to speech for each person, generating a realistic conversation scene.
ai-avatar/multi-text

MultiTalk model generates a multi-person conversation video from an image and text inputs. Converts text to speech for each person, generating a realistic conversation scene.

stylized
transform
image-to-video
FLUX.1 [dev] Redux is a high-performance endpoint for the FLUX.1 [dev] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.
Black Forest Labs logo
flux-1/dev/redux

FLUX.1 [dev] Redux is a high-performance endpoint for the FLUX.1 [dev] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.

image-to-image
Marlin is a 2B video VLM tuned for the two questions developers actually want to ask of their videos: what is happening, and when?
marlin/find

Marlin is a 2B video VLM tuned for the two questions developers actually want to ask of their videos: what is happening, and when?

utility
editing
vision
Showing 1401 to 1428 of 1504 results