Train custom LoRAs for Wan-2.1 T2V 1.3B
Alibaba logo
wan-trainer/t2v

Train custom LoRAs for Wan-2.1 T2V 1.3B

lora
training
PIDI (Pidinet) preprocessor.
image-preprocessors/pidi

PIDI (Pidinet) preprocessor.

detection
preprocess
utility
image-to-image
Generate full portrait from a cropped face photo
Alibaba logo
qwen-image-edit-2509-lora-gallery/face-to-full-portrait

Generate full portrait from a cropped face photo

stylized
transform
image-to-image
Inpaint high-quality video using LTX-2.3  with lora
LTX logo
ltx-2.3-quality/inpaint/lora

Inpaint high-quality video using LTX-2.3 with lora

inpaint
video-to-video
Text to Audio high-quality using LTX-2.3 with Lora
LTX logo
ltx-2.3-quality/text-to-audio/lora

Text to Audio high-quality using LTX-2.3 with Lora

text-to-audio
One-to-All Animation is a pose driven video model that animates characters from a single reference image, enabling flexible, alignment-free motion transfer across diverse styles and scenes
one-to-all-animation/1.3b

One-to-All Animation is a pose driven video model that animates characters from a single reference image, enabling flexible, alignment-free motion transfer across diverse styles and scenes

video to video
motion
video-to-video
Generate video with audio from reference video, text and images using LTX-2.3 and custom LoRA
LTX logo
ltx-2.3-22b/reference-video-to-video/lora

Generate video with audio from reference video, text and images using LTX-2.3 and custom LoRA

video-to-video
Outpaint high-quality video using LTX-2.3 with Lora
LTX logo
ltx-2.3-quality/outpaint/lora

Outpaint high-quality video using LTX-2.3 with Lora

outpaint
outpainting
video-to-video
Reduce color saturation using different methods (luminance Rec.709, luminance Rec.601, average, lightness) with adjustable factor.
post-processing/desaturate

Reduce color saturation using different methods (luminance Rec.709, luminance Rec.601, average, lightness) with adjustable factor.

stylized
transform
image-to-image
PersonaPlex is a real-time, full-duplex speech-to-speech conversational model that enables persona control through text-based role prompts and audio-based voice conditioning.
personaplex/realtime

PersonaPlex is a real-time, full-duplex speech-to-speech conversational model that enables persona control through text-based role prompts and audio-based voice conditioning.

realtime
conversational
audio-to-audio
Sa2VA is an MLLM capable of question answering, visual prompt understanding, and dense object segmentation at both image and video levels
sa2va/4b/video

Sa2VA is an MLLM capable of question answering, visual prompt understanding, and dense object segmentation at both image and video levels

multimodal
vision
MultiTalk model generates a multi-person conversation video from an image and text inputs. Converts text to speech for each person, generating a realistic conversation scene.
ai-avatar/multi-text

MultiTalk model generates a multi-person conversation video from an image and text inputs. Converts text to speech for each person, generating a realistic conversation scene.

stylized
transform
image-to-video
Structured Instructions Generation endpoint for Fibo Edit, Bria's newest editing model.
Bria AI logo
bria/fibo-edit/edit/structured_instruction

Structured Instructions Generation endpoint for Fibo Edit, Bria's newest editing model.

structured-prompt-generation
fibo-edit
json
text-to-json
Decompression / Denoise high-quality video using LTX-2.3
LTX logo
ltx-2.3-quality/decompression

Decompression / Denoise high-quality video using LTX-2.3

decompression
denoise
video-to-video
Deblur high-quality video using LTX-2.3
LTX logo
ltx-2.3-quality/deblur

Deblur high-quality video using LTX-2.3

deblur
denoise
video-to-video
Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
florence-2-large/region-to-description

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

multimodal
vision
Day to Night for high-quality video using LTX-2.3
LTX logo
ltx-2.3-quality/day-to-night

Day to Night for high-quality video using LTX-2.3

day
night
video-to-video
Generate video with audio from images using LTX-2 Distilled and custom LoRA
LTX logo
ltx-2-19b/distilled/image-to-video/lora

Generate video with audio from images using LTX-2 Distilled and custom LoRA

image-to-video
SAM 3D enables full scene reconstructions, placing objects and humans in a shared context together.
sam-3/3d-align

SAM 3D enables full scene reconstructions, placing objects and humans in a shared context together.

align
3d
3d-to-3d
Generate video with audio from reference videos using LTX-2.3 Distilled
LTX logo
ltx-2.3-22b/distilled/reference-video-to-video

Generate video with audio from reference videos using LTX-2.3 Distilled

video-to-video
A high-fidelity capability for erasing unwanted objects, people, or visual elements from videos while maintaining aesthetic quality and temporal consistency.
Bria AI logo
bria/bria_video_eraser/erase/keypoints

A high-fidelity capability for erasing unwanted objects, people, or visual elements from videos while maintaining aesthetic quality and temporal consistency.

bria
erase
video-to-video
Generate video clips from your prompts using Kling 1.5 (pro)
Kling logo
kling-video/v1.5/pro/effects

Generate video clips from your prompts using Kling 1.5 (pro)

text-to-video
Generate video with audio from reference videos using LTX-2.3 Distilled and custom LoRA
LTX logo
ltx-2.3-22b/distilled/reference-video-to-video/lora

Generate video with audio from reference videos using LTX-2.3 Distilled and custom LoRA

video-to-video
Generate subject consistent videos using Lynx from ByteDance!
Bytedance logo
bytedance/lynx

Generate subject consistent videos using Lynx from ByteDance!

subject
image-to-video
Showing 1393 to 1416 of 1491 results