Heygen Text to Video Generation Model
Heygen logo
heygen/v2/video-agent

Heygen Text to Video Generation Model

text-to-video
Extend videos using LTX Video-0.9.7 13B Distilled and custom LoRA
LTX logo
ltx-video-13b-distilled/extend

Extend videos using LTX Video-0.9.7 13B Distilled and custom LoRA

video
ltx-video
extend-video
video-to-video
Rapidly create image variations with Ideogram V2 Turbo Remix. Fast and efficient reimagining of existing images while maintaining creative control through prompt guidance.
Ideogram logo
ideogram/v2/turbo/remix

Rapidly create image variations with Ideogram V2 Turbo Remix. Fast and efficient reimagining of existing images while maintaining creative control through prompt guidance.

realism
typography
image-to-image
Scribble preprocessor.
image-preprocessors/scribble

Scribble preprocessor.

preprocess
utility
editing
image-to-image
Extend videos using LTX Video-0.9.8 13B Distilled and custom LoRA
LTX logo
ltxv-13b-098-distilled/extend

Extend videos using LTX Video-0.9.8 13B Distilled and custom LoRA

ltx-video
extend
video-to-video
Blend products into backgrounds with automatic perspective and lighting correction
Alibaba logo
qwen-image-edit-2509-lora-gallery/integrate-product

Blend products into backgrounds with automatic perspective and lighting correction

stylized
transform
image-to-image
InfinityStar’s unified 8B spacetime autoregressive engine to turn any text prompt into crisp 720p videos - 10× faster than diffusion models.
infinity-star/text-to-video

InfinityStar’s unified 8B spacetime autoregressive engine to turn any text prompt into crisp 720p videos - 10× faster than diffusion models.

text-to-video
Generate high-quality video from a text prompt with Bernini-R, ByteDance's unified video generation and editing model.
Bytedance logo
bernini-r/text-to-video

Generate high-quality video from a text prompt with Bernini-R, ByteDance's unified video generation and editing model.

cinematic
text-to-video
Train a LoRA that generates audio from a text prompt — the audio counterpart of text-to-video — learning a sound or style from your clips.
LTX logo
ltx23-trainer-v2/t2a

Train a LoRA that generates audio from a text prompt — the audio counterpart of text-to-video — learning a sound or style from your clips.

training
An open source, community-driven and native audio turn detection model by Pipecat AI.
smart-turn

An open source, community-driven and native audio turn detection model by Pipecat AI.

speech-to-text
Convert a textured 3D model into a multicolor model suited for multicolor 3D printing with Hi3D.
hitem3d/hi3d/multicolor

Convert a textured 3D model into a multicolor model suited for multicolor 3D printing with Hi3D.

multicolor
3d-to-3d
Generate video with audio from text using LTX-2 and custom LoRA
LTX logo
ltx-2-19b/text-to-video/lora

Generate video with audio from text using LTX-2 and custom LoRA

text-to-video
Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
florence-2-large/region-to-segmentation

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

multimodal
vision
segmentation
image-to-image
Train LTX-2.3 22B for video transformation or video-conditioned generation.
LTX logo
ltx23-v2v-trainer

Train LTX-2.3 22B for video transformation or video-conditioned generation.

ltx2-video
fine-tuning
video-to-video
training
Removes harsh shadows and light spots from images, replacing them with soft, even, natural-looking illumination.
Alibaba logo
qwen-image-edit-plus-lora-gallery/lighting-restoration

Removes harsh shadows and light spots from images, replacing them with soft, even, natural-looking illumination.

stylized
transform
image-to-image
Heygen Avatar 4 Digital Twin Model
Heygen logo
heygen/avatar4/digital-twin

Heygen Avatar 4 Digital Twin Model

text-to-video
Train a MiniMax H3 LoRA on first/last/both keyframe signatures, teaching it to generate video with audio that starts on one image and lands on another.
Minimax logo
minimax/h3/flf2v/trainer

Train a MiniMax H3 LoRA on first/last/both keyframe signatures, teaching it to generate video with audio that starts on one image and lands on another.

utility
editing
training
Convert your assets into lottie using Omnilottie.
omnilottie

Convert your assets into lottie using Omnilottie.

lottie
json
Extend any sound effect with seamless, natural tails.
mirelo-ai/sfx1.6/extend-audio

Extend any sound effect with seamless, natural tails.

sfx
audio-to-audio
Stable Audio 3 LoRA Trainer fine-tunes Stable Audio 3 base models on paired audio-caption datasets, producing compact LoRA weights that adapt generation toward a custom music style, sound palette, or domain.
stable-audio-3-trainer

Stable Audio 3 LoRA Trainer fine-tunes Stable Audio 3 base models on paired audio-caption datasets, producing compact LoRA weights that adapt generation toward a custom music style, sound palette, or domain.

music
audio
sfx
training
OpenAI spec compatible endpoint of Isaac-01 which is a multimodal vision-language model from Perceptron for various vision language tasks.
perceptron/isaac-01/openai/v1/chat/completions

OpenAI spec compatible endpoint of Isaac-01 which is a multimodal vision-language model from Perceptron for various vision language tasks.

multimodal
vision
Generate video with audio from text using LTX-2.3 Distilled and custom LoRA
LTX logo
ltx-2.3-22b/distilled/text-to-video/lora

Generate video with audio from text using LTX-2.3 Distilled and custom LoRA

text-to-video
The Avatar X API offers access to Mirage's most advanced generation model yet, delivering
industry-leading identity preservation and expressivity in AI video
mirage-api/avatar-x/text-to-video

The Avatar X API offers access to Mirage's most advanced generation model yet, delivering industry-leading identity preservation and expressivity in AI video

avatar
lipsync
talking-head
text-to-video
Invisible Watermark is a model that can add an invisible watermark to an image.
invisible-watermark

Invisible Watermark is a model that can add an invisible watermark to an image.

utility
editing
image-to-image
Showing 1369 to 1392 of 1491 results