LongCat-Video-Avatar is an audio-driven video generation model that can generates super-realistic, lip-synchronized long video generation with natural dynamics and consistent identity.
longcat-single-avatar/image-audio-to-video

LongCat-Video-Avatar is an audio-driven video generation model that can generates super-realistic, lip-synchronized long video generation with natural dynamics and consistent identity.

image-to-video
audio-to-video
Generates vector images that hold a consistent style, from either a saved style ID or reference images attached directly.
new
recraft/v4/style/pro/text-to-vector

Generates vector images that hold a consistent style, from either a saved style ID or reference images attached directly.

stylized
transform
editing
text-to-image
Photo restoration model that automatically denoises, deblurs, and enhances old or damaged photos - removes imperfections while preserving original character.
Bria AI logo
bria/fibo-edit/restore

Photo restoration model that automatically denoises, deblurs, and enhances old or damaged photos - removes imperfections while preserving original character.

image-restoration
fibo-edit
bria
image-to-image
Turn images into pixel-perfect retro art
image2pixel

Turn images into pixel-perfect retro art

post-processing
pixel-art
image-to-image
Use NAFNet to fix issues like blurriness and noise in your images. This model specializes in image restoration and can help enhance the overall quality of your photography.
nafnet/deblur

Use NAFNet to fix issues like blurriness and noise in your images. This model specializes in image restoration and can help enhance the overall quality of your photography.

image-restoration
deblur
denoise
image-to-image
LoRA trainer for FLUX.1 Kontext [dev]
Black Forest Labs logo
flux-kontext-trainer

LoRA trainer for FLUX.1 Kontext [dev]

training
Photorealistic Text-to-Image
kolors

Photorealistic Text-to-Image

realism
diffusion
text-to-image
Generate short video clips from your images using SVD v1.1
stable-video

Generate short video clips from your images using SVD v1.1

image-to-video
Create natural HeyGen Avatar V digital twin videos from text or audio, with lip-sync, optional backgrounds, captions, and MP4/WebM output.
Heygen logo
heygen/avatar5/digital-twin

Create natural HeyGen Avatar V digital twin videos from text or audio, with lip-sync, optional backgrounds, captions, and MP4/WebM output.

avatar
digital-twin
talking-avatar
text-to-video
Unified image generation with HiDream-O1-Image. Create, edit, and personalize high-resolution images up to 2K—single native model handles text-to-image, editing, and custom subjects without external components.
hidream-o1-image/dev/edit

Unified image generation with HiDream-O1-Image. Create, edit, and personalize high-resolution images up to 2K—single native model handles text-to-image, editing, and custom subjects without external components.

image-to-image
Generate video with audio from text using LTX-2.3
LTX logo
ltx-2.3-22b/text-to-video

Generate video with audio from text using LTX-2.3

text-to-video
Generate realistic images.
realistic-vision

Generate realistic images.

realism
diffusion
text-to-image
Extend the beginning or end of provided audio with lyrics and/or style using ACE-Step
ace-step/audio-outpaint

Extend the beginning or end of provided audio with lyrics and/or style using ACE-Step

audio-outpaint
audio-extend
audio-to-audio
Remove background from videos filmed using chromakey, with automatic green spill suppression for clean, professional edges.
new
Bria AI logo
bria/video/background-removal/green-screen-despill

Remove background from videos filmed using chromakey, with automatic green spill suppression for clean, professional edges.

video-to-video
Replace your photo's background with any scene you desire, from beach sunsets to urban landscapes, with perfect lighting and shadows
image-editing/background-change

Replace your photo's background with any scene you desire, from beach sunsets to urban landscapes, with perfect lighting and shadows

stylized
transform
image-to-image
ControlLight is a LoRA fine-tune of FLUX.2 [klein] 9B that enhances low-light images while preserving scene structure and fine details, with a single alpha parameter that gives continuous control over enhancement strength from subtle to full brightening.
control-light

ControlLight is a LoRA fine-tune of FLUX.2 [klein] 9B that enhances low-light images while preserving scene structure and fine details, with a single alpha parameter that gives continuous control over enhancement strength from subtle to full brightening.

stylized
transform
image-to-image
Generate long videos from images using LongCat Video Distilled
longcat-video/distilled/image-to-video/480p

Generate long videos from images using LongCat Video Distilled

image-to-video
Generate 3D models from your images using Hunyuan 3D. A native 3D generative model enabling versatile and high-quality 3D asset creation.
hunyuan3d/v2/turbo

Generate 3D models from your images using Hunyuan 3D. A native 3D generative model enabling versatile and high-quality 3D asset creation.

stylized
image-to-3d
Stable Audio 3 Small Music Base is the foundational 459 million parameter checkpoint generating full music compositions up to 2 minutes from text prompts, intended as the unmodified base for fine-tuning.
stable-audio-3/small/music/base/text-to-audio

Stable Audio 3 Small Music Base is the foundational 459 million parameter checkpoint generating full music compositions up to 2 minutes from text prompts, intended as the unmodified base for fine-tuning.

music
on-device
lightweight
text-to-audio
Discover ultimate control with Pikaframes key frame interpolation, a stunning image-to-video feature that allows you to upload up to 5 keyframes, customize their transition length and prompt, and see their images come to life as seamless videos.
pika/v2.2/pikaframes

Discover ultimate control with Pikaframes key frame interpolation, a stunning image-to-video feature that allows you to upload up to 5 keyframes, customize their transition length and prompt, and see their images come to life as seamless videos.

image-to-video
Vision
llava-next

Vision

multimodal
vision
Image to Video for the high-quality Hunyuan Video I2V model.
hunyuan-video-image-to-video

Image to Video for the high-quality Hunyuan Video I2V model.

motion
image-to-video
State of the art Image to 3D Object generation
triposr

State of the art Image to 3D Object generation

image-to-3d
Generate high-quality images from depth maps using Flux.1 [dev] depth estimation model. The model produces accurate depth representations for scene understanding and 3D visualization.
Black Forest Labs logo
flux-lora-depth

Generate high-quality images from depth maps using Flux.1 [dev] depth estimation model. The model produces accurate depth representations for scene understanding and 3D visualization.

depth
lora
utility
image-to-image
Showing 817 to 840 of 1493 results