Apply realistic makeup styles with adjustable intensity.
image-apps-v2/makeup-application

Apply realistic makeup styles with adjustable intensity.

makeup
transform
image-to-image
A fast and natural-sounding Japanese text-to-speech model optimized for smooth pronunciation.
kokoro/japanese

A fast and natural-sounding Japanese text-to-speech model optimized for smooth pronunciation.

speech
text-to-audio
Generate short video clips from your images using SVD v1.1 at Lightning Speed
fast-svd-lcm

Generate short video clips from your images using SVD v1.1 at Lightning Speed

turbo
image-to-video
Foley Control is a video-to-audio model that automatically generates synchronized sound effects for videos, using text prompts to shape the type of sound while matching the timing and action on screen.
controlfoley

Foley Control is a video-to-audio model that automatically generates synchronized sound effects for videos, using text prompts to shape the type of sound while matching the timing and action on screen.

stylized
transform
lipsync
video-to-video
A high-quality Italian text-to-speech model delivering smooth and expressive speech synthesis.
kokoro/italian

A high-quality Italian text-to-speech model delivering smooth and expressive speech synthesis.

speech
text-to-audio
Interpolate between image frames
amt-interpolation/frame-interpolation

Interpolate between image frames

interpolation
editing
image-to-video
Generate Images with ControlNet.
fast-sdxl-controlnet-canny

Generate Images with ControlNet.

diffusion
controlnet
manipulation
text-to-image
Segment Anything Model (SAM) preprocessor.
image-preprocessors/sam

Segment Anything Model (SAM) preprocessor.

segmentation
preprocess
utility
image-to-image
Vidu Q1 Start-End to Video generates smooth transition 1080p videos between specified start and end images.
vidu/q1/start-end-to-video

Vidu Q1 Start-End to Video generates smooth transition 1080p videos between specified start and end images.

stylized
transform
image-to-video
Stable Audio 3 Medium Base audio-to-audio is the foundational 1.4 billion parameter checkpoint that transforms input audio into new stereo variations up to 6 minutes guided by text prompts.
stable-audio-3/medium/base/audio-to-audio

Stable Audio 3 Medium Base audio-to-audio is the foundational 1.4 billion parameter checkpoint that transforms input audio into new stereo variations up to 6 minutes guided by text prompts.

music
style-transfer
remix
audio-to-audio
Hunyuan Video 1.5 is Tencent's latest and best video model
hunyuan-video-v1.5/image-to-video

Hunyuan Video 1.5 is Tencent's latest and best video model

image-to-video
An efficent SDXL multi-controlnet text-to-image model.
sdxl-controlnet-union

An efficent SDXL multi-controlnet text-to-image model.

diffusion
controlnet
composition
text-to-image
Precise camera position and angle control (rotation, zoom, vertical movement)
Alibaba logo
qwen-image-edit-2509-lora-gallery/multiple-angles

Precise camera position and angle control (rotation, zoom, vertical movement)

stylized
transform
image-to-image
Photorealistic Image-to-Image
kolors/image-to-image

Photorealistic Image-to-Image

realism
editing
diffusion
image-to-image
Generate video with audio from videos using LTX-2
LTX logo
ltx-2-19b/video-to-video

Generate video with audio from videos using LTX-2

video-to-video
Transform your consistent character into different art styles, settings, or scenarios while maintaining their distinctive appearance and identity
Ideogram logo
ideogram/character/remix

Transform your consistent character into different art styles, settings, or scenarios while maintaining their distinctive appearance and identity

character-consistency
image-to-image
FLUX.1 [dev] Redux is a high-performance endpoint for the FLUX.1 [dev] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.
Black Forest Labs logo
flux/dev/redux

FLUX.1 [dev] Redux is a high-performance endpoint for the FLUX.1 [dev] model that enables rapid transformation of existing images, delivering high-quality style transfers and image modifications with the core FLUX capabilities.

image-to-image
ZoeDepth preprocessor.
image-preprocessors/zoe

ZoeDepth preprocessor.

depth
preprocess
utility
image-to-image
Generate videos from prompts using CogVideoX-5B
cogvideox-5b

Generate videos from prompts using CogVideoX-5B

text-to-video
Default parameters with automated optimizations and quality improvements.
fooocus

Default parameters with automated optimizations and quality improvements.

stylized
text-to-image
Juggernaut Base Flux LoRA by RunDiffusion is a drop-in replacement for Flux [Dev] that delivers sharper details, richer colors, and enhanced realism to all your LoRAs and LyCORIS with full compatibility.
rundiffusion-fal/juggernaut-flux-lora

Juggernaut Base Flux LoRA by RunDiffusion is a drop-in replacement for Flux [Dev] that delivers sharper details, richer colors, and enhanced realism to all your LoRAs and LyCORIS with full compatibility.

image generation
text-to-image
Vidu Reference to Video creates videos by using a reference images and combining them with a prompt.
vidu/reference-to-video

Vidu Reference to Video creates videos by using a reference images and combining them with a prompt.

motion
reference
image-to-video
Generate video with audio from audio, text and images using LTX-2
LTX logo
ltx-2.3-22b/audio-to-video

Generate video with audio from audio, text and images using LTX-2

audio-to-video
Generate videos from prompts using LTX Video-0.9.5
LTX logo
ltx-video-v095

Generate videos from prompts using LTX Video-0.9.5

video
text-video
text-to-video
Showing 1009 to 1032 of 1493 results