LTX-2.5 is Lightricks' open-source audio-video model. This endpoint generates video timed to a supplied audio clip in a quality-optimized mode, for final visuals synchronized to music, dialogue, or a soundtrack.
LTX logo
lightricks/ltx-2.5/audio-to-video/pro

LTX-2.5 is Lightricks' open-source audio-video model. This endpoint generates video timed to a supplied audio clip in a quality-optimized mode, for final visuals synchronized to music, dialogue, or a soundtrack.

stylized
transform
lip-sync
audio-to-video
Generate video prompts using a variety of techniques including camera direction, style, pacing, special effects and more.
video-prompt-generator

Generate video prompts using a variety of techniques including camera direction, style, pacing, special effects and more.

motion
transformation
chat
llm
Wan Effects generates high-quality videos with popular effects from images
Alibaba logo
wan-effects

Wan Effects generates high-quality videos with popular effects from images

motion
effects
image-to-video
Generate video with audio from images using LTX-2 Distilled
LTX logo
ltx-2-19b/distilled/image-to-video

Generate video with audio from images using LTX-2 Distilled

image-to-video
Wan 2.2's 14B model generates high-resolution, photorealistic images with powerful prompt understanding and fine-grained visual detail
Alibaba logo
wan/v2.2-a14b/text-to-image

Wan 2.2's 14B model generates high-resolution, photorealistic images with powerful prompt understanding and fine-grained visual detail

text-to-image
Interpolate videos with FILM - Frame Interpolation for Large Motion
film/video

Interpolate videos with FILM - Frame Interpolation for Large Motion

interpolation
video-to-video
Stable Audio 3 Small Music Base audio outpainting is the foundational 459 million parameter checkpoint that extends music tracks via causal continuation guided by text prompts.
stable-audio-3/small/music/base/audio-outpainting

Stable Audio 3 Small Music Base audio outpainting is the foundational 459 million parameter checkpoint that extends music tracks via causal continuation guided by text prompts.

music
extension
continuation
audio-to-audio
Adds synchronized, royalty-free, commercial-use-safe sound effects to a video. Returns the finished video with the generated audio mixed in.
sonilo/v1.1/video-to-video-sound-effects

Adds synchronized, royalty-free, commercial-use-safe sound effects to a video. Returns the finished video with the generated audio mixed in.

sfx
audio
effects
video-to-video
Generate high-quality video with audio from images using LTX-2.3
LTX logo
ltx-2.3-quality/image-to-video

Generate high-quality video with audio from images using LTX-2.3

image-to-video
Moondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.
moondream3-preview/point

Moondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.

vision
vision
FLUX.1 Krea [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.
Black Forest Labs logo
flux/krea/image-to-image

FLUX.1 Krea [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.

image-to-image
Image based high quality Virtual Try-On
cat-vton

Image based high quality Virtual Try-On

try-on
fashion
clothing
image-to-image
Tuning-free ID customization.
pulid

Tuning-free ID customization.

editing
customization
personalization
image-to-image
Recraft V4.1 Utility is a faster, lighter variant of V4.1 made for high-volume creative workflows. Ideal for ideation, A/B exploration, and content pipelines, it keeps Recraft's design sensibility while optimizing for throughput and cost.
recraft/v4.1/utility/text-to-image

Recraft V4.1 Utility is a faster, lighter variant of V4.1 made for high-volume creative workflows. Ideal for ideation, A/B exploration, and content pipelines, it keeps Recraft's design sensibility while optimizing for throughput and cost.

stylized
transform
typography
text-to-image
Text-to-image generation with FLUX.2 [klein] 4B Base from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.
Black Forest Labs logo
flux-2/klein/4b/base

Text-to-image generation with FLUX.2 [klein] 4B Base from Black Forest Labs. Enhanced realism, crisper text generation, and native editing capabilities.

text-to-image
Text To Image Model using Boogu-Image
boogu-image

Text To Image Model using Boogu-Image

text-to-image
High-fidelity mask-based video object removal with strong temporal consistency. Erase unwanted objects, people, or elements while preserving aesthetic quality. Trained on licensed data for risk-free commercial use.
Bria AI logo
bria/video/erase/mask

High-fidelity mask-based video object removal with strong temporal consistency. Erase unwanted objects, people, or elements while preserving aesthetic quality. Trained on licensed data for risk-free commercial use.

bria
video
erase
video-to-video
The GenFill Route enables the generation of objects by prompt in a specific region of an image.
You can define the area for object generation by using a mask that outlines the region where the object will be created. Our model is optimized to work seamlessly with blob-shaped masks.
Bria AI logo
bria/genfill/v2

The GenFill Route enables the generation of objects by prompt in a specific region of an image. You can define the area for object generation by using a mask that outlines the region where the object will be created. Our model is optimized to work seamlessly with blob-shaped masks.

image-to-image
Change a video’s lighting using a relit reference frame while preserving the scene, subjects, and original performance. ID-V2V Relight propagates the new illumination across the video.
new
id-v2v/relight

Change a video’s lighting using a relit reference frame while preserving the scene, subjects, and original performance. ID-V2V Relight propagates the new illumination across the video.

relighting
editing
cinematic
video-to-video
Wan 2.2's 5B FastVideo model produces up to 5 seconds of video 720p at 24FPS with fluid motion and powerful prompt understanding
Alibaba logo
wan/v2.2-5b/text-to-video/fast-wan

Wan 2.2's 5B FastVideo model produces up to 5 seconds of video 720p at 24FPS with fluid motion and powerful prompt understanding

text to video
motion
text-to-video
Ideogram V4.0q Tiling generates seamless, edge-matching textures and patterns that repeat infinitely in any direction, ideal for backgrounds, surfaces, and wallpapers.
Ideogram logo
ideogram/v4/tiling

Ideogram V4.0q Tiling generates seamless, edge-matching textures and patterns that repeat infinitely in any direction, ideal for backgrounds, surfaces, and wallpapers.

stylized
transform
realism
image-to-image
Recraft 20b is a new and affordable text-to-image model.
recraft-20b

Recraft 20b is a new and affordable text-to-image model.

image generation
vector art
typograph
text-to-image
A natural-sounding Spanish text-to-speech model optimized for Latin American and European Spanish.
kokoro/spanish

A natural-sounding Spanish text-to-speech model optimized for Latin American and European Spanish.

speech
text-to-audio
State-of-the-art open-source model in aesthetic quality
playground-v25

State-of-the-art open-source model in aesthetic quality

artistic
style
text-to-image
Showing 769 to 792 of 1493 results