Image-to-image editing with LoRA support for FLUX.2 [dev] from Black Forest Labs. Specialized style transfer and domain-specific modifications.
Black Forest Labs logo
flux-2/lora/edit

Image-to-image editing with LoRA support for FLUX.2 [dev] from Black Forest Labs. Specialized style transfer and domain-specific modifications.

image-to-image
econstructs a high-fidelity textured 3D model from multiple angle views of one object, with game-ready topology and polygon control
meshy/v7/multi-image-to-3d

econstructs a high-fidelity textured 3D model from multiple angle views of one object, with game-ready topology and polygon control

stylized
transform
image-to-3d
LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.
LTX logo
ltx-2.3/image-to-video

LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.

stylized
transform
lipsync
image-to-video
Design a personalized voice from a text description, and generate speech from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality text-to-speech.
Minimax logo
minimax/voice-design

Design a personalized voice from a text description, and generate speech from text prompts using the MiniMax model, which leverages advanced AI techniques to create high-quality text-to-speech.

speech
text-to-speech
Create stunningly realistic sound effects in seconds - CassetteAI's Sound Effects Model generates high-quality SFX up to 30 seconds long in just 1 second of processing time
cassetteai/sound-effects-generator

Create stunningly realistic sound effects in seconds - CassetteAI's Sound Effects Model generates high-quality SFX up to 30 seconds long in just 1 second of processing time

sound
sfx
sound-effects
text-to-audio
Qwen Image 2512 is an improved version of Qwen Image with better text rendering, finer natural textures, and more realistic human generation.
Alibaba logo
qwen-image-2512

Qwen Image 2512 is an improved version of Qwen Image with better text rendering, finer natural textures, and more realistic human generation.

qwen
2512
text-to-image
Fix distorted or blurred photos of people with CodeFormer.
codeformer

Fix distorted or blurred photos of people with CodeFormer.

image-restoration
faces
utility
image-to-image
Veo 3.1 Lite balances practical utility with professional capabilities, supporting Text-to-Video and Image-to-Video
Google logo
veo3.1/lite/first-last-frame-to-video

Veo 3.1 Lite balances practical utility with professional capabilities, supporting Text-to-Video and Image-to-Video

stylized
transform
lipsync
image-to-video
Transform your photos into ultra-high-resolution 3D models in seconds. Film-quality geometry with PBR textures, ready for games, e-commerce, and 3D printing.
hunyuan3d-v3/image-to-3d

Transform your photos into ultra-high-resolution 3D models in seconds. Film-quality geometry with PBR textures, ready for games, e-commerce, and 3D printing.

image-to-3d
Use SeedVR2 to upscale images, retaining seamless tiling
seedvr/upscale/image/seamless

Use SeedVR2 to upscale images, retaining seamless tiling

upscale
seamless
tiling
image-to-image
Add automatic subtitles to videos
workflow-utilities/auto-subtitle

Add automatic subtitles to videos

auto-subtitle
captioning
video-to-video
Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.
Alibaba logo
wan/v2.7/text-to-video

Wan 2.7 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

stylized
transform
lipsync
text-to-video
Generate realistic videos using Kling O3 from Kling Team!
Kling logo
kling-video/o3/standard/text-to-video

Generate realistic videos using Kling O3 from Kling Team!

text-to-video
Latest object erasing model from Black Forest Labs. Remove undesired objects, texts from images.
Black Forest Labs logo
flux-pro/v1/erase

Latest object erasing model from Black Forest Labs. Remove undesired objects, texts from images.

utility
editing
image-to-image
Recraft V4.1 Vector turns prompts into fully editable SVGs with structured layers and clean geometry. Built for logos, icons, and illustration systems, it produces artwork that goes straight from generation into Figma or Illustrator.
recraft/v4.1/text-to-vector

Recraft V4.1 Vector turns prompts into fully editable SVGs with structured layers and clean geometry. Built for logos, icons, and illustration systems, it produces artwork that goes straight from generation into Figma or Illustrator.

stylized
transform
typography
text-to-image
State of the art Image to 3D Object generation. Generate 3D model from a single image!
tripo3d/tripo/v2.5/image-to-3d

State of the art Image to 3D Object generation. Generate 3D model from a single image!

stylized
image-to-3d
Directional outpainting. Choose edges to expand. left, right, top, or center (uniform all sides). Only expanded areas are generated; an optional zoom-out pulls the frame back by the chosen amount.
image-apps-v2/outpaint

Directional outpainting. Choose edges to expand. left, right, top, or center (uniform all sides). Only expanded areas are generated; an optional zoom-out pulls the frame back by the chosen amount.

outpainting
image-to-image
Generate videos from reference images using Google's Veo 3.1 Fast
Google logo
veo3.1/fast/reference-to-video

Generate videos from reference images using Google's Veo 3.1 Fast

image-to-video
Kling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling
Kling logo
kling-video/o3/4k/image-to-video

Kling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling

stylized
transform
lipsync
image-to-video
Generate music from a simple prompt using ACE-Step
ace-step/prompt-to-audio

Generate music from a simple prompt using ACE-Step

text-to-music
text-to-audio
MiniMax Hailuo-2.3 Image To Video API (Pro, 1080p): Advanced image-to-video generation model with 1080p resolution
Minimax logo
minimax/hailuo-2.3/pro/image-to-video

MiniMax Hailuo-2.3 Image To Video API (Pro, 1080p): Advanced image-to-video generation model with 1080p resolution

image-to-video
Kling 3.0 Turbo Standard is a fast, cost-efficient video generation model that turns text prompts directly into 720P video with native audio, optimized for rapid iteration and high-volume production
Kling logo
kling-video/v3/turbo/standard/text-to-video

Kling 3.0 Turbo Standard is a fast, cost-efficient video generation model that turns text prompts directly into 720P video with native audio, optimized for rapid iteration and high-volume production

stylized
transform
lipsync
text-to-video
Generate realistic lipsync animations from audio using advanced algorithms for high-quality synchronization with PixVerse Lipsync model
Pixverse logo
pixverse/lipsync

Generate realistic lipsync animations from audio using advanced algorithms for high-quality synchronization with PixVerse Lipsync model

animation
lip sync
video-to-video
Isolate audio tracks using ElevenLabs advanced audio isolation technology.
ElevenLabs logo
elevenlabs/audio-isolation

Isolate audio tracks using ElevenLabs advanced audio isolation technology.

audio
audio-to-audio
Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.
chatterbox/text-to-speech

Whether you're working on memes, videos, games, or AI agents, Chatterbox brings your content to life. Use the first tts from resemble ai.

text-to-speech
LTX-2.5 is Lightricks' open-source audio-video model. This endpoint animates a still image into video with synchronized audio in a single pass, in a quality-optimized mode for high-fidelity final output.
LTX logo
lightricks/ltx-2.5/image-to-video/pro

LTX-2.5 is Lightricks' open-source audio-video model. This endpoint animates a still image into video with synchronized audio in a single pass, in a quality-optimized mode for high-fidelity final output.

stylized
transform
lip-sync
image-to-video
FLUX LoRA training optimized for portrait generation, with bright highlights, excellent prompt following and highly detailed results.
Black Forest Labs logo
flux-lora-portrait-trainer

FLUX LoRA training optimized for portrait generation, with bright highlights, excellent prompt following and highly detailed results.

lora
personalization
training
Wan 2.5 text-to-video model.
Alibaba logo
wan-25-preview/text-to-video

Wan 2.5 text-to-video model.

text-to-video
Showing 309 to 336 of 1504 results