Remove unwanted objects or people from your photos while seamlessly blending the background.
image-editing/object-removal

Remove unwanted objects or people from your photos while seamlessly blending the background.

stylized
transform
image-to-image
FLUX Control LoRA Canny is a high-performance endpoint that uses a control image to transfer structure to the generated image, using a Canny edge map.
Black Forest Labs logo
flux-control-lora-canny

FLUX Control LoRA Canny is a high-performance endpoint that uses a control image to transfer structure to the generated image, using a Canny edge map.

lora
style transfer
text-to-image
Juggernaut Lightning Flux by RunDiffusion provides blazing-fast, high-quality images rendered at five times the speed of Flux. Perfect for mood boards and mass ideation, this model excels in both realism and prompt adherence.
rundiffusion-fal/juggernaut-flux/lightning

Juggernaut Lightning Flux by RunDiffusion provides blazing-fast, high-quality images rendered at five times the speed of Flux. Perfect for mood boards and mass ideation, this model excels in both realism and prompt adherence.

image generation
text-to-image
Qwen Image 2512 LoRA training
Alibaba logo
qwen-image-2512-trainer

Qwen Image 2512 LoRA training

lora
personalization
training
Maya1 is a state-of-the-art speech model by Maya Research for expressive voice generation, built to capture real human emotion and precise voice design.
maya

Maya1 is a state-of-the-art speech model by Maya Research for expressive voice generation, built to capture real human emotion and precise voice design.

tts
text-to-speech
Pixverse's latest v6 Model.
Pixverse logo
pixverse/v6/extend

Pixverse's latest v6 Model.

extend
video-to-video
Professional photo denoising powered by Topaz Labs. Normal, Strong and Extreme presets clean noise at source resolution; Denoise Max adds generative detail recovery. Best for high-ISO and night photography.
Topaz Labs logo
topaz/denoise/image

Professional photo denoising powered by Topaz Labs. Normal, Strong and Extreme presets clean noise at source resolution; Denoise Max adds generative detail recovery. Best for high-ISO and night photography.

restore
image
image-to-image
Meshy-6-Preview is the latest model from Meshy. It generates realistic and production ready 3D models.
meshy/v6-preview/text-to-3d

Meshy-6-Preview is the latest model from Meshy. It generates realistic and production ready 3D models.

text-to-3d
Generate images from text, an image and a mask using Z-Image Turbo, Tongyi-MAI's super-fast 6B model.
Alibaba logo
z-image/turbo/inpaint

Generate images from text, an image and a mask using Z-Image Turbo, Tongyi-MAI's super-fast 6B model.

inpainting
image-to-image
Audio reasoning variant of NVIDIA's Nemotron 3 Nano Omni. 30B A3B hybrid Transformer-Mamba MoE - accepts audio plus a prompt and returns text.
nvidia/nemotron-3-nano-omni/audio

Audio reasoning variant of NVIDIA's Nemotron 3 Nano Omni. 30B A3B hybrid Transformer-Mamba MoE - accepts audio plus a prompt and returns text.

nemotron
nvidia
audio-understanding
audio-to-text
Heygen Translate Model with Extreme Speed
Heygen logo
heygen/v2/translate/speed

Heygen Translate Model with Extreme Speed

video-to-video
Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
florence-2-large/detailed-caption

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

captioning
multimodal
vision
Bria GenFill enables high-quality object addition or visual transformation. Trained exclusively on licensed data for safe and risk-free commercial use. Access the model's source code and weights: https://bria.ai/contact-us
Bria AI logo
bria/genfill

Bria GenFill enables high-quality object addition or visual transformation. Trained exclusively on licensed data for safe and risk-free commercial use. Access the model's source code and weights: https://bria.ai/contact-us

image editing
image-to-image
Transform images, elements, and text into consistent, high-quality video scenes, ensuring stable character identity, object details, and environments.
Kling logo
kling-video/o1/standard/reference-to-video

Transform images, elements, and text into consistent, high-quality video scenes, ensuring stable character identity, object details, and environments.

image-to-video
Fine-tune FLUX.2 [klein] 9B from Black Forest Labs with custom datasets. Create specialized LoRA adaptations for specific editing tasks.
Black Forest Labs logo
flux-2-klein-9b-base-trainer

Fine-tune FLUX.2 [klein] 9B from Black Forest Labs with custom datasets. Create specialized LoRA adaptations for specific editing tasks.

training
Pony V7 is a finetuned text to image for superior aesthetics and prompt following.
pony-v7

Pony V7 is a finetuned text to image for superior aesthetics and prompt following.

diffusion
style
text-to-image
Endpoint for Qwen's Image Editing Plus model also known as Qwen-Image-Edit-2509. Has superior text editing capabilities and multi-image support.
Alibaba logo
qwen-image-edit-2509

Endpoint for Qwen's Image Editing Plus model also known as Qwen-Image-Edit-2509. Has superior text editing capabilities and multi-image support.

image-editing
high-quality-text
image-to-image
Generate film-grade videos from text prompts with native audio, up to 1080p and 15 seconds, using PixVerse C1.
Pixverse logo
pixverse/c1/text-to-video

Generate film-grade videos from text prompts with native audio, up to 1080p and 15 seconds, using PixVerse C1.

video-generation
pixverse
cinematic
text-to-video
LoRA endpoint for the Qwen Image Edit Plus model.
Alibaba logo
qwen-image-edit-plus-lora

LoRA endpoint for the Qwen Image Edit Plus model.

image-editing
image-to-image
Generate video with audio from images using LTX-2
LTX logo
ltx-2-19b/image-to-video

Generate video with audio from images using LTX-2

image-to-video
Generate high quality music and sound effects using Stable Audio 2.5 from StabilityAI
stable-audio-25/audio-to-audio

Generate high quality music and sound effects using Stable Audio 2.5 from StabilityAI

audio
audio-to-audio
Hunyuan World 1.0 turns a single image into a panorama or a 3D world. It creates realistic scenes from the image, allowing you to explore and view it from different angles.
hunyuan_world

Hunyuan World 1.0 turns a single image into a panorama or a 3D world. It creates realistic scenes from the image, allowing you to explore and view it from different angles.

image-to-image
Generate images from your prompts using Luma Photon. Photon is the most creative, personalizable, and intelligent visual models for creatives, bringing a step-function change in the cost of high-quality image generation.
Luma AI logo
luma-photon

Generate images from your prompts using Luma Photon. Photon is the most creative, personalizable, and intelligent visual models for creatives, bringing a step-function change in the cost of high-quality image generation.

text-to-image
Image editing endpoint for Qwen-Image-Max. Qwen Image Max improves upon the Qwen Image Plus series by enhancing the realism and naturalness of images.
Alibaba logo
qwen-image-max/edit

Image editing endpoint for Qwen-Image-Max. Qwen Image Max improves upon the Qwen Image Plus series by enhancing the realism and naturalness of images.

qwen-image
max
image-to-image
Showing 697 to 720 of 1493 results