Transfer expression from a video to a portrait.
live-portrait

Transfer expression from a video to a portrait.

expression
animation
image-to-video
Luma Uni-1 turns a text prompt into a single high-fidelity image, with control over aspect ratio and visual style, plus optional web-sourced and reference-image guidance for sharper grounding.
Luma AI logo
luma/agent/uni-1/v1/text-to-image

Luma Uni-1 turns a text prompt into a single high-fidelity image, with control over aspect ratio and visual style, plus optional web-sourced and reference-image guidance for sharper grounding.

realism
typography
stylized
text-to-image
Moondream2 is a highly efficient open-source vision language model that combines powerful image understanding capabilities with a remarkably small footprint.
moondream2

Moondream2 is a highly efficient open-source vision language model that combines powerful image understanding capabilities with a remarkably small footprint.

vision
vision
Restyle a video’s scene, lighting, and visual style from edited keyframes while preserving the source subjects’ identity, expressions, gaze, and motion. Developed by Eyeline Labs and Netflix researchers.
new
id-v2v

Restyle a video’s scene, lighting, and visual style from edited keyframes while preserving the source subjects’ identity, expressions, gaze, and motion. Developed by Eyeline Labs and Netflix researchers.

stylized
transform
editing
video-to-video
Wan 2.6 text-to-image model.
Alibaba logo
wan/v2.6/text-to-image

Wan 2.6 text-to-image model.

text-to-image
State of the art Multiview to 3D Object generation. Generate 3D models from multiple images!
tripo3d/tripo/v2.5/multiview-to-3d

State of the art Multiview to 3D Object generation. Generate 3D models from multiple images!

stylized
multiview
image-to-3d
Create high-quality images with accurate text rendering and rich knowledge details—supports editing, style transfer, and maintaining consistent characters across multiple images.
glm-image/image-to-image

Create high-quality images with accurate text rendering and rich knowledge details—supports editing, style transfer, and maintaining consistent characters across multiple images.

image-to-image
Generate images from text and a reference image using MiniMax Image-01 for consistent character appearance.
Minimax logo
minimax/image-01/subject-reference

Generate images from text and a reference image using MiniMax Image-01 for consistent character appearance.

stylized
transform
image-to-image
Moondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.
moondream3-preview/caption

Moondream 3 is a vision language model that brings frontier-level visual reasoning with native object detection, pointing, and OCR capabilities to real-world applications requiring fast, inexpensive inference at scale.

vision
vision
Modify a face to look younger or older while keeping identity realistic.
image-apps-v2/age-modify

Modify a face to look younger or older while keeping identity realistic.

age-transformation
face-editing
image-to-image
Luma Uni-1 Max generates a single image at the model's highest fidelity, delivering richer detail and stronger prompt adherence than the base tier for hero-quality stills.
Luma AI logo
luma/agent/uni-1/v1/max

Luma Uni-1 Max generates a single image at the model's highest fidelity, delivering richer detail and stronger prompt adherence than the base tier for hero-quality stills.

realism
typography
stylized
text-to-image
FireRed Image Edit v1.1 is an updated version of FireRed Image Edit, with improved image editing capabilities.
firered-image-edit-v1.1

FireRed Image Edit v1.1 is an updated version of FireRed Image Edit, with improved image editing capabilities.

firered-image-edit
image-to-image
Makes images more photorealistic and natural
Black Forest Labs logo
flux-2-lora-gallery/realism

Makes images more photorealistic and natural

stylized
transform
text-to-image
Generate synced sounds for any video, and return the new sound track (like MMAudio)
mirelo-ai/sfx-v1.5/video-to-audio

Generate synced sounds for any video, and return the new sound track (like MMAudio)

sfx
video-to-audio
Bria Extract Object uses text prompts to isolate a selected object from an image and return it as an RGBA PNG with a transparent background. Ideal for product, ecommerce, advertising, and creative editing workflows. Bria's Extract Object API leads in product shot extraction, outperforming SAM 3.1 where it counts most for commercial use.
Bria AI logo
bria/extract-object

Bria Extract Object uses text prompts to isolate a selected object from an image and return it as an RGBA PNG with a transparent background. Ideal for product, ecommerce, advertising, and creative editing workflows. Bria's Extract Object API leads in product shot extraction, outperforming SAM 3.1 where it counts most for commercial use.

image-to-image
Automatically retouches faces to smooth skin and remove blemishes.
retoucher

Automatically retouches faces to smooth skin and remove blemishes.

editing
image-to-image
ImagineArt 1.5 Pro is an advanced text-to-image model that creates ultra-high-fidelity 4K visuals with lifelike realism, refined aesthetics, and powerful creative output suited for professional use.
imagineart/imagineart-1.5-pro-preview/text-to-image

ImagineArt 1.5 Pro is an advanced text-to-image model that creates ultra-high-fidelity 4K visuals with lifelike realism, refined aesthetics, and powerful creative output suited for professional use.

visuals
imagineart
realism
text-to-image
Predict poses from images.
dwpose

Predict poses from images.

pose
utility
image-to-image
Flux Vision Upscaler for magnify/upscaling images with high fidelity and creativity.
Black Forest Labs logo
flux-vision-upscaler

Flux Vision Upscaler for magnify/upscaling images with high fidelity and creativity.

image-to-image
Generate videos from prompts using LTX Video-0.9.7 13B Distilled and custom LoRA
LTX logo
ltx-video-13b-distilled

Generate videos from prompts using LTX Video-0.9.7 13B Distilled and custom LoRA

video
ltx-video
text-to-video
Text-to-image generation with LoRA support for FLUX.2 [klein] 9B Base from Black Forest Labs. Custom style adaptation and fine-tuned model variations.
Black Forest Labs logo
flux-2/klein/9b/base/lora

Text-to-image generation with LoRA support for FLUX.2 [klein] 9B Base from Black Forest Labs. Custom style adaptation and fine-tuned model variations.

text-to-image
Wan 2.5 image-to-image model.
Alibaba logo
wan-25-preview/image-to-image

Wan 2.5 image-to-image model.

image-to-image
Stable Diffusion 3 Medium (Text to Image) is a Multimodal Diffusion Transformer (MMDiT) model that improves image quality, typography, prompt understanding, and efficiency.
stable-diffusion-v3-medium

Stable Diffusion 3 Medium (Text to Image) is a Multimodal Diffusion Transformer (MMDiT) model that improves image quality, typography, prompt understanding, and efficiency.

diffusion
style
text-to-image
Change hairstyles and hair colors in photos realistically.
image-apps-v2/hair-change

Change hairstyles and hair colors in photos realistically.

hair-edit
style-change
image-to-image
Showing 649 to 672 of 1491 results