Creates a reusable style from your reference images for use with Recraft V4 Styles Pro generation
new
recraft/v4/pro/create-style

Creates a reusable style from your reference images for use with Recraft V4 Styles Pro generation

stylized
transform
editing
training
Stable Cascade: Image generation on a smaller & cheaper latent space.
stable-cascade

Stable Cascade: Image generation on a smaller & cheaper latent space.

diffusion
lcm
text-to-image
Generate video with audio from audio, text and images using LTX-2 Distilled
LTX logo
ltx-2-19b/distilled/audio-to-video

Generate video with audio from audio, text and images using LTX-2 Distilled

audio-to-video
Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
florence-2-large/dense-region-caption

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

multimodal
vision
image-to-image
Wan-2.1 Pro is a premium text-to-video model that generates high-quality 1080p videos at 30fps with up to 6 seconds duration, delivering exceptional visual quality and motion diversity from text prompts
Alibaba logo
wan-pro/text-to-video

Wan-2.1 Pro is a premium text-to-video model that generates high-quality 1080p videos at 30fps with up to 6 seconds duration, delivering exceptional visual quality and motion diversity from text prompts

text to video
motion
text-to-video
Place your subject in any scene you imagine, from enchanted forests to urban settings, with professional composition and lighting
image-editing/scene-composition

Place your subject in any scene you imagine, from enchanted forests to urban settings, with professional composition and lighting

stylized
transform
image-to-image
Use USO to perform subject driven generations using reference image.
uso

Use USO to perform subject driven generations using reference image.

image-to-image
Run inference on LoRA adapters for TRELLIS.2 model
trellis-2-lora

Run inference on LoRA adapters for TRELLIS.2 model

image-to-3d
Juggernaut Base Flux by RunDiffusion is a drop-in replacement for Flux [Dev] that delivers sharper details, richer colors, and enhanced realism, while instantly boosting LoRAs and LyCORIS with full compatibility.
rundiffusion-fal/juggernaut-flux/base

Juggernaut Base Flux by RunDiffusion is a drop-in replacement for Flux [Dev] that delivers sharper details, richer colors, and enhanced realism, while instantly boosting LoRAs and LyCORIS with full compatibility.

image generation
text-to-image
MultiTalk model generates a talking avatar video from an image and text. Converts text to speech automatically, then generates the avatar speaking with lip-sync.
ai-avatar/single-text

MultiTalk model generates a talking avatar video from an image and text. Converts text to speech automatically, then generates the avatar speaking with lip-sync.

stylized
transform
image-to-video
Fast, low-latency text-to-image model with high-quality output and full JSON-structured controllability. Open-source, trained on licensed data, and optimized for production-scale generation.
Bria AI logo
bria/fibo-lite/generate

Fast, low-latency text-to-image model with high-quality output and full JSON-structured controllability. Open-source, trained on licensed data, and optimized for production-scale generation.

bria
fibo
lite
text-to-image
Train a MiniMax H3 LoRA with first-frame conditioning, so a still image animates into video with audio; captions optional.
Minimax logo
minimax/h3/i2v/trainer

Train a MiniMax H3 LoRA with first-frame conditioning, so a still image animates into video with audio; captions optional.

utility
editing
training
Structured Prompt Generation endpoint for Fibo, Bria's SOTA Open source model.
Bria AI logo
bria/fibo/generate/structured_prompt

Structured Prompt Generation endpoint for Fibo, Bria's SOTA Open source model.

bria
fibo
structured-prompting
text-to-json
Generate video with audio from text using LTX-2.3 Distilled
LTX logo
ltx-2.3-22b/distilled/text-to-video

Generate video with audio from text using LTX-2.3 Distilled

text-to-video
Any pose, any style, any identity
omni-zero

Any pose, any style, any identity

style transfer
image-to-image
M-LSD line segment detection preprocessor.
image-preprocessors/mlsd

M-LSD line segment detection preprocessor.

preprocess
utility
controlnet
image-to-image
LoRA inference endpoint for the Qwen Image Editing model.
Alibaba logo
qwen-image-edit-lora

LoRA inference endpoint for the Qwen Image Editing model.

image-editing
lora
image-to-image
Generate videos from images and prompts using CogVideoX-5B
cogvideox-5b/image-to-video

Generate videos from images and prompts using CogVideoX-5B

image-to-video
Train a MiniMax H3 LoRA on your own captioned clips for pure text-to-video generation with matching audio.
Minimax logo
minimax/h3/t2v/trainer

Train a MiniMax H3 LoRA on your own captioned clips for pure text-to-video generation with matching audio.

utility
editing
training
FLUX.1 Krea [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.
Black Forest Labs logo
flux-1/krea/image-to-image

FLUX.1 Krea [dev] is a 12 billion parameter flow transformer that generates high-quality images from text with incredible aesthetics. It is suitable for personal and commercial use.

image-to-image
Texture an existing geometry mesh using a reference image with Hi3D.
hitem3d/hi3d/texture

Texture an existing geometry mesh using a reference image with Hi3D.

3d-to-3d
Stable Audio 3 Small Music audio outpainting is a 459 million parameter latent diffusion model that extends music compositions beyond their original endpoint via causal continuation.
stable-audio-3/small/music/audio-outpainting

Stable Audio 3 Small Music audio outpainting is a 459 million parameter latent diffusion model that extends music compositions beyond their original endpoint via causal continuation.

music
extension
continuation
audio-to-audio
Accelerated image generation with Ideogram V2 Turbo. Create high-quality visuals, posters, and logos with enhanced speed while maintaining Ideogram's signature quality.
Ideogram logo
ideogram/v2/turbo

Accelerated image generation with Ideogram V2 Turbo. Create high-quality visuals, posters, and logos with enhanced speed while maintaining Ideogram's signature quality.

realism
typography
text-to-image
Train custom LoRAs for personalization, styles or other use cases on top of Ideogram V4.
Ideogram logo
ideogram/v4/trainer

Train custom LoRAs for personalization, styles or other use cases on top of Ideogram V4.

training
Showing 1129 to 1152 of 1496 results