Run Any Stable Diffusion model with customizable LoRA weights.
lora

Run Any Stable Diffusion model with customizable LoRA weights.

diffusion
lora
customization
text-to-image
Generate text embeddings using OpenAI-compatible API. Access embedding models like text-embedding-3-small, text-embedding-3-large (OpenAI), and other embedding models available through OpenRouter. Drop-in replacement for the OpenAI embeddings API. Powered by OpenRouter.
openrouter/router/openai/v1/embeddings

Generate text embeddings using OpenAI-compatible API. Access embedding models like text-embedding-3-small, text-embedding-3-large (OpenAI), and other embedding models available through OpenRouter. Drop-in replacement for the OpenAI embeddings API. Powered by OpenRouter.

llm
Generate videos from prompts using LTX Video
LTX logo
ltx-video

Generate videos from prompts using LTX Video

text-to-video
SAM 3.1 builds comes with Object Multiplex, a shared-memory approach for joint multi-object tracking that delivers faster speeds with larger number of objects tracked.
sam-3-1/image-rle

SAM 3.1 builds comes with Object Multiplex, a shared-memory approach for joint multi-object tracking that delivers faster speeds with larger number of objects tracked.

segmentation
mask
real-time
image-to-image
Generate images with transparent backgrounds using Ideogram Transparent model
Ideogram logo
ideogram/v3/generate-transparent

Generate images with transparent backgrounds using Ideogram Transparent model

stylized
transform
typography
text-to-image
FLUX 3 is Black Forest Labs' frontier video model. This endpoint builds video from a sequence of keyframes, generating the motion between each anchor point for precise control over how a shot progresses.
Black Forest Labs logo
blackforestlabs/flux-3/keyframes-to-video

FLUX 3 is Black Forest Labs' frontier video model. This endpoint builds video from a sequence of keyframes, generating the motion between each anchor point for precise control over how a shot progresses.

stylized
transform
lipsync
image-to-video
Generate high quality video clips from text and image prompts using PixVerse v5.5
Pixverse logo
pixverse/v5.5/image-to-video

Generate high quality video clips from text and image prompts using PixVerse v5.5

image-to-video
Extend videos with xAI's Grok Imagine video model
xAI logo
xai/grok-imagine-video/extend-video

Extend videos with xAI's Grok Imagine video model

video-edit
v2v
grok
video-to-video
Vision reasoning variant of NVIDIA's Nemotron 3 Nano Omni. 30B A3B hybrid Transformer-Mamba MoE - accepts an image plus a prompt and returns text.
nvidia/nemotron-3-nano-omni/vision

Vision reasoning variant of NVIDIA's Nemotron 3 Nano Omni. 30B A3B hybrid Transformer-Mamba MoE - accepts an image plus a prompt and returns text.

nemotron
nvidia
vision-language
image-to-text
Generate video clips from your prompts using MiniMax model
Minimax logo
minimax/video-01

Generate video clips from your prompts using MiniMax model

motion
transformation
text-to-video
Generate natural multilingual speech from text with fast voice and language control using Qwen Audio 3.0 TTS Flash.
Alibaba logo
alibaba/qwen-audio-3-tts

Generate natural multilingual speech from text with fast voice and language control using Qwen Audio 3.0 TTS Flash.

audio
speech-synthesis
multilingual
text-to-speech
Generate dubbed videos or audios using ElevenLabs Dubbing feature!
ElevenLabs logo
elevenlabs/dubbing

Generate dubbed videos or audios using ElevenLabs Dubbing feature!

dubbing
audio-to-audio
audio-to-video
References into video with synchronized audio using MiniMax H3
Minimax logo
minimax/h3/reference-to-video/lora

References into video with synchronized audio using MiniMax H3

utility
editing
video-to-video
Extract seamless tiling textures with PBR attribute maps from images
patina/material/extract

Extract seamless tiling textures with PBR attribute maps from images

material
pbr
extraction
image-to-image
FFMPEG Untility for Extracting nth Frame
workflow-utilities/extract-nth-frame

FFMPEG Untility for Extracting nth Frame

image-to-image
Start with a simple text input to create dynamic generations that defy expectations in up to 1080p. Experience better image clarity and crisper, sharper visuals.
pika/v2.2/text-to-video

Start with a simple text input to create dynamic generations that defy expectations in up to 1080p. Experience better image clarity and crisper, sharper visuals.

editing
effects
animation
text-to-video
Vidu's Q3 Turbo Model
vidu/q3/image-to-video/turbo

Vidu's Q3 Turbo Model

image-to-video
Image-to-image editing with FLUX.2 [klein] 4B Base from Black Forest Labs. Precise modifications using natural language descriptions and hex color control.
Black Forest Labs logo
flux-2/klein/4b/base/edit

Image-to-image editing with FLUX.2 [klein] 4B Base from Black Forest Labs. Precise modifications using natural language descriptions and hex color control.

image-to-image
Generate high-fidelity images extremely fast from text with Krea 2 Medium Turbo, supporting aspect ratio, creativity, seed controls, and optional style references.
Krea logo
krea/v2/medium/turbo/text-to-image

Generate high-fidelity images extremely fast from text with Krea 2 Medium Turbo, supporting aspect ratio, creativity, seed controls, and optional style references.

stylized
transform
typography
text-to-image
Create seamless transition between images using PixVerse v5
Pixverse logo
pixverse/v5/transition

Create seamless transition between images using PixVerse v5

stylized
transform
image-to-video
Generate premium-quality images from text prompts using the enhanced WAN 2.7 Pro model with superior detail and composition.
Alibaba logo
wan/v2.7/pro/text-to-image

Generate premium-quality images from text prompts using the enhanced WAN 2.7 Pro model with superior detail and composition.

wan
pro
text-to-image
Modify consistent characters while preserving their core identity. Edit poses, expressions, or clothing without losing recognizable character features
Ideogram logo
ideogram/character/edit

Modify consistent characters while preserving their core identity. Edit poses, expressions, or clothing without losing recognizable character features

character-consistency
image-to-image
Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
florence-2-large/ocr-with-region

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

ocr
multimodal
vision
image-to-image
The OpenRouter Responses API with fal, powered by OpenRouter, provides unified access to a wide range of large language models - including GPT, Claude, Gemini, and many others through a single API interface.
openrouter/router/openai/v1/responses

The OpenRouter Responses API with fal, powered by OpenRouter, provides unified access to a wide range of large language models - including GPT, Claude, Gemini, and many others through a single API interface.

llm
Pixelcut's Video Background Remover is an AI segmentation model that erases backgrounds frame by frame, with seamless temporal consistency.
pixelcut/video-background-removal

Pixelcut's Video Background Remover is an AI segmentation model that erases backgrounds frame by frame, with seamless temporal consistency.

transform
utility
rembg
video-to-video
FLUX.1 [dev] is a 12 billion parameter flow transformer that generates high-quality images from text. It is suitable for personal and commercial use.
Black Forest Labs logo
flux-1/dev

FLUX.1 [dev] is a 12 billion parameter flow transformer that generates high-quality images from text. It is suitable for personal and commercial use.

text-to-image
Place any product in any scenery with just a prompt or reference image while maintaining high integrity of the product. Trained exclusively on licensed data for safe and risk-free commercial use and optimized for eCommerce.
Bria AI logo
bria/product-shot

Place any product in any scenery with just a prompt or reference image while maintaining high integrity of the product. Trained exclusively on licensed data for safe and risk-free commercial use and optimized for eCommerce.

product photography
image-to-image
VOID removes objects from videos along with all interactions they induce on the scene
void-video-inpainting

VOID removes objects from videos along with all interactions they induce on the scene

utility
editing
video-to-video
Showing 505 to 532 of 1491 results