Text-to-image generation with FLUX.2 [klein] 9B from Black Forest Labs and custom LoRA.
Black Forest Labs logo
flux-2/klein/9b/lora

Text-to-image generation with FLUX.2 [klein] 9B from Black Forest Labs and custom LoRA.

text-to-image
Run any LLM (Large Language Model) with fal, powered by OpenRouter.
openrouter/router/enterprise

Run any LLM (Large Language Model) with fal, powered by OpenRouter.

llm
Wan-2.1 is a text-to-video model that generates high-quality videos with high visual quality and motion diversity from text prompts
Alibaba logo
wan-t2v

Wan-2.1 is a text-to-video model that generates high-quality videos with high visual quality and motion diversity from text prompts

text to video
motion
text-to-video
VEED Fabric 1.0 is an image-to-video API that turns any image into a talking video
Veed logo
veed/fabric-1.0/fast

VEED Fabric 1.0 is an image-to-video API that turns any image into a talking video

lipsync
avatar
image-to-video
Generate character-consistent videos from reference images using PixVerse C1, with subject and background references.
Pixverse logo
pixverse/c1/reference-to-video

Generate character-consistent videos from reference images using PixVerse C1, with subject and background references.

video-generation
reference-to-video
pixverse
image-to-video
Prompt-free object removal from an image and mask, erasing objects with their shadows and reflections and reconstructing the scene cleanly.
Ideogram logo
ideogram/object-removal

Prompt-free object removal from an image and mask, erasing objects with their shadows and reflections and reconstructing the scene cleanly.

utility
editing
image-to-image
Extend Veo-Created Videos up to 30 seconds
Google logo
veo3.1/extend-video

Extend Veo-Created Videos up to 30 seconds

extend-video
video-to-video
Generate high-fidelity, design-ready images with precise typography, strong prompt alignment, and rich visual detail using Microsoft's flagship MAI Image 2.5 Pro.
microsoft/mai-image-2.5-pro

Generate high-fidelity, design-ready images with precise typography, strong prompt alignment, and rich visual detail using Microsoft's flagship MAI Image 2.5 Pro.

photorealism
typography
illustration
text-to-image
FLUX Control LoRA Depth is a high-performance endpoint that uses a control image to transfer structure to the generated image, using a depth map.
Black Forest Labs logo
flux-control-lora-depth

FLUX Control LoRA Depth is a high-performance endpoint that uses a control image to transfer structure to the generated image, using a depth map.

lora
style transfer
text-to-image
Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
florence-2-large/open-vocabulary-detection

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

multimodal
vision
detection
image-to-image
Pixal3D turns a single image into a high-fidelity 3D model with detailed geometry and realistic textures.
pixal3d

Pixal3D turns a single image into a high-fidelity 3D model with detailed geometry and realistic textures.

stylized
transform
image-to-3d
Enhances a given raster image using the 'creative upscale' tool, increasing image resolution, making the image sharper and cleaner.
recraft/upscale/creative

Enhances a given raster image using the 'creative upscale' tool, increasing image resolution, making the image sharper and cleaner.

upscaling
image-to-image
MuseTalk is a real-time high quality audio-driven lip-syncing model. Use MuseTalk to animate a face with your own audio.
musetalk

MuseTalk is a real-time high quality audio-driven lip-syncing model. Use MuseTalk to animate a face with your own audio.

animation
lip sync
real-time
image-to-video
Precise camera position and angle control (rotation, zoom, vertical movement)
Alibaba logo
qwen-image-edit-plus-lora-gallery/multiple-angles

Precise camera position and angle control (rotation, zoom, vertical movement)

stylized
transform
image-to-image
Generate 3D models from a single image using Tripo P1.
tripo3d/p1/image-to-3d

Generate 3D models from a single image using Tripo P1.

3d
3d-generation
tripo
image-to-3d
Audio-driven talking avatar generation powered by the SoulX-FlashTalk 14B model.
flashtalk

Audio-driven talking avatar generation powered by the SoulX-FlashTalk 14B model.

avatar
talking-head
audio-driven
audio-to-video
Generate videos from prompts and images using LTX Video-0.9.7 13B Distilled and custom LoRA
LTX logo
ltx-video-13b-distilled/image-to-video

Generate videos from prompts and images using LTX Video-0.9.7 13B Distilled and custom LoRA

video
ltx-video
image-to-video
Luma Uni-1 Max Edit applies text-guided edits to a source image at maximum fidelity, holding the original structure while honoring reference images for precise, high-detail revisions.
Luma AI logo
luma/agent/uni-1/v1/max/edit

Luma Uni-1 Max Edit applies text-guided edits to a source image at maximum fidelity, holding the original structure while honoring reference images for precise, high-detail revisions.

realism
typography
stylized
image-to-image
Edit an existing video using natural-language instructions, transforming subjects, settings, and style while retaining the original motion structure.
Kling logo
kling-video/o1/standard/video-to-video/edit

Edit an existing video using natural-language instructions, transforming subjects, settings, and style while retaining the original motion structure.

video-to-video
US hosted version of ByteDance's most advanced text-to-video model. Cinematic output with native audio, multi-shot editing, real-world physics, and director-level camera control.
new
Bytedance logo
bytedance/seedance-2.0/us/text-to-video

US hosted version of ByteDance's most advanced text-to-video model. Cinematic output with native audio, multi-shot editing, real-world physics, and director-level camera control.

stylized
transform
lipsync
text-to-video
Generate long videos from prompts and images using LTX Video-0.9.8 13B Distilled and custom LoRA
LTX logo
ltxv-13b-098-distilled/image-to-video

Generate long videos from prompts and images using LTX Video-0.9.8 13B Distilled and custom LoRA

video
ltx-video
image-to-video
Image2SVG transforms raster images into clean vector graphics, preserving visual quality while enabling scalable, customizable SVG outputs with precise control over detail levels.
image2svg

Image2SVG transforms raster images into clean vector graphics, preserving visual quality while enabling scalable, customizable SVG outputs with precise control over detail levels.

utility
editing
image-to-image
Image editing with HY-WU. Transfer outfits, swap faces, and blend textures instantly—no finetuning needed, just describe what you want and provide reference images.
hy-wu-edit

Image editing with HY-WU. Transfer outfits, swap faces, and blend textures instantly—no finetuning needed, just describe what you want and provide reference images.

image-to-image
Pixverse Effects
Pixverse logo
pixverse/v5.5/effects

Pixverse Effects

image-to-video
Removes mask-selected objects and their visual effects, seamlessly reconstructing the scene with contextually appropriate content.
object-removal/mask

Removes mask-selected objects and their visual effects, seamlessly reconstructing the scene with contextually appropriate content.

utility
editing
image-to-image
Happy Horse 1.1 is Alibaba's #1-ranked video model. This text-to-video endpoint generates 1080p video with synchronized native audio and multilingual lip-sync from a text prompt alone.
Alibaba logo
alibaba/happy-horse/v1.1/text-to-video

Happy Horse 1.1 is Alibaba's #1-ranked video model. This text-to-video endpoint generates 1080p video with synchronized native audio and multilingual lip-sync from a text prompt alone.

happy-horse
video
text
text-to-video
Recraft V4.1 Pro Vector generates large-format, fully editable SVGs with the structural clarity professional illustrators expect. Built for poster art, complex brand assets, and detailed scene illustration, it scales without losing geometric integrity.
recraft/v4.1/pro/text-to-vector

Recraft V4.1 Pro Vector generates large-format, fully editable SVGs with the structural clarity professional illustrators expect. Built for poster art, complex brand assets, and detailed scene illustration, it scales without losing geometric integrity.

stylized
transform
typography
text-to-image
Wan-2.1 flf2v generates dynamic videos by intelligently bridging a given first frame to a desired end frame through smooth, coherent motion sequences.
Alibaba logo
wan-flf2v

Wan-2.1 flf2v generates dynamic videos by intelligently bridging a given first frame to a desired end frame through smooth, coherent motion sequences.

image to video
motion
image-to-video
Showing 561 to 588 of 1491 results