Upscale videos to 1080p, 2K, or 4K via API. FLUX 3 powered super-resolution with a precise mode and a creative detail-enhancement mode.
new
Black Forest Labs logo
blackforestlabs/flux-video-upscale

Upscale videos to 1080p, 2K, or 4K via API. FLUX 3 powered super-resolution with a precise mode and a creative detail-enhancement mode.

utility
editing
upscale
video-to-video
SAM 3D enables precise 3D reconstruction of objects from real images, while accurately reconstructing their geometry and texture.
sam-3/3d-objects

SAM 3D enables precise 3D reconstruction of objects from real images, while accurately reconstructing their geometry and texture.

3d
object
image-to-3d
 Run any audio capable LLM with fal. Process audio files — transcription, analysis, understanding, understand— using Gemini (Google) models. Supports wav, mp3, aiff, aac, ogg, flac, m4a. Powered by OpenRouter.
openrouter/router/audio

Run any audio capable LLM with fal. Process audio files — transcription, analysis, understanding, understand— using Gemini (Google) models. Supports wav, mp3, aiff, aac, ogg, flac, m4a. Powered by OpenRouter.

unknown
Use Gemini TTS Models to convert your prompts to real audio.
Google logo
gemini-tts

Use Gemini TTS Models to convert your prompts to real audio.

text-to-speech
audio
gemini
text-to-audio
Turn any flat ad image into fully editable layers —background, product and logo cutouts, live text with typography, and vector shapes. Commercial-safe, structured JSON output
new
Bria AI logo
bria/ad-delayer

Turn any flat ad image into fully editable layers —background, product and logo cutouts, live text with typography, and vector shapes. Commercial-safe, structured JSON output

utility
editing
image-to-json
Recraft V4 was developed with designers to bring true visual taste to AI image generation. Built for brand systems and production-ready workflows, it goes beyond prompt accuracy — delivering stronger composition, refined lighting, realistic materials, and a cohesive aesthetic. The result is imagery shaped by professional design judgment, ready for immediate real-world use without additional post-processing.
recraft/v4/pro/text-to-image

Recraft V4 was developed with designers to bring true visual taste to AI image generation. Built for brand systems and production-ready workflows, it goes beyond prompt accuracy — delivering stronger composition, refined lighting, realistic materials, and a cohesive aesthetic. The result is imagery shaped by professional design judgment, ready for immediate real-world use without additional post-processing.

text-to-image
MiniMax Hailuo-2.3 Image To Video API (Standard, 768p): Advanced image-to-video generation model with 768p resolution
Minimax logo
minimax/hailuo-2.3/standard/image-to-video

MiniMax Hailuo-2.3 Image To Video API (Standard, 768p): Advanced image-to-video generation model with 768p resolution

image-to-video
MAI-Image-2.5 is Microsoft's photorealistic image generation and editing model that turns text prompts or uploaded images into high-quality, design-ready visuals with fine-grained, pixel-level control.
microsoft/mai-image-2.5/edit

MAI-Image-2.5 is Microsoft's photorealistic image generation and editing model that turns text prompts or uploaded images into high-quality, design-ready visuals with fine-grained, pixel-level control.

realism
typography
stylized
image-to-image
Generate videos from prompts with audio using xAI's Grok Imagine 1.5 Video model.
xAI logo
xai/grok-imagine-video/v1.5/text-to-video

Generate videos from prompts with audio using xAI's Grok Imagine 1.5 Video model.

stylized
transform
lipsync
text-to-video
Generate video clips from your images using Kling 1.0
Kling logo
kling-video/v1/standard/image-to-video

Generate video clips from your images using Kling 1.0

motion
image-to-video
MiniMax Hailuo-02 Text To Video API (Standard, 768p): Advanced video generation model with 768p resolution
Minimax logo
minimax/hailuo-02/standard/text-to-video

MiniMax Hailuo-02 Text To Video API (Standard, 768p): Advanced video generation model with 768p resolution

text-to-video
Kling 2.1 Master: The premium endpoint for Kling 2.1, designed for top-tier image-to-video generation with unparalleled motion fluidity, cinematic visuals, and exceptional prompt precision.
Kling logo
kling-video/v2.1/master/image-to-video

Kling 2.1 Master: The premium endpoint for Kling 2.1, designed for top-tier image-to-video generation with unparalleled motion fluidity, cinematic visuals, and exceptional prompt precision.

image-to-video
Rig humanoid 3D models from GLB URLs with Meshy, returning rigged GLB/FBX files plus basic   animations.
meshy/rigging

Rig humanoid 3D models from GLB URLs with Meshy, returning rigged GLB/FBX files plus basic animations.

rigging
3d-to-3d
Generate high fidelity, studio quality videos of your avatar speaking or singing using the Aurora from Creatify team!
creatify/aurora

Generate high fidelity, studio quality videos of your avatar speaking or singing using the Aurora from Creatify team!

lipsync
image-to-video
Generate high-quality images from text with Krea 2 Medium, supporting aspect ratio, creativity controls, seeds, and optional style references.
Krea logo
krea/v2/medium/text-to-image

Generate high-quality images from text with Krea 2 Medium, supporting aspect ratio, creativity controls, seeds, and optional style references.

image-generation
style-reference
krea
text-to-image
FFMPEG Utilities to Scale Videos
workflow-utilities/scale-video

FFMPEG Utilities to Scale Videos

video-to-video
Happy Horse 1.1 is Alibaba's #1-ranked video model. This image-to-video endpoint animates a still image into 1080p video with synchronized native audio and multilingual lip-sync
Alibaba logo
alibaba/happy-horse/v1.1/image-to-video

Happy Horse 1.1 is Alibaba's #1-ranked video model. This image-to-video endpoint animates a still image into 1080p video with synchronized native audio and multilingual lip-sync

happy-horse
video
image
image-to-video
Image editing endpoint for Hunyuan Image 3.0 Instruct.
hunyuan-image/v3/instruct/edit

Image editing endpoint for Hunyuan Image 3.0 Instruct.

tencent
hunyuan-image
instruct
image-to-image
Wan 2.2's 5B model produces up to 5 seconds of video 720p at 24FPS with fluid motion and powerful prompt understanding
Alibaba logo
wan/v2.2-5b/image-to-video

Wan 2.2's 5B model produces up to 5 seconds of video 720p at 24FPS with fluid motion and powerful prompt understanding

image-to-video
Change the voices in your audios with voices in ElevenLabs!
ElevenLabs logo
elevenlabs/voice-changer

Change the voices in your audios with voices in ElevenLabs!

voice-change
audio-to-audio
Professional creative image upscaling powered by Topaz Labs. Bloom 2 reinvents detail with adjustable creativity and color preservation. Best for AI-generated images that need striking enhancement.
Topaz Labs logo
topaz/upscale/image/creative

Professional creative image upscaling powered by Topaz Labs. Bloom 2 reinvents detail with adjustable creativity and color preservation. Best for AI-generated images that need striking enhancement.

upscale
image
image-to-image
Rodin by Hyper3D generates realistic and production ready 3D models from text or images.
hyper3d/rodin

Rodin by Hyper3D generates realistic and production ready 3D models from text or images.

stylized
image-to-3d
LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.
LTX logo
ltx-2.3/text-to-video/fast

LTX-2.3 is a high-quality, fast AI video model available in Pro and Fast variants for text-to-video, image-to-video, and audio-to-video.

stylized
transform
lipsync
text-to-video
Meshy-6 is the latest model from Meshy. It generates realistic and production ready 3D models.
meshy/v6/image-to-3d

Meshy-6 is the latest model from Meshy. It generates realistic and production ready 3D models.

image-to-3d
Generate complete seamlessly tiling PBR materials including normal, roughness, basecolor, height and metalness maps up to 8K
patina/material

Generate complete seamlessly tiling PBR materials including normal, roughness, basecolor, height and metalness maps up to 8K

material
pbr
displacement
text-to-image
OmniHuman generates video using an image of a human figure paired with an audio file. It produces vivid, high-quality videos where the character’s emotions and movements maintain a strong correlation with the audio.
Bytedance logo
bytedance/omnihuman

OmniHuman generates video using an image of a human figure paired with an audio file. It produces vivid, high-quality videos where the character’s emotions and movements maintain a strong correlation with the audio.

lipsync
image-to-video
Generate consistent character appearances across multiple images. Maintain facial features, proportions, and distinctive traits for cohesive storytelling and branding
Ideogram logo
ideogram/character

Generate consistent character appearances across multiple images. Maintain facial features, proportions, and distinctive traits for cohesive storytelling and branding

character-consistency
image-to-image
Run SDXL at the speed of light
fast-lightning-sdxl

Run SDXL at the speed of light

diffusion
lightning
real-time
text-to-image
Showing 337 to 364 of 1504 results