Lucy-5B is a model that can create 5-second I2V videos in under 5 seconds, achieving >1x RTF end-to-end
Decart logo
decart/lucy-5b/image-to-video

Lucy-5B is a model that can create 5-second I2V videos in under 5 seconds, achieving >1x RTF end-to-end

image-to-video
Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks
florence-2-large/region-to-category

Florence-2 is an advanced vision foundation model that uses a prompt-based approach to handle a wide range of vision and vision-language tasks

multimodal
vision
Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.
Black Forest Labs logo
flux-krea-lora/stream

Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

lora
personalization
text-to-image
Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.
Black Forest Labs logo
flux-lora/stream

Super fast endpoint for the FLUX.1 [dev] model with LoRA support, enabling rapid and high-quality image generation using pre-trained LoRA adaptations for personalization, specific styles, brand identities, and product-specific outputs.

lora
personalization
text-to-image
Train Ideogram on your photos, your style, your subject, your look, from a small set of reference images to images that feel consistently yours
Ideogram logo
ideogram/custom-models

Train Ideogram on your photos, your style, your subject, your look, from a small set of reference images to images that feel consistently yours

stylized
transform
training
Extend videos with audio using LTX-2 Distilled and custom LoRA
LTX logo
ltx-2-19b/distilled/extend-video/lora

Extend videos with audio using LTX-2 Distilled and custom LoRA

video-to-video
Generate video with audio from videos using LTX-2 Distilled
LTX logo
ltx-2-19b/distilled/video-to-video

Generate video with audio from videos using LTX-2 Distilled

video-to-video
Generate video with audio from videos using LTX-2 Distilled and custom LoRA
LTX logo
ltx-2-19b/distilled/video-to-video/lora

Generate video with audio from videos using LTX-2 Distilled and custom LoRA

video-to-video
Extend video with audio using LTX-2 and custom LoRA
LTX logo
ltx-2-19b/extend-video/lora

Extend video with audio using LTX-2 and custom LoRA

video-to-video
Generate video with audio from videos using LTX-2 and custom LoRA
LTX logo
ltx-2-19b/video-to-video/lora

Generate video with audio from videos using LTX-2 and custom LoRA

video-to-video
Cross-eyes for high-quality video using LTX-2.3
LTX logo
ltx-2.3-quality/cross-eyed

Cross-eyes for high-quality video using LTX-2.3

eyes
video-to-video
Train a LoRA that transforms one audio clip into another, learning a reference→target mapping from paired audio examples.
LTX logo
ltx23-trainer-v2/a2a

Train a LoRA that transforms one audio clip into another, learning a reference→target mapping from paired audio examples.

training
Train a LoRA that continues an audio clip forward in time, generating the audio that follows a short clean prefix.
LTX logo
ltx23-trainer-v2/audio-extend-prefix

Train a LoRA that continues an audio clip forward in time, generating the audio that follows a short clean prefix.

training
Train a LoRA that generates the lead-in to an audio clip, extending audio backward in time from its ending.
LTX logo
ltx23-trainer-v2/audio-extend-suffix

Train a LoRA that generates the lead-in to an audio clip, extending audio backward in time from its ending.

training
Train a LoRA that regenerates masked time spans of an audio clip while keeping the rest unchanged.
LTX logo
ltx23-trainer-v2/audio-inpaint

Train a LoRA that regenerates masked time spans of an audio clip while keeping the rest unchanged.

training
Train a LoRA for a joint audio+video transformation, conditioned on a reference clip (its video and audio) to produce a matching target clip.
LTX logo
ltx23-trainer-v2/av2av

Train a LoRA for a joint audio+video transformation, conditioned on a reference clip (its video and audio) to produce a matching target clip.

training
Train a LoRA that regenerates a masked video region (guided by kept pixels and a video reference) while jointly generating audio from an audio reference.
LTX logo
ltx23-trainer-v2/av2av-masked

Train a LoRA that regenerates a masked video region (guided by kept pixels and a video reference) while jointly generating audio from an audio reference.

training
Train a LoRA that continues a video forward in time — supply an opening clip at inference and the model generates what comes next.
LTX logo
ltx23-trainer-v2/extend-prefix

Train a LoRA that continues a video forward in time — supply an opening clip at inference and the model generates what comes next.

training
Train a LoRA that generates the lead-in to a video, extending a clip backward in time from its ending.
LTX logo
ltx23-trainer-v2/extend-suffix

Train a LoRA that generates the lead-in to a video, extending a clip backward in time from its ending.

training
Train an IC-LoRA that transforms one audio clip into another, conditioned at inference on a reference audio clip.
LTX logo
ltx23-trainer-v2/ic-lora/a2a

Train an IC-LoRA that transforms one audio clip into another, conditioned at inference on a reference audio clip.

training
Train an IC-LoRA for a joint audio+video transformation, conditioned on a reference clip's video and audio to produce a matching target.
LTX logo
ltx23-trainer-v2/ic-lora/av2av

Train an IC-LoRA for a joint audio+video transformation, conditioned on a reference clip's video and audio to produce a matching target.

training
Train an IC-LoRA that regenerates a masked video region (guided by kept pixels and a video reference) while jointly generating audio from an audio reference.
LTX logo
ltx23-trainer-v2/ic-lora/av2av-masked

Train an IC-LoRA that regenerates a masked video region (guided by kept pixels and a video reference) while jointly generating audio from an audio reference.

training
Train an IC-LoRA that regenerates only the masked region of a video, guided by the kept pixels and a separate reference/control video.
LTX logo
ltx23-trainer-v2/ic-lora/v2v-masked

Train an IC-LoRA that regenerates only the masked region of a video, guided by the kept pixels and a separate reference/control video.

training
Train a LoRA that generates the video between keyframes — supply first/last (and optional middle) frames at inference and the model fills the in-between motion.
LTX logo
ltx23-trainer-v2/interpolate

Train a LoRA that generates the video between keyframes — supply first/last (and optional middle) frames at inference and the model fills the in-between motion.

training
Train a LoRA that regenerates only the masked region of a video, guided by both the kept pixels and a separate reference/control video.
LTX logo
ltx23-trainer-v2/v2v-masked

Train a LoRA that regenerates only the masked region of a video, guided by both the kept pixels and a separate reference/control video.

training
PersonaPlex is a real-time, full-duplex speech-to-speech conversational model that enables persona control through text-based role prompts and audio-based voice conditioning.
personaplex/realtime

PersonaPlex is a real-time, full-duplex speech-to-speech conversational model that enables persona control through text-based role prompts and audio-based voice conditioning.

realtime
conversational
audio-to-audio
Turn photos into mind-blowing, dynamic videos. Your images can can come to life with sharp details, impressive character control and cinematic camera moves.
pika/v2.1/image-to-video

Turn photos into mind-blowing, dynamic videos. Your images can can come to life with sharp details, impressive character control and cinematic camera moves.

editing
effects
animation
image-to-video
Turbo is the model to use when you feel the need for speed. Turn your image to stunning video up to 3x faster – all with high quality outputs.
pika/v2/turbo/image-to-video

Turbo is the model to use when you feel the need for speed. Turn your image to stunning video up to 3x faster – all with high quality outputs.

editing
effects
animation
image-to-video
Showing 1429 to 1456 of 1472 results