
Train a LoRA that generates the lead-in to an audio clip, extending audio backward in time from its ending.

Train a LoRA that regenerates masked time spans of an audio clip while keeping the rest unchanged.

Train a LoRA for a joint audio+video transformation, conditioned on a reference clip (its video and audio) to produce a matching target clip.

Train a LoRA that regenerates a masked video region (guided by kept pixels and a video reference) while jointly generating audio from an audio reference.

Train a LoRA that continues a video forward in time — supply an opening clip at inference and the model generates what comes next.

Train a LoRA that generates the lead-in to a video, extending a clip backward in time from its ending.

Train an IC-LoRA that transforms one audio clip into another, conditioned at inference on a reference audio clip.

Train an IC-LoRA for a joint audio+video transformation, conditioned on a reference clip's video and audio to produce a matching target.

Train an IC-LoRA that regenerates a masked video region (guided by kept pixels and a video reference) while jointly generating audio from an audio reference.

Train an IC-LoRA that regenerates only the masked region of a video, guided by the kept pixels and a separate reference/control video.

Train a LoRA that regenerates a masked region of a video while keeping the rest unchanged, blending the new content with its surroundings.

Train a LoRA that generates the video between keyframes — supply first/last (and optional middle) frames at inference and the model fills the in-between motion.

Train a LoRA that expands the video frame outward, keeping an inner rectangle fixed and generating the surrounding region.

Train a LoRA that generates audio (foley / sound design) for a silent video, learning a soundtrack that matches the on-screen action.

Train a LoRA that learns a video-to-video transformation from paired before/after clips, steered at inference by a reference (control) video.

Train a LoRA that regenerates only the masked region of a video, guided by both the kept pixels and a separate reference/control video.

Turn photos into mind-blowing, dynamic videos. Your images can can come to life with sharp details, impressive character control and cinematic camera moves.

Start with a simple text input to create dynamic generations that defy expectations. Anything you dream can come to life with sharp details, impressive character control and cinematic camera moves.

Pika v2 Turbo creates videos from a text prompt with high quality output.

Create chromatic aberration by shifting red, green, and blue channels horizontally or vertically with customizable shift amounts.

Apply dodge and burn effects with multiple modes and adjustable intensity.

Apply a parabolic distortion effect with configurable coefficient and vertex position.

Apply solarization effect by inverting pixel values above a threshold

Leverage the rapid processing capabilities of AI models to enable accurate and efficient real-time speech-to-text transcription.