5s Living Oil Painting | 16:9 | Impressionist Post-Painterly Texture | Sea-Cliff Wind and Scattering Blossoms FIRST FRAME: strictly preserve the reference image's original composition, palette, and brushwork. Nothing is re-rendered as photoreal — the animation must look like an oil painting in motion, visible impasto strokes and palette-knife ridges holding their identity as they move. SUBJECT: young dark-haired woman in a sage-olive long-sleeved blouse with a patterned sash, head tilted down and to her right, eyes lowered, expression calm and inward. Both bare forearms lift toward a great bundle of white chrysanthemum blooms breaking apart in the air before her. Face structure, hair mass, sleeve folds, and skin tone stay consistent throughout.
moreAI Video Models Production-ready video generation APIs
Build with the latest AI video models on fal. Generate production-ready video from text, images, and references through fast, scalable APIs.
Choose the right video model for your workflow
Compare current video generation models from leading providers, selected from the active fal catalog.

fal's H3 Max is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality

fal's H3 Max is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality

Wan 3.0 Prime Text-to-Video transforms written prompts into polished videos with accelerated generation, fluid motion, strong scene fidelity, and coherent visual storytelling. Built for fast creative iteration, it brings complex ideas to life while preserving visual detail and cinematic consistency throughout each shot.

Wan 3.0 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

Generate high quality 1080p videos using Kling's Turbo 3.0 model, with improved lipsync and multishot generation capabilities.

Kling's Native 4K is a video generation model that directly outputs professional-grade 4K video in one step, eliminating the need for post-production upscaling

Generate film-grade videos from text prompts with native audio, up to 1080p and 15 seconds, using PixVerse C1.

Veo 3.1 by Google, the most advanced AI video generation model in the world. With sound on!

Wan 3.0 is the latest generation AI video model, delivering enhanced motion smoothness, superior scene fidelity, and greater visual coherence.

The Avatar X API offers access to Mirage's most advanced generation model yet, delivering industry-leading identity preservation and expressivity in AI video

Generate high quality 1080p videos from images using Kling's Turbo 3.0 model, with improved lipsync and multishot generation capabilities.

FLUX 3 is Black Forest Labs' frontier video model. This endpoint builds video from a sequence of keyframes, generating the motion between each anchor point for precise control over how a shot progresses.

Animate images into cinematic videos with PixVerse C1, supporting 1080p resolution and native audio generation.

Veo 3.1 is the latest state-of-the art video generation model from Google DeepMind
From an idea to a production video
Use the playground for a quick test or call the same model from your application through the API.
See what the latest AI video models can create
Explore generations from the featured models, then open an example to inspect its inputs and run it yourself.
A distinguished man in a tailored charcoal three-piece suit, crisp white shirt, and silk tie stands behind a polished mahogany desk in a warmly lit studio. He holds an artisanal glass paperweight delicately between the fingertips of both hands, presenting it forward at chest height like a master jeweler revealing a rare gem. His posture is confident and composed, a subtle knowing smile on his face as his eyes remain fixed on the sphere.The camera opens on a medium shot framing the man from the waist up, the paperweight centered between his hands. It then begins a slow, smooth 360-degree orbit around him, circling from his right side around to his back and returning to the front, keeping the paperweight locked in the center of the frame. As the orbit completes, the man gently rotates the sphere between his thumbs and forefingers, turning it toward the camera.The camera then initiates a deliberate push-in, gliding past his hands and zooming steadily toward the glass orb until it fills the entire frame. The studio's track lighting catches in the paperweight's crystalline depths, refracting into prismatic patterns that dance across the polished mahogany desk below. Within its perfectly spherical form, suspended bubbles and swirls of cobalt blue and amber form a miniature cosmos that shifts with every subtle movement of his hands.In the final macro close-up, the camera drifts through the glass itself, exploring the frozen bubbles and color ribbons like a submersible navigating an alien ocean, with soft bokeh highlights from the studio lights blooming in the background and the suited man's silhouette softly out of focus behind.Cinematic luxury product showcase aesthetic, warm key light with cool rim light, shallow depth of field, photorealistic, 24fps film look, elegant and refined atmosphere
moreTight close-up portrait of a 20 years old Gen Z fashion model, early twenties, dewy glass-skin complexion with a light dusting of faux freckles, glossy bitten-lip stain in a muted berry tone, fluffy laminated brows, and a single pearl-studded graphic liner flick in chrome silver across one eyelid. Her hair is slicked back into a sleek low bun with two soft face-framing tendrils, small silver butterfly clips catching the light. She wears oversized vintage-style chrome chandelier earrings and a sheer mesh high-neck top layered under a cropped leather moto jacket. The camera holds an extreme close-up on her face, then slowly arcs around her in a subtle 45-degree orbit as she tilts her chin down, cuts her eyes directly into the lens with a confident deadpan stare, and exhales softly. A gentle breeze lifts the loose strands of hair across her cheekbone. Lighting is soft neon-tinged — cool lavender key light from the left blending into a warm peach rim light from the right, creating a duotone gradient across her skin. Background is an out-of-focus wash of deep magenta and teal bokeh. Cinematic editorial fashion aesthetic, shallow depth of field, 85mm lens compression, subtle film grain, photorealistic, 24fps, reminiscent of a modern i-D Magazine or Vogue Beauty film.
moreHe explodes upward out of the black. Both arms drive toward the falling apple, body twisting, white linen sleeves snapping taut then bunching as his shoulders roll. The apple tumbles fast, end over end, its stem whipping. He mistimes it — the fruit glances off his fingertips and spins away; he lunges after it, throwing his weight sideways, hair swinging across his face, eyes flashing wide. He catches it hard against his palm, the impact shoving his arm back and down, linen collapsing in loose folds. He pulls it in against his chest, breathing hard, and breaks into a grin as he looks straight into the lens. The candlelight jumps with every movement, flaring across his forearms and guttering into darkness behind him.
moreA lighthouse beam sweeps through heavy rain over black rocks at night, waves detonating against the cliff.
moreA lighthouse beam sweeps through heavy rain over black rocks at night, waves detonating against the cliff.
moreCreate a 5-second live-action Hollywood film shot in a bright university library. A thoughtful professor stands in a medium shot, turns slightly toward an unseen student, and calmly says, “What is the real?” Natural daylight, professional cinema camera, realistic performance, subtle camera drift, no close-up, no dark lighting, perfect lip sync.
morethe camera hangs back and ascends to a high angle. As a police car speeds fowards with it's lights on entering the frame. The camera finishes at rear tracking shot.
moreTight close-up portrait of a Gen Z fashion model, early twenties, dewy glass-skin complexion with a light dusting of faux freckles, glossy bitten-lip stain in a muted berry tone, fluffy laminated brows, and a single pearl-studded graphic liner flick in chrome silver across one eyelid. Her hair is slicked back into a sleek low bun with two soft face-framing tendrils, small silver butterfly clips catching the light. She wears oversized vintage-style chrome chandelier earrings and a sheer mesh high-neck top layered under a cropped leather moto jacket. The camera holds an extreme close-up on her face, then slowly arcs around her in a subtle 45-degree orbit as she tilts her chin down, cuts her eyes directly into the lens with a confident deadpan stare, and exhales softly. A gentle breeze lifts the loose strands of hair across her cheekbone. Lighting is soft neon-tinged — cool lavender key light from the left blending into a warm peach rim light from the right, creating a duotone gradient across her skin. Background is an out-of-focus wash of deep magenta and teal bokeh. Cinematic editorial fashion aesthetic, shallow depth of field, 85mm lens compression, subtle film grain, photorealistic, 24fps, reminiscent of a modern i-D Magazine or Vogue Beauty film.
moreUltra-high-velocity anime action sequence, masterpiece quality. A tall woman knight with short pale-gold hair and calm luminous eyes, wearing weightless white-and-silver plate armor with a long drifting translucent cape, wielding a single ethereal longsword of pure white light. She fights alone against thirty enemy knights of a rival order — real warriors in charcoal-grey lacquered plate with crimson underlayers, tattered war-banners on their backs, bare faces and half-visors, armed with spears, twin curved blades, banner-polearms, greatswords, warhammers and lightning lances. They fight in disciplined formations, shouting orders and flanking her across a chain of shattered marble temple fragments floating high above an endless sea of clouds at dusk. She never stops moving — sprinting, sliding, vaulting, wall-running, leaping between islands, defeating each knight in one or two strikes without breaking stride, armor shearing, sparks bursting, banners cut, bodies falling away into the clouds behind her. Rain falls upward in the wind, aurora ribbons shimmer overhead, wet marble glows, volumetric god rays cut through the cloud sea. Relentless forward momentum, elegant lethal choreography, dynamic orbital camera chained to her shoulder, whip-pans, speed lines, energy waves, radiant sparks on every clash, marble debris and glowing dust drifting in slow arcs, dramatic close-ups on her calm luminous eyes and on the fear in theirs, cinematic lighting, anime movie quality, highly detailed backgrounds, ultra smooth fluid animation, breathtaking final attack
moreThe model strides forward, the sculptural gown moving with her, camera flashes popping in the dark crowd, reflection sliding on the floor. The camera tracks backward ahead of her. Photorealistic, dramatic, glamorous.
moreA colossal, ancient library with impossibly high shelves, where books fly and pages turn on their own. Audio: the rustle of thousands of pages turning, the soft whoosh of flying books, distant, echoing whispers, a grand, magical orchestral piece with swirling harps and enchanted woodwinds.
moreWhat can you build with AI video models?
Generate and transform video without managing model infrastructure.
Turn concepts into moving scenes
Generate cinematic shots, storyboards, and visual sequences from scripts, prompts, and reference frames.
Produce campaign video faster
Create social clips, product videos, and localized campaign variations without managing model infrastructure.
Build video generation into your product
Add text-to-video, image animation, reference-guided generation, and video utilities to automated workflows.
Prepare generated video for production
Trim, resize, blend, and reverse video with hosted utility endpoints that compose with your generation workflow.

FFMPEG Utility for Trim Video

FFMPEG Utilities to Scale Videos

FFMPEG Utility for Blending Videos

FFMPEG Utility to Reverse Videos
Common questions about AI video models
Which AI video model should I choose?
Choose based on your input, output quality, speed, duration, audio, and control requirements. Open a featured model to compare its schema, pricing, and examples before integrating it.
Can these models generate video from text and images?
Yes. fal hosts text-to-video, image-to-video, reference-to-video, and keyframe-to-video endpoints. The supported inputs and controls vary by model.
Can I use AI video models through an API?
Yes. Every featured model has a hosted fal API and an interactive playground. Open a model to review its request schema, pricing, and code examples, or create an API key from your dashboard.
Do AI-generated videos include audio?
Some models generate synchronized audio natively, while others return silent video. Check the selected endpoint's schema and model description for its audio capabilities.
What video utilities are available?
fal provides hosted utilities for common post-processing tasks including trimming, scaling, blending, and reversing video. They can be composed with generation endpoints in a production workflow.