
google/gemini-omni-flash/v1.1/reference-to-video
Gemini Omni Flash 1.1 is Google's multimodal video model. This endpoint generates video from combined multimodal references, images, videos and text together. Reasoning across all inputs to produce a single coherent result, with characters retaining their face, clothing, and voice throughout
Inference
Commercial use
Partner
Prompt examples
Examples are generated using the Gemini Omni Flash 1.1 Reference to Video. You can customize them by clicking on the "Playground" button.
