Gemini Omni 1.1 Flash
Google·
Overview
Successor to Gemini Omni Flash (2026-06-30) for video generation and editing, released as the model id gemini-omni-1.1-flash. Takes any combination of text, image, audio and video as input and returns video. Scene extension now conditions on up to 10 seconds of preceding footage instead of the tail of the clip and continues in 10-second increments up to 40 seconds total; a first-and-last-frame control renders a single continuous shot between two supplied frames, and a video reference of up to 3 seconds carries character and style across generations. Native output resolution is 720p, with 1080p and 4K delivered as upscales, plus a new 360p draft mode that Google reports as up to 60% faster at a third of the 720p cost. Per-second API pricing: $0.03 (360p), $0.10 (720p), $0.15 (1080p), $0.30 (4K). Live via the Gemini API, Google AI Studio, Flow and the Gemini Enterprise Agent Platform; scene extension is rolling out to Google AI Plus, Pro and Ultra subscribers in the Gemini app globally. The predecessor's gemini-omni-flash-preview endpoint is scheduled for deprecation on 2026-09-30.
Capabilities and innovations
Benchmarks
No tracked academic benchmark applies to video models; the timeline carries no scores and no leadership tracking for them.
- GPQA Diamond: not reported by the vendor
- Humanity’s Last Exam (no tools): not reported by the vendor
- SWE-bench Verified: not reported by the vendor
- SWE-bench Pro: not reported by the vendor
- MMLU: not reported by the vendor
- MMLU-Pro: not reported by the vendor
- Terminal-Bench 2.x: not reported by the vendor
- MMMU-Pro: not reported by the vendor
Links
More from Google
Data curated by AI Model Timeline. See the methodology for admission criteria, benchmark eras and source priority.