Gemini 3.8 Flash Cyber
Google·
Overview
Cybersecurity-specialised variant of Gemini 3.8 Flash and successor to Gemini 3.5 Flash Cyber, built to discover, validate and patch software vulnerabilities autonomously across codebases in 20 programming languages. Not a public model: access runs through Google's new Fairwind Program, which pairs the model with the CodeMender harness for government agencies and national cyber authorities, critical-infrastructure operators (healthcare, telecom, energy, finance), maintainers of widely used software platforms and vetted security partners, under operational requirements such as multi-factor authentication and access limited to internal security staff. Google says it will not be generally released. Vendor-reported, none of it a tracked academic benchmark: CyberGym (autonomous vulnerability discovery) above 3.5 Flash Cyber and larger frontier models, no figure published; a success rate above 70% on an internal 20-language vulnerability-discovery benchmark; CWE-Bench (Collinear) patching at 47.2% pass@1; and 2.6x more correct Chrome patches than the best commercial models in the Chrome Security team's own evaluation. No standard academic benchmarks were reported.
Capabilities and innovations
Benchmarks
No benchmark values are recorded for this model.
- GPQA Diamond: not reported by the vendor
- Humanity’s Last Exam (no tools): not reported by the vendor
- SWE-bench Pro: not reported by the vendor
- MMLU-Pro: not reported by the vendor
- Terminal-Bench 2.x: not reported by the vendor
- MMMU-Pro: not reported by the vendor
Links
More from Google
Data curated by AI Model Timeline. See the methodology for admission criteria, benchmark eras and source priority.