AI Model Timeline

Tracking the accelerating release frequency of frontier AI models.

Gemini 3.8 Flash Cyber

Google·

LLMsspecializedcloud

Overview

Cybersecurity-specialised variant of Gemini 3.8 Flash and successor to Gemini 3.5 Flash Cyber, built to discover, validate and patch software vulnerabilities autonomously across codebases in 20 programming languages. Not a public model: access runs through Google's new Fairwind Program, which pairs the model with the CodeMender harness for government agencies and national cyber authorities, critical-infrastructure operators (healthcare, telecom, energy, finance), maintainers of widely used software platforms and vetted security partners, under operational requirements such as multi-factor authentication and access limited to internal security staff. Google says it will not be generally released. Vendor-reported, none of it a tracked academic benchmark: CyberGym (autonomous vulnerability discovery) above 3.5 Flash Cyber and larger frontier models, no figure published; a success rate above 70% on an internal 20-language vulnerability-discovery benchmark; CWE-Bench (Collinear) patching at 47.2% pass@1; and 2.6x more correct Chrome patches than the best commercial models in the Chrome Security team's own evaluation. No standard academic benchmarks were reported.

Capabilities and innovations

Vulnerability discovery, validation and patching across 20 programming languagesRuns inside the CodeMender harnessRestricted access: Fairwind Program only (governments, critical infrastructure, core software maintainers, vetted security partners)Cybersecurity-specialised variant of Gemini 3.8 Flash; successor to Gemini 3.5 Flash CyberFirst model distributed through Google's Fairwind Program, a gated access scheme with operational requirements instead of a general release

Benchmarks

No benchmark values are recorded for this model.

  • GPQA Diamond: not reported by the vendor
  • Humanity’s Last Exam (no tools): not reported by the vendor
  • SWE-bench Pro: not reported by the vendor
  • MMLU-Pro: not reported by the vendor
  • Terminal-Bench 2.x: not reported by the vendor
  • MMMU-Pro: not reported by the vendor

Links

More from Google

Data curated by AI Model Timeline. See the methodology for admission criteria, benchmark eras and source priority.