AI Model Timeline

Tracking the accelerating release frequency of frontier AI models.

Gemini 3.5 Flash Cyber

Google·

LLMsspecializedcloud

Overview

Security-specialized model fine-tuned from Gemini 3.5 Flash to identify, validate and patch software vulnerabilities at low cost per token. Deployed inside Google's CodeMender security agent, which calls it many times in rapid succession to explore alternative execution paths across a codebase. Availability is limited: an initial pilot for governments and trusted partners via CodeMender, not general API access. Google reports internal results of 55 confirmed unique vulnerabilities on Chrome's V8 JavaScript engine (vs 47 for standard 3.5 Flash and 36 for Claude Opus 4.6), a 42% improvement on long-range multi-turn cyber benchmarks over the Flash 3 predecessor, and frontier-competitive performance on CyberGym within CodeMender. No standard academic benchmarks were reported.

Capabilities and innovations

Vulnerability detection, validation and patchingRuns inside the CodeMender security agentLimited pilot: governments and trusted partners onlyCybersecurity-specialized fine-tune of Gemini 3.5 Flash for automated vulnerability discovery and patching

Benchmarks

No benchmark values are recorded for this model.

  • GPQA Diamond: not reported by the vendor
  • Humanity’s Last Exam (no tools): not reported by the vendor
  • SWE-bench Pro: not reported by the vendor
  • MMLU-Pro: not reported by the vendor
  • Terminal-Bench 2.x: not reported by the vendor
  • MMMU-Pro: not reported by the vendor

Links

More from Google

Data curated by AI Model Timeline. See the methodology for admission criteria, benchmark eras and source priority.