Key Info
Google DeepMind says the new Gemini 3.8 Flash Cyber model excels at autonomously finding weaknesses on security benchmarks like CyberGym while remaining fast and efficient. In real-world testing on Google Chrome codebases, it produced 2.6 times more valid fixes, helping protect software faster.
Highlights
- Leads on CyberGym and similar benchmarks for autonomous vulnerability discovery, with strong speed and efficiency.
- In tests across Google Chrome codebases, the model generated 2.6x more valid fixes than the baseline.
- The results point to practical use of AI for faster, more scalable software security and patching.