Gemini 3.8 Flash and 3.8 Flash Cyber(blog.google)
1154 points by bratao 8 days ago | 662 comments
tl;dr: Google released Gemini 3.8 Flash, a reasoning/coding model priced at $0.75/$3.75 per million input/output tokens, claiming performance approaching larger frontier models on benchmarks like DeepSWE, HLE-Verified, and legal/finance agent tasks—though it uses more tokens per task than 3.7 Flash. A specialized variant, Gemini 3.8 Flash Cyber, targets vulnerability discovery and automated patching, reportedly outperforming larger models on CyberGym and CWE-Bench, and is restricted to vetted defenders via Google's new Fairwind Program. Google says it's already using it internally, including finding a critical Cloud vulnerability in under two hours.
HN Discussion:
  • Impressed by speed and coding/HTML generation capabilities at low cost
  • Gemini excels at real-world knowledge, multimodal tasks, and document parsing
  • Benchmark performance is remarkable for a Flash-tier model, rivaling frontier models
  • Rapid release cadence and incremental improvements strategy is paying off
  • ~Concern that Google hasn't successfully pretrained a new base model since January 2025