Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

Google drops new Gemini models, but the comments are asking: is anyone actually impressed

TLDR: Google unveiled faster and cheaper Gemini AI models, with 3.6 Flash positioned as the main upgrade and Flash-Lite as the budget speed option. But commenters were unconvinced, questioning the confusing numbers, weak comparisons, and whether Google’s new toys actually beat the competition.

Google rolled out a fresh trio of AI models — Gemini 3.6 Flash, 3.5 Flash-Lite, and a security-focused Cyber version — promising they’re faster, cheaper, and better for the kind of behind-the-scenes tasks companies use to build AI helpers. On paper, 3.6 Flash is the big star: Google says it uses fewer words to get the job done, costs less, and performs better on coding, document reading, and other work tasks. Flash-Lite is being pitched as the speed demon of the family, while the Cyber model is aimed at security teams.

But the real fireworks were in the comments, where the launch got a collective eyebrow raise instead of a standing ovation. One of the loudest reactions was basically: cool story, but why would I use this? Critics said the benchmark scores didn’t feel exciting, especially because Google mostly compared the new models to its own older models instead of rivals. That drew some blunt side-eye, with one commenter accusing the company of being deeply inward-looking and acting like outside competition barely exists.

Then came the math drama. A commenter pounced on Google’s numbers, asking how the model could be “up to 65%” better in one place but only show “49% vs. 37%” elsewhere. Translation for non-nerds: the crowd smelled possible marketing spin and immediately started fact-checking the fine print. Another hot take went straight for the jugular, claiming the new model is less smart, pricier, and closed off compared to a Chinese competitor. In other words, Google wanted applause for a shiny upgrade — and the internet responded with receipts, sarcasm, and a group chat full of skepticism.

Key Points

  • Google introduced Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and Gemini 3.5 Flash Cyber in CodeMender for production AI agent workloads.
  • Google says Gemini 3.6 Flash uses 17% fewer output tokens than 3.5 Flash on the Artificial Analysis Index, with lower pricing of $1.50 per 1M input tokens and $7.50 per 1M output tokens.
  • The article reports 3.6 Flash benchmark gains over 3.5 Flash in DeepSWE, MLE Bench, OSWorld-Verified, and GDPval-AA v2.
  • Gemini 3.5 Flash-Lite is positioned for low-latency and high-throughput tasks, with Artificial Analysis measuring it at 350 output tokens per second and pricing of $0.30 input / $2.50 output per 1M tokens.
  • Google says Gemini 3.5 Pro is in partner testing and that pre-training for Gemini 4 has already begun.

Hottest takes

"not clear why would I use it now" — yanis_t
"So which one is it? 65% or 49%?" — dumberquestions
"less intelligent and more expensive than GLM-5.2" — jgbuddy
Made with <3 by @siedrix and @shesho from CDMX. Powered by Forge&Hive.