Remember when Google released its Gemini 3 AI model last year and it seemed like it was going to force itself into the conversation among the frontier labs with major benchmarking achievements and major fanfare around its image generation model? That all faded mighty fast, huh? Well, Google is back to remind everyone that it is competing—though it’s not exactly blowing anyone away. On Tuesday, the company introduced Gemini 3.6 Flash, the “workhorse” version of the company’s flagship model that is supposed to provide the best balance of quality and efficiency. And it does seem to eat up fewer tokens, which is currently a meaningful benchmark in the post-tokenamxxing era. Per Google, Gemini 3.6 Flash cuts output token by as much as 65% compared to 3.5 Flash in some uses, and is uses 17% fewer output tokens overall than the previous model. That’s noteworthy as people become more aware of the costs associated with AI use, but the performance of Gemini 3.6 Flash isn’t going to blow anyone away. It appears to be behind Anthropic’s Claude Sonnet 5 and OpenAI’s GPT-5.6 in most major benchmarking tests, and it’s even started to lag behind the latest Grok model, 4.5, in tasks like agentic coding (Grok can probably thank the Cursor team for the boost there, but the score is the score). Considering that Gemini 3.6 Flash isn’t that much cheaper than most of its competitors—it’s about in line with Grok 4.5 and GPT-5.6—it’s going to have a hard task carving out a niche in the model war.