GOOGLE · DEEPMIND

Google's newest AI model makes fast, cheap coding help even cheaper

Google released Gemini 3.7 Flash, a faster model built for coding and everyday AI-agent tasks — just three weeks after its predecessor, at introductory pricing about half of the previous model's cost per token.

Open in the XNewsAi app →

Why this matters

For the 806

The fast, cheap tier of AI is where most everyday work actually gets done — quotes, code fixes, customer replies. When it gets faster and cheaper this quickly, the tools built on top of it get better and less expensive too.

What we know

  • Gemini 3.7 Flash shipped August 13, 2026, three weeks after Gemini 3.6 Flash.
  • Google's own benchmarks show substantial gains in coding and agentic-workflow tasks over the prior Flash model.
  • Introductory API pricing is about half of the prior Flash model's per-token cost, through the end of 2026.

What we don't know

  • Whether these gains hold up in real-world, non-benchmark use the way Google's own numbers suggest.
  • When Google's larger flagship model, Gemini 3.5 Pro, will actually ship — it remains behind competitors' release pace.

Sources

  • Primary
    Google — Gemini models blog ↗

    Google's own announcement of Gemini 3.7 Flash, including pricing and benchmark claims.

  • Reporting
    Axios ↗

    Independent coverage of the release timing and competitive context against the delayed Gemini 3.5 Pro.

Related stories

Open in the XNewsAi app →