Gemini 3.8 Flash Review: Full Benchmarks, Flash Speed & What It Really Costs

Gemini 3.8 Flash review cover art with a glowing lightning bolt racing past a benchmark data tower

Google has turned its “fast model” tier into something much harder to dismiss. Released on September 2, 2026, Gemini 3.8 Flash is positioned as the company’s new workhorse for coding, agents, multimodal tasks, and complex reasoning. It keeps the headline economics of Flash, with introductory API pricing of $0.75 per million input tokens and $3.75 … Read more

Stealing Reasoning: How Smaller AI Models Exposed Claude, GPT and Gemini’s Hidden Chain-of-Thought

Stealing Reasoning cover image showing hidden AI chain-of-thought leaking from a sealed data block to a smaller model.

AI companies have spent years making their best reasoning models harder to imitate. One increasingly important defense is simple in principle: don’t show users the model’s full chain of thought. A new paper, Stealing Reasoning Traces from Proprietary LLM APIs, found a surprisingly indirect way around that protection. Researchers discovered that encrypted reasoning blocks produced … Read more

Gemini 3.6 Flash: The 1M Context, Pricing, and Vision Breakthroughs You Missed

Gemini 3.6 Flash feature image showing a glowing data stream through a 1M token context window

Introduction Google just shipped Gemini 3.6 Flash, and the reaction split almost immediately along job description. Developers who went straight to the coding benchmarks shrugged. Enterprise teams running document pipelines and agent workflows did not. That gap tells you most of what you need to know about where Google is placing its bets with this … Read more