topic

inference

1 dispatch on inference.

Fewer Tokens, Lower Price, Ships Now: Google's New Inference Playbook

Google's Gemini 3.6 Flash launched with a 17% output token reduction and a price cut on outputs. It signals that the real AI model competition has moved from headline benchmarks to cost per task.

AI Google Gemini