topic

LLM

1 dispatch on LLM.

Fewer Tokens, Lower Price, Ships Now: Google's New Inference Playbook

Google's Gemini 3.6 Flash launched with a 17% output token reduction and a price cut on outputs. It signals that the real AI model competition has moved from headline benchmarks to cost per task.

AI Google Gemini