topic
LLM
1 dispatch on LLM.
Fewer Tokens, Lower Price, Ships Now: Google's New Inference Playbook
Google's Gemini 3.6 Flash launched with a 17% output token reduction and a price cut on outputs. It signals that the real AI model competition has moved from headline benchmarks to cost per task.