topic

compute

2 dispatches on compute.

Fewer Tokens, Lower Price, Ships Now: Google's New Inference Playbook

Google's Gemini 3.6 Flash launched with a 17% output token reduction and a price cut on outputs. It signals that the real AI model competition has moved from headline benchmarks to cost per task.

AI Google Gemini

Apple's Siri Runs on Google Now. That's Not a Bug — It's the Compute Crisis Made Visible.

Apple's WWDC Siri AI reveal and Google's deal with SpaceX arrived days apart. Together, they expose the structural reality of the AI era: even the biggest device makers and hyperscalers can't build fast enough alone.

Apple AI Siri