Find a story

Search Spins

Search titles, summaries, and missing voices across published articles — press releases, announcements, and media coverage.

3 results for “inference costs”

SPIN Processed News Frame: The Hype

Google's "Frozen v2" chip reportedly bakes Gemini's architecture directly into silicon for efficiency gains

Google is reportedly developing a custom server chip called 'Frozen v2' that hardcodes Gemini's architecture into silicon, aiming for 6–10× efficiency gains over current TPUs by 2028 to reduce inference costs and gain competitive pricing leverage.

Spin 85% Claim Present in Source AI Risk High
The Decoder

Jul 21, 2026

SPIN Processed News Frame: The Cushion

OpenAI Halves Inference Costs With Software Alone: GPUs Drop to Hundreds - Tech Times

OpenAI claims to have reduced AI inference costs by 50% using only software optimizations, enabling deployment on cheaper hardware like sub-$1,000 GPUs.

Spin 81% Claim Present in Source AI Risk High
Google News: OpenAI

Published Jul 3, 2026 · Analyzed Jul 6, 2026

SPIN Processed News Frame: The Hype

OpenAI Discovers New Way to Cut Inference Costs in Half - The Information

OpenAI claims to have developed a novel method that reduces AI inference costs by 50%, potentially improving model deployment economics and scalability.

Spin 85% Needs Evidence AI Risk High
The Information AI via Google News

Published Jun 30, 2026 · Analyzed Jul 4, 2026