Find a story
Search Spins
Search titles, summaries, and missing voices across published articles — press releases, announcements, and media coverage.
7 results for “token usage”
Something is wrong with Codex since July 22
Users report a persistent, undocumented regression in OpenAI's Codex API causing infinite loops and doubled token consumption since July 22, linked to version 0.146 and the 'wait_agent_enabled' configuration setting.
Aug 6, 2026
Presentation: The Five Stages of AI Maturity in Engineering Organizations - Where and Why Teams Get Stuck
Quotient CEO Lizzie Matusov introduces a five-stage AI maturity framework for engineering organizations to diagnose why high AI investment isn’t translating into improved software delivery, emphasizing organizational alignment and outcome-based metrics over token-centric vanity metrics.
Aug 4, 2026
A look at OpenAI's open-source agent harness that now powers Codex and ChatGPT Work, as the company works to optimize the harness to cut runaway token usage (Jason Hiner/The Deep View)
OpenAI has developed and open-sourced an agent harness that now underpins Codex and ChatGPT Work, with internal optimization efforts underway to reduce excessive token consumption.
Jul 29, 2026
Any guide for optimising token usage in codex/work?
A Reddit user on the r/OpenAI forum expresses frustration with token limits on their OpenAI Plus subscription while attempting AI-assisted CAD work, highlighting a real constraint in practical usage.
Jul 27, 2026
Chinese AI models account for ~60% of token usage by US companies on OpenRouter, making restrictions harder to impose without disrupting US users and businesses (Nectar Gan/Bloomberg)
Chinese AI models now constitute approximately 60% of token usage by US companies on the OpenRouter API marketplace, creating practical constraints on US export controls or regulatory restrictions due to dependency risk.
Jul 22, 2026
Do we really need all the information that Frontier models give us?
A GitHub-hosted open-source tool called 'Sir Shortoken' claims to reduce LLM token consumption by filtering inputs to core concepts without lossy compression, tested on technical domains like APIs and Kubernetes.
Jul 19, 2026
10 Billion token usage in 20 days, (orange is claude) (Grey is Gpt)
A Reddit user shared an unattributed, unlabeled chart showing token usage metrics for Claude and GPT models over 20 days, with no source, methodology, or verification provided.
Jul 14, 2026