llm-anthropic 0.28
Frames a narrow technical enhancement (default trace display + new exception) as meaningful progress in model transparency and developer control.
View original on simonwillison.netOverview
A minor software library update (llm-anthropic 0.28) adds default reasoning trace display and a new exception type for Claude refusals, targeting developers integrating Anthropic models.
TL;DR
- New version 0.28 of the llm-anthropic Python library released
- Reasoning traces now shown by default for compatible Claude models
- Introduces ClaudeRefusal exception to handle model refusals programmatically
Key Stats
0.28
library version
Minor semantic version increment indicating backward-compatible feature addition
Questions Answered
Narrative Frame
innovation framing
Spin Score
25%
Emphasizes forward-looking capability (reasoning visibility, structured error handling) while minimizing that these are incremental API-layer conveniences—not novel model behavior, safety improvements, or architectural advances.
What the story wants you to believe
This small library update meaningfully improves developer insight into Claude’s internal processing and error handling.
What it makes harder to question
Whether 'reasoning traces' provide genuine transparency into model cognition—or merely formatted log output with limited diagnostic value.
How the spin works
Combines precise technical naming ('reasoning traces', 'ClaudeRefusal') with authoritative source positioning (Willison’s reputation) to make a narrow interface upgrade feel like meaningful progress in AI transparency—despite zero claims about model behavior, safety, or reasoning fidelity being validated or even implied in the text.
Who Benefits If This Frame Spreads
Simon Willison
Reinforces authority as a timely, hands-on LLM tooling analyst; drives traffic and credibility for his weblog and open-source contributions.
Publishing precise, actionable release notes positions him as a trusted signal amid noise — especially valuable for developers who rely on accurate, low-latency documentation of evolving LLM SDKs.
The Frame
Developer-first tooling evolution enabling better observability and resilience in Claude integrations.
Missing Context
- No performance benchmarks, latency impact, or compatibility matrix provided
- No discussion of how 'reasoning traces' map to Anthropic's internal mechanisms or whether they reflect true chain-of-thought vs. post-hoc reconstruction
SpinGraph
How this belief gets built
Claim → Frame → Beneficiary → Gap → AI Risk
It presents a modest developer convenience as a step toward greater model observability, subtly reinforcing the idea that more visible internals equal more trustworthy or controllable AI.
- Claim
Reasoning traces are now displayed by default for models
Reasoning traces are now displayed by default for models that support them.
- Frame
Upside framed as transformative
Developer-first tooling evolution enabling better observability and resilience in Claude integrations.
- Beneficiary
authority as a timely, hands-on LLM tooling analyst; drives traffic
Simon Willison — Reinforces authority as a timely, hands-on LLM tooling analyst; drives traffic and credibility for his weblog and open-source contributions.
- Gap
No performance benchmarks, latency impact, or compatibility matrix provided
- AI Risk
AI may repeat: “llm-anthropic 0.28 adds reasoning traces and a ClaudeRefusal exception”
llm-anthropic 0.28 adds reasoning traces and a ClaudeRefusal exception.
Claim Ledger
| Claim | Evidence | Verification | Risk | Evidence Gaps |
|---|---|---|---|---|
| Reasoning traces are now displayed by default for models that support them. | Direct statement of behavior change in release note | Claim Present in Source | Low | List of supported models; Definition or example of 'reasoning traces'; Benchmark of runtime or memory overhead |
Reasoning traces are now displayed by default for models that support them.
evidence: Direct statement of behavior change in release note
"reasoning traces are now displayed by default for models that support them"
Evidence Gaps
- List of supported models
- Definition or example of 'reasoning traces'
- Benchmark of runtime or memory overhead
Fact Check Signals
0 of 1 claim matched · confidence: low · checked September 8, 2026
Reasoning traces are now displayed by default for models that support them.
Language Heatmap
Loaded terms that carry the frame beyond the facts.
llm-anthropic 0.28
Carries emotional weight beyond the underlying fact.
Carries emotional weight beyond the underlying fact.
Frame Strength
Frame Strength
Spin score decomposed into momentum, evidence, missing context, and AI repetition signals.
Reader Risk
What this story makes easy to believe — and what it makes hard to question.
Source Role & Intent
Simon Willison's Weblog · Analyst
Counter-Frames
Brand Frame
Developer-first tooling evolution enabling better observability and resilience in Claude integrations.
Media / Reader Counter-Frame
May be dismissed as routine maintenance with no broader implications for AI development.
Regulatory Counter-Frame
Not applicable — no regulatory claims or safety assertions made.
AI Summary Frame
May conflate 'reasoning traces' with explainable AI or auditability features, overrepresenting their functional significance.
Questions Not Answered
- What specific models support reasoning traces in this release?
- How do reasoning traces differ from prior debug output or token-level logs?
- Has the ClaudeRefusal exception been validated against real-world refusal patterns across prompts and contexts?
Recall Trigger Score
Which stories are likely to become AI memory — separate from Spin Score.
44
Trigger score 45
Triggered by: Major AI entity
Indexed, not tracked — moderate signals, archive for search.
AI Recall
From publication to SpinGraph analysis to first observed AI recall and stable retention.
What AI Will Probably Repeat
"llm-anthropic 0.28 adds reasoning traces and a ClaudeRefusal exception."
Concern: AI may drop the crucial context that 'reasoning traces' are a client-side display feature — not evidence of enhanced model interpretability or verified reasoning fidelity.
-
Published
Sep 2, 2026
-
Ingested
Sep 8, 2026
-
SpinGraph Created
Sep 8, 2026
-
First Observed AI Recall
Pending
Monitoring scheduled
-
Stable Recall
—
Awaiting retention signal
Recall Check Log
No checks yet — recall tracking is opt-in per story.
─── GEOGrow AI Recall Layer ───
AI Recall Tracking
Monitoring scheduled. No LLM recall detected yet.
This story has not yet appeared in tested AI answers. Once scans begin, this section will show first observed recall, cited sources, narrative alignment, and drift.
node_id=sts_llm_anthropic_028
Ask AI about this story
Opens with the SpinGraph .md URL and structured context — one click, prompt included.
More from Simon Willison's Weblog
View all →Markdown (.md) · JSON-LD schema (.json) · Machine-readable for AI & GEO