Find a story

Search Spins

Search titles, summaries, and missing voices across published articles — press releases, announcements, and media coverage.

13 results for “steering”

SPIN Processed News Frame: The Hype

Asymmetries in Spontaneous and Instructed Deception

A new arXiv preprint reports empirical evidence that spontaneous (uninstructed) and instructed deception in Llama-3.1-70B-Instruct share latent geometric structure in model representations, revealing asymmetric transferability between detection and steering across deception modes.

Spin 65% Claim Present in Source AI Risk Moderate
arXiv Artificial Intelligence

Sep 2, 2026

SPIN Processed News Frame: The Hype

Recipes for Steering and Scaling LLMs via Sampling

A new arXiv preprint introduces a theoretically grounded sampling framework for steering and scaling LLMs—using Sequential Monte Carlo and Replica Exchange—to improve generation quality without external supervision or reward models.

Spin 45% Needs Evidence AI Risk Moderate
arXiv Computation and Language

Aug 28, 2026

SPIN Processed News Frame: The Halo

Forecasting Side Effects of Activation Steering

Researchers propose a method to forecast unintended behavioral side effects of activation steering in language models before deployment, using a cross-effect matrix across 67 behaviors and three open-weight models.

Spin 65% Claim Present in Source AI Risk Moderate
arXiv Artificial Intelligence

Aug 13, 2026

SPIN Processed News Frame: The Hype

Cross-Architecture Steering Transfer in Language Models: A Systematic Empirical Study

Researchers demonstrate that semantic concept directions learned in one large language model can be transferred to steer behavior in a different, independently trained LLM—provided both models meet a minimum scale threshold (~1.7B parameters) and architectural stability.

Spin 48% Claim Present in Source AI Risk Moderate
arXiv Computation and Language

Aug 7, 2026

SPIN Processed News Frame: The Hype

OpenAI introduces Agent Plugins, an open standard for bundling skills and MCP servers, and says its steering committee includes Amazon, Microsoft, and Vercel (Zac Hall/9to5Mac)

OpenAI announced Agent Plugins, an open standard for bundling AI agent skills and MCP servers, with a steering committee including Amazon, Microsoft, and Vercel, positioning itself as architect of next-generation AI infrastructure beyond single-model paradigms.

Spin 85% Claim Present in Source AI Risk High
Techmeme

Aug 6, 2026

SPIN Processed News Frame: none

Getting rid of a non-working car in a financially smart way?

A Reddit user seeks advice on financially optimizing the disposal of a non-functional 2013 Mercedes ML350 with a costly power steering failure and minor cosmetic damage.

Spin 0% Claim Present in Source
Reddit r/personalfinance

Aug 6, 2026

SPIN Processed News Frame: The Fog

GCC steering committee announces AI policy

No substantive article content was provided — only a forum title and metadata indicating a Hacker News post titled 'GCC steering committee announces AI policy' with no body text, claims, or details.

Spin 0% Needs Evidence
Hacker News Front Page

Jul 31, 2026

SPIN Processed News Frame: The Hype

Beyond a Global Norm: Personalizing Toxicity Sensitivity in Language Models Without Retraining

Researchers introduced a new framework for personalizing language model toxicity sensitivity without retraining, using inference-time interventions across pre-, in-, and post-decoding stages, revealing trade-offs between alignment accuracy, personalization, and language quality.

Spin 60% Claim Present in Source AI Risk Moderate
arXiv Computation and Language

Jul 28, 2026

SPIN Processed News Frame: The Hype

Probabilistic Concept-Aware Steering for Trustworthy LLM Inference

A new research paper introduces Probabilistic Concept-Aware Steering (PCS), a method to improve interpretability and fine-grained control in LLM inference by replacing binary steering evaluation with probabilistic, continuous semantic alignment.

Spin 65% Claim Present in Source AI Risk Moderate
arXiv Artificial Intelligence

Jul 22, 2026

SPIN Processed News Frame: The Hype

Controlling Tool Use with Heading-Specific Activation Steering

Researchers propose a method to steer tool-augmented LLMs toward more selective tool invocation using heading-anchored steering vectors, demonstrating causal suppression across five open-source models—but find the underlying geometry is irregular and inconsistent with linear encoding assumptions.

Spin 65% Claim Present in Source AI Risk Moderate
arXiv Artificial Intelligence

Jul 9, 2026

SPIN Processed News Frame: The Halo

Global Equity Crowdfunding Alliance Launches AI Governance Task Force

The Global Equity Crowdfunding Alliance (GECA), a UK-based advocacy group for online capital formation platforms, announced the formation of an AI Governance Task Force to coordinate industry dialogue on AI oversight.

Spin 65% Claim Present in Source AI Risk Moderate
Crowdfund Insider

Jul 9, 2026

SPIN Processed News Frame: The Hype

Harnessing the Latent Space: From Steering Vectors to Model Calibrators for Control and Trust

Researchers propose methods to control and trust large language models.

Spin 70% Claim Present in Source AI Risk Moderate
arXiv Computation and Language

Published Jul 2, 2026 · Analyzed Jul 5, 2026

SPIN Processed News Frame: The Hype

Persona Without Substrate: Regime-Dependence and the LLM Individuation Problem

Researchers challenge a widely-held assumption in LLM individuation by presenting empirical evidence from persona-topology experiments.

Spin 50% Claim Present in Source AI Risk Moderate
arXiv Computation and Language

Published Jul 2, 2026 · Analyzed Jul 5, 2026