Find a story
Search Spins
Search titles, summaries, and missing voices across published articles — press releases, announcements, and media coverage.
13 results for “steering”
Asymmetries in Spontaneous and Instructed Deception
A new arXiv preprint reports empirical evidence that spontaneous (uninstructed) and instructed deception in Llama-3.1-70B-Instruct share latent geometric structure in model representations, revealing asymmetric transferability between detection and steering across deception modes.
Sep 2, 2026
Recipes for Steering and Scaling LLMs via Sampling
A new arXiv preprint introduces a theoretically grounded sampling framework for steering and scaling LLMs—using Sequential Monte Carlo and Replica Exchange—to improve generation quality without external supervision or reward models.
Aug 28, 2026
Forecasting Side Effects of Activation Steering
Researchers propose a method to forecast unintended behavioral side effects of activation steering in language models before deployment, using a cross-effect matrix across 67 behaviors and three open-weight models.
Aug 13, 2026
Cross-Architecture Steering Transfer in Language Models: A Systematic Empirical Study
Researchers demonstrate that semantic concept directions learned in one large language model can be transferred to steer behavior in a different, independently trained LLM—provided both models meet a minimum scale threshold (~1.7B parameters) and architectural stability.
Aug 7, 2026
OpenAI introduces Agent Plugins, an open standard for bundling skills and MCP servers, and says its steering committee includes Amazon, Microsoft, and Vercel (Zac Hall/9to5Mac)
OpenAI announced Agent Plugins, an open standard for bundling AI agent skills and MCP servers, with a steering committee including Amazon, Microsoft, and Vercel, positioning itself as architect of next-generation AI infrastructure beyond single-model paradigms.
Aug 6, 2026
Getting rid of a non-working car in a financially smart way?
A Reddit user seeks advice on financially optimizing the disposal of a non-functional 2013 Mercedes ML350 with a costly power steering failure and minor cosmetic damage.
Aug 6, 2026
GCC steering committee announces AI policy
No substantive article content was provided — only a forum title and metadata indicating a Hacker News post titled 'GCC steering committee announces AI policy' with no body text, claims, or details.
Jul 31, 2026
Beyond a Global Norm: Personalizing Toxicity Sensitivity in Language Models Without Retraining
Researchers introduced a new framework for personalizing language model toxicity sensitivity without retraining, using inference-time interventions across pre-, in-, and post-decoding stages, revealing trade-offs between alignment accuracy, personalization, and language quality.
Jul 28, 2026
Probabilistic Concept-Aware Steering for Trustworthy LLM Inference
A new research paper introduces Probabilistic Concept-Aware Steering (PCS), a method to improve interpretability and fine-grained control in LLM inference by replacing binary steering evaluation with probabilistic, continuous semantic alignment.
Jul 22, 2026
Controlling Tool Use with Heading-Specific Activation Steering
Researchers propose a method to steer tool-augmented LLMs toward more selective tool invocation using heading-anchored steering vectors, demonstrating causal suppression across five open-source models—but find the underlying geometry is irregular and inconsistent with linear encoding assumptions.
Jul 9, 2026
Global Equity Crowdfunding Alliance Launches AI Governance Task Force
The Global Equity Crowdfunding Alliance (GECA), a UK-based advocacy group for online capital formation platforms, announced the formation of an AI Governance Task Force to coordinate industry dialogue on AI oversight.
Jul 9, 2026
Harnessing the Latent Space: From Steering Vectors to Model Calibrators for Control and Trust
Researchers propose methods to control and trust large language models.
Published Jul 2, 2026 · Analyzed Jul 5, 2026
Persona Without Substrate: Regime-Dependence and the LLM Individuation Problem
Researchers challenge a widely-held assumption in LLM individuation by presenting empirical evidence from persona-topology experiments.
Published Jul 2, 2026 · Analyzed Jul 5, 2026