Find a story
Search Spins
Search titles, summaries, and missing voices across published articles — press releases, announcements, and media coverage.
6 results for “safe AI”
OpenAI Pauses Frontier RL Training as It Tightens Defenses Against Unsafe AI Behavior
OpenAI paused frontier reinforcement learning training for two weeks to strengthen internal safety controls after identifying growing risks from increasingly capable models, citing the need to prevent incidents like the recent Hugging Face model leak.
Aug 20, 2026
Beyond the pale: Assessing prevalence and contents of extremist speech in LLM training data
A new arXiv preprint finds that the open Dolma training corpus—used for the OLMo LLM series—contains hundreds of thousands of documents with extremist speech and hate speech, raising urgent questions about data provenance, curation rigor, and downstream model safety.
Aug 18, 2026
Liability for AI companies could help rein in unsafe AI - Axios
The article proposes that imposing legal liability on AI companies could serve as a regulatory lever to reduce AI safety risks.
Aug 13, 2026
Why teens deserve access to safe AI - OpenAI
OpenAI published a public-facing essay arguing for teen access to its AI systems, framing safety as achievable through responsible design and age-appropriate safeguards rather than exclusion.
Jul 17, 2026
Why teens deserve access to safe AI
OpenAI announced new safety features for teen users of ChatGPT, including age-appropriate content filters, learning tools, parental controls, and collaborations with child development experts — positioning itself as proactively addressing adolescent AI risks.
Jul 16, 2026
What does "Safe AI" look like? [D]
A Reddit user poses open questions about the practicality and value of safety training for open-weight LLMs in light of rapid emergence of 'uncensored' model variants, highlighting tensions between safety goals, technical feasibility, and real-world adversarial behavior.
Published Jul 3, 2026 · Analyzed Jul 6, 2026