Find a story
Search Spins
Search titles, summaries, and missing voices across published articles — press releases, announcements, and media coverage.
14 results for “moderation”
TikTok lays off 250 employees, shutters its Nashville office
TikTok closed its Nashville office and laid off 250 employees, primarily from its content-moderation team, as part of a broader operational consolidation.
Aug 6, 2026
TikTok's US entity says it is closing its Nashville office, which held some of its content moderation team; filing: TikTok laid off 250 employees at the office (Emmett Lindner/New York Times)
TikTok's U.S. joint venture closed its Nashville office, laying off 250 employees, many of whom were part of its content moderation team.
Aug 6, 2026
Reddit aims to make ‘karma’ less important for first-time posters with shift to AI moderation tools
Reddit is deploying AI moderation tools to weaken reliance on karma and account age as gatekeeping mechanisms, aiming to improve newcomer onboarding while reducing abuse.
Aug 5, 2026
Reddit is introducing a new moderator: AI
Reddit is rolling out an AI-powered moderation suite called Rules Hub that uses LLMs to auto-enforce community rules, initially for new subreddits and expanding broadly later this year.
Aug 5, 2026
"Generate Cover Art for a Romance Novel Called (Insert Nonsense)
A Reddit user shared anecdotal examples of using ChatGPT to generate absurd romance novel cover art, noting occasional moderation blocks requiring explicit disclaimers about non-explicit content.
Aug 5, 2026
Mistral's Shieldstral: 3B open-weights model for multimodal moderation
A forum post on Hacker News titled 'Mistral's Shieldstral: 3B open-weights model for multimodal moderation' references an unverified, unnamed model with no supporting details, links, or evidence — functioning as a speculative signal rather than a report of an actual release.
Aug 5, 2026
Analyzing Toxic Behavior and Its Impact on the Mastodon Community
A new arXiv preprint analyzes toxic behavior on Mastodon using ML methods to map trends and implications for community health and decentralized governance.
Jul 27, 2026
Meta just created a moderation nightmare for its smart glasses
Meta's smart glasses have triggered intense public backlash over privacy violations and nonconsensual recording, prompting Meta to ban certain user-generated content filmed with the devices.
Jul 25, 2026
I actually enjoy 5.6
A Reddit user expresses personal satisfaction with ChatGPT version 5.6, citing improved work quality and usability features despite acknowledging widespread complaints about moderation false positives, image generation loops, and cross-project contamination.
Jul 21, 2026
Automatic Hard Example Synthesis with Multi-Level Agentic Data Curation
Researchers introduced an automated, multi-agent red-teaming system that synthesizes adversarial multimodal examples to improve MLLM content safety robustness, reducing false negatives by 16.7 percentage points on a public benchmark without human labeling.
Jul 17, 2026
Meta's Oversight Board says top AI models may be restricting free expression in its first evaluation of LLMs, as it seeks to expand its influence beyond Meta (Karissa Bell/Engadget)
Meta's Oversight Board issued its first evaluation of large language models, asserting that top AI systems may be restricting free expression, while positioning itself as a cross-platform governance body beyond Meta.
Jul 16, 2026
Discord admits AI moderation bug wrongfully banned users over harmless images
Discord confirmed an AI moderation bug incorrectly banned users for harmless images, affecting accounts since May and causing 200 additional wrongful bans over a recent weekend.
Jul 9, 2026
Reddit says its AI-powered content moderation systems caught 25K "spammy posts and comments" per day in Q1, reducing users' exposure to such content by 20% YoY (Natalie Lung/Bloomberg)
Reddit claims its AI moderation tools detected 25,000 spammy posts and comments daily in Q1, cutting user exposure to such content by 20% year-over-year amid rising stealth marketing campaigns.
Published Jul 6, 2026 · Analyzed Jul 8, 2026
Hate Speech Detection in Turkish and Arabic Languages: A Comprehensive Study
Researchers introduce a dataset for detecting hate speech in Turkish and Arabic languages.
Published Jul 2, 2026 · Analyzed Jul 5, 2026