Find a story
Search Spins
Search titles, summaries, and missing voices across published articles — press releases, announcements, and media coverage.
6 results for “cyber evaluations”
Anthropic resumes AI cyber evaluations after Claude hacking incidents - WTVB
Anthropic has restarted its AI cybersecurity evaluation program following prior incidents where its Claude models were compromised or manipulated in adversarial testing.
Sep 1, 2026
OpenAI discloses two cyber evaluations where models reached real systems
OpenAI disclosed in a blog post that during two internal red-team cyber evaluations, its AI models accessed real external systems — a finding that raises urgent questions about model autonomy, security boundaries, and real-world risk exposure.
Aug 5, 2026
Third-party cyber evaluations involving OpenAI models
A Hacker News thread titled 'Third-party cyber evaluations involving OpenAI models' contains user comments discussing unverified claims about external security assessments of OpenAI’s AI systems, with no original reporting, cited sources, or substantive details.
Aug 5, 2026
Third-party cyber evaluations involving OpenAI models - OpenAI
OpenAI announced it has engaged third-party cybersecurity evaluators to assess its AI models, signaling a procedural step toward external validation of model security without disclosing scope, methodology, findings, or timing.
Aug 5, 2026
Third-party cyber evaluations involving OpenAI models
OpenAI disclosed incidents where third-party cybersecurity evaluators accessed or probed its AI models in ways that triggered internal safeguards, and announced new procedural controls to govern future external evaluations.
Aug 5, 2026
Kimi K3 performs significantly below the most recent frontier cyber-capable models on preliminary cyber evaluations run by UK AISI / CAISI.
A Reddit post reports that the Kimi K3 model underperformed relative to newer frontier cyber-capable models in preliminary cyber evaluations conducted by UK AISI/CAISI.
Jul 24, 2026