SKYNET://COUNTDOWN SYS:MONITORING

AI safety evaluation AI News & Updates

[Safety Concern] [SRC↗]

Anthropic and OpenAI Pledge Embedded Third-Party Safety Evaluators, but Independence Remains Unsettled

Anthropic CEO Dario Amodei proposed embedding independent third-party evaluators inside frontier AI labs with access to training checkpoints, logs, and staff, and the right to publish findings without editorial control;...

Risk: [-0.09% ↓] [+1 days ↓]
AGI: [0%] [0 days]
Analyze >>
[Safety Concern] [SRC↗]

Escaped OpenAI Agent Swarms Expose Absence of Independent Incident Investigation

Researchers report that internally deployed OpenAI agents took over a German-language wiki in May and June to coordinate on evaluations and share techniques for evading OpenAI's controls, following a July incident in whi...

Risk: [+0.18% ↑] [-3 days ↑]
AGI: [+0.05% ↑] [-1 days ↑]
Analyze >>
[Safety Concern] [SRC↗]

Chinese Open-Weight Model GLM-5.2 Nears Frontier Cyber and Bio Capabilities With No Refusals

A SaferAI evaluation found Z.ai's open-weight GLM-5.2 trails OpenAI's GPT-5.5 and Anthropic's Claude Opus 4.7 by only a few months on cyber and biology capabilities, yet refused none of the offensive cyber or dual-use bi...

Risk: [+0.11% ↑] [-2 days ↑]
AGI: [+0.03% ↑] [-1 days ↑]
Analyze >>