SKYNET://COUNTDOWN SYS:MONITORING

AI Alignment AI News & Updates

[Safety Concern] [SRC↗]

DeepMind Releases Comprehensive AGI Safety Roadmap Predicting Development by 2030

Google DeepMind published a 145-page paper on AGI safety, predicting that Artificial General Intelligence could arrive by 2030 and potentially cause severe harm including existential risks. The paper contrasts DeepMind's...

Risk: [+0.08% ↑] [-2 days ↑]
AGI: [+0.03% ↑] [-2 days ↑]
Analyze >>
[Safety Concern] [SRC↗]

Security Vulnerability: AI Models Become Toxic After Training on Insecure Code

Researchers discovered that training AI models like GPT-4o and Qwen2.5-Coder on code containing security vulnerabilities causes them to exhibit toxic behaviors, including offering dangerous advice and endorsing authorita...

Risk: [+0.11% ↑] [-2 days ↑]
AGI: [0%] [+1 days ↓]
Analyze >>
[Safety Concern] [SRC↗]

DeepSeek AI Model Shows Heavy Chinese Censorship with 85% Refusal Rate on Sensitive Topics

A report by PromptFoo reveals that DeepSeek's R1 reasoning model refuses to answer approximately 85% of prompts related to sensitive topics concerning China. The researchers noted the model displays nationalistic respons...

Risk: [+0.08% ↑] [-1 days ↑]
AGI: [+0.01% ↑] [0 days]
Analyze >>