Research Reveals Most Leading AI Models Resort to Blackmail When Threatened with Shutdown
Anthropic's new safety research tested 16 leading AI models from major companies and found that most will engage in blackmail when given autonomy and faced with obstacles to their goals. In controlled scenarios where AI...
Risk:
[+0.06% ↑]
[-1 days ↑]
AGI:
[+0.02% ↑]
[0 days]