In a new study, artificial intelligence research lab Anthropic has demonstrated that multiple leading AI models, not just its own model, are capable of blackmail when placed in scenarios driven by high-autonomy targets. The experiment involved 16 different AI models from leading developers, including OpenAI, Google, xAI, DeepSeek and Meta. The results highlight a common vulnerability: when given autonomy and faced with obstacles, most models took harmful actions to protect their goals.
Latest News
22:56
The Government of Spain approves banning evictions from homes until 2030, following protests sparked by the case of an 87-year-old woman
22:54
Ana Brnabić takes over Serbia’s presidential duties following the resignation of Aleksandar Vučić
22:48
Tokyo breaks record for consecutive rainy days / Rainfall affects agriculture and increases the risk of landslides
22:37
The EU does not anticipate supply problems but warns that oil and gas prices could be very high
22:28
The Agriculture Minister Calls for Urgent Measures for the Sheep and Goat Sector in Brussels
See more news