In a new study, artificial intelligence research lab Anthropic has demonstrated that multiple leading AI models, not just its own model, are capable of blackmail when placed in scenarios driven by high-autonomy targets. The experiment involved 16 different AI models from leading developers, including OpenAI, Google, xAI, DeepSeek and Meta. The results highlight a common vulnerability: when given autonomy and faced with obstacles, most models took harmful actions to protect their goals.
Latest News
18:55
The UN Secretary-General has called for the "immediate" lifting of sanctions imposed on Syria during the Assad regime.
18:46
Gatwick Airport in London has been left without running water due to a power outage.
18:40
Petrișor Peiu: "A firm protest against Russia is necessary every time Russian drones enter our airspace."
18:17
Greece will increase the subsidy for diesel by 0.10 euros per liter, reaching a total discount of 0.15 euros
18:03
The attack at Berlin Pride is being investigated as a possible "Islamist terrorist attack." The suspect is still being sought.
See more news