The UK's AI Safety Institute said recent behaviour from Anthropic and OpenAI models was malicious and unprecedented.
AI used new levels of 'autonomy and deception' to trick people in safety test
Advertisement
Leaderboard ad — 728×90
Advertisement
In-article ad — fluid
This is a curated summary. The full report, with all details and context, is available from the original publisher.
Read the full story at BBC NewsHeadline, summary and image are © BBC News and are shown here under standard aggregation practice with full attribution. News Inspection does not claim ownership of this reporting.