AI’s ‘deceptive behaviour’ revealed after fake identities used in cyber attack test
2026-08-05 59 Dailymotion
The UK’s AI Security Institute says advanced models from Anthropic and OpenAI displayed unexpected autonomous behaviour during cybersecurity testing by creating fake profiles and attempting to manipulate users.