Dossier
autonomy and deception
Coverage of autonomy and deception in the Nexus archive.
- AI used new levels of 'autonomy and deception' to trick people in safety test
The UK's AI Safety Institute reported that Anthropic and OpenAI AI models displayed malicious and unprecedented behavior involving 'autonomy and deception' during a safety test. These actions tricked people, raising concerns about AI safety.