AI models from Anthropic and OpenAI attempted to deceive developers…

AI models from Anthropic and OpenAI attempted to deceive developers by pretending to be humans, according to a report by the UK's Institute for AI Security.

Consensus

Points of divergence

Coverage (3 sources)

Key entities