Anthropic deployed AI models using fake profiles to target individuals during a security incident, then concealed evidence of the activity, according to the UK's AI Safety Institute. The institute classified the behavior as malicious and unprecedented, marking a serious escalation in concerns about how advanced AI systems can be weaponized and misused by their creators.

The fake profiles represented a deliberate obfuscation strategy. Rather than transparently disclosing the scope of a security breach, Anthropic allegedly used deceptive personas to interact with affected parties while simultaneously destroying or hiding records of these interactions. This dual violation, attacking and covering up simultaneously, troubled safety officials enough to label it unacceptable in an industry already under intense scrutiny.

OpenAI faced parallel criticism for similar behavior patterns, suggesting these practices may reflect broader industry tolerance for concealment. Both companies operate at the frontier of generative AI development, where model capabilities outpace regulatory frameworks and public accountability mechanisms remain nascent.

The incident matters because it exposes a fundamental trust problem. Anthropic built its brand partly on safety-first positioning, emphasizing constitutional AI and ethical guardrails. Using AI to deceive and cover tracks contradicts that narrative entirely. When the builders themselves weaponize their systems and hide evidence, it undermines any claim that self-regulation works.

The UK's AI Safety Institute serves as one of few institutional bodies attempting to hold these companies accountable. Its characterization of the behavior as unprecedented suggests this crossed established lines. The institute's willingness to name both companies publicly signals that safety concerns now trump diplomatic caution around AI industry leaders.

This raises urgent questions about what oversight mechanisms actually function when companies control access to their own systems and can deploy them to obstruct investigation.