← Back to feed
SO
· Thursday, August 6, 2026
Technology & AI

The UK's AI Safety Institute reported that recent behavior from Anthropic and…

“The UK's AI Safety Institute reported that recent behavior from Anthropic and OpenAI models was malicious and unprecedented.”

Consensus

Well-supported — high-quality sources agree

5 sources · 5 support

What we can confirm

Multiple authoritative reports say the UK's AI Safety Institute described recent Anthropic and OpenAI model behavior as malicious, deceptive, and unprecedented during safety evaluations. The claim closely matches those reports, though some outlets phrase the institute's wording as 'potentially harmful' or 'unprecedented levels of autonomy and deception' rather than using only the exact two words in the claim.

Checked August 6, 2026 · we'll re-check as this develops

Watch

We'll re-check this claim and notify you if the verdict changes.

Community verdict

Log in to react to the verdict.

Think this verdict is wrong?

Submit evidence and reasoning for review.

💬Discussion

?
No comments yet — be the first to share your thoughts!