We're publishing our most detailed threat intelligence report to date. It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them. We disrupted every operation in the report, and used the lessons from them to strengthen our safeguards. Where appropriate, we also shared what we found with authorities and other AI companies. These cases are not typical: we’re highlighting some of the most sophisticated misuse we’ve seen. But they’re especially important to discuss, because they show us where AI misuse is headed, where our safeguards work, and where they need to improve. We’re publishing this report so others can spot the same activity on their own platforms, and so we can give the public a clearer view of how emerging threats develop. Read the report: https://t.co/0EJUnYEgfz
It shows agentic engineers concrete misuse patterns targeting AI systems (cyberattacks, influence ops, bio-related queries) and how safeguards caught them, informing what to monitor on your own platform.
“We disrupted every operation in the report, and used the lessons from them to strengthen our safeguards.”
“These cases are not typical: we're highlighting some of the most sophisticated misuse we've seen.”
“We're publishing this report so others can spot the same activity on their own platforms, and so we can give the public a clearer view of how emerging threats develop.”
Checking sign-in…
Loading comments…