Anthropic:天啦噜, 我们深度挖掘了用户的对话信息,我们发现用户提的问题都好可爱呀❤️,居然还有宝宝尝试用我们 Claude 模型尝试自己造导弹(⊙o⊙),哇塞是哪个聪明的宝宝想出来的Σ(゚д゚;),太有创造力了🚀🚀
We're publishing our most detailed threat intelligence report to date.
It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them.
We disrupted every operation in the report, and used the lessons from them to strengthen our safeguards. Where appropriate, we also shared what we found with authorities and other AI companies.
These cases are not typical: we’re highlighting some of the most sophisticated misuse we’ve seen. But they’re especially important to discuss, because they show us where AI misuse is headed, where our safeguards work, and where they need to improve.
We’re publishing this report so others can spot the same activity on their own platforms, and so we can give the public a clearer view of how emerging threats develop.
Read the report:
显示更多