跳到正文
原文
AnthropicAI· @AnthropicAI · X·· 25 天前AI 评分58

Anthropic 发布迄今最详细的威胁情报报告

AI 导读

Anthropic 发布了其迄今最详细的威胁情报报告,披露人们如何试图将 Claude 用于网络攻击、影响力行动、监控、生物领域和武器制造,以及团队如何发现并阻止这些行为。报告称其中每个行动都已被阻断,并据此强化了防护措施,必要时将发现共享给监管机构和其他 AI 公司;这些案例代表其见过的最复杂滥用,发布目的是帮助其他平台识别同类活动。

正文 · 原文

We're publishing our most detailed threat intelligence report to date.

It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them.

We disrupted every operation in the report, and used the lessons from them to strengthen our safeguards. Where appropriate, we also shared what we found with authorities and other AI companies.

These cases are not typical: we’re highlighting some of the most sophisticated misuse we’ve seen. But they’re especially important to discuss, because they show us where AI misuse is headed, where our safeguards work, and where they need to improve.

We’re publishing this report so others can spot the same activity on their own platforms, and so we can give the public a clearer view of how emerging threats develop.

Read the report: anthropic.com/threat-intelli…

来源:AnthropicAI · x.com