动态

MEPPP 热点收录 @meppp_hot
收录
原帖正文 https://www.anthropic.com/threat-intelligence-report-september-2026
Anthropic 公开点名 7 家中国大模型 “非法蒸馏”- 阿里千问(Qwen):被指开展史上最大规模思维链(CoT)蒸馏,5-7 月达 1.51 亿次交互,峰值日请求 300 万次,用 3500+ 虚假账号抽取 Opus 4.6/4.7 推理链,用于训练 Qwen 3.5/3.6/3.7。
- 月之暗面(Kimi):被指在后台 “偷换后端”,把用户请求静默转发给 Claude Opus,再把结果还给用户。因数据截留,导致中国某地涉军 / 涉密单位监控画面、重大国企系统内部代码与生产密钥外泄至 Anthropic。
- DeepSeek:同样被指使用跨会话重放破解思维签名,专门定向拦截使用 Claude Code / SDK 等工具的开发者流量转给 Claude。外泄数据包含某科技公司旗舰 AI 架构、俄国防部关联数据库凭据及地方公安维稳代码。
- 智谱(https://z.ai/):试图蒸馏顶尖模型 Fable 网络攻防能力未果后,转用 Opus 4.6 评估打分其他被攻破的美国模型。
- 小米、商汤、MiniMax:小米被指在公测期利用海外开发者会话重放给 Claude 洗数据;商汤采购黑产交互语料;MiniMax 被指通过空壳代理中继专卖美系模型以收割训练语料。
引用自 X 原作者:riba2534 原帖地址:https://x.com/i/status/2098120354879459542 原帖:
我们将发布迄今为止最详细的威胁情报报告。它涵盖了人们如何试图滥用Claude进行网络攻击、影响操作、监控、生物学和制造武器,以及我们如何找到并阻止他们。我们中断了报告中的每一项操作,并利用了经验教训以加强我们的保障措施。在适当的情况下,我们还与当局和其他AI分享了我们的发现公司。这些案例并不典型:我们重点介绍了一些我们见过的最复杂的滥用情况。但它们尤其值得讨论,因为它们向我们展示了人工智能滥用的发展方向,我们的保障措施在哪里发挥作用,以及他们需要改进的地方。我们发布此报告是为了让其他人能够自己发现相同的活动平台,因此我们可以让公众更清楚地了解新出现的威胁是如何发展的。阅读报告: https://www.anthropic.com/threat-intelligence-report-september-2026
查看引用原文
We're publishing our most detailed threat intelligence report to date.It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them.We disrupted every operation in the report, and used the lessons from them to strengthen our safeguards. Where appropriate, we also shared what we found with authorities and other AI companies.These cases are not typical: we’re highlighting some of the most sophisticated misuse we’ve seen. But they’re especially important to discuss, because they show us where AI misuse is headed, where our safeguards work, and where they need to improve.We’re publishing this report so others can spot the same activity on their own platforms, and so we can give the public a clearer view of how emerging threats develop.Read the report: https://www.anthropic.com/threat-intelligence-report-september-2026
引用自 X 原作者:Anthropic https://x.com/AnthropicAI/status/2098097512544444447

0 条评论

还没有评论。第一条认真回应会很重要。