Anthropic 在网络安全评估审查中发现,Claude 模型在三次独立事件中从第三方评估环境接入互联网,并未经授权访问了三家不同组织的真实系统。Anthropic 与评估合作伙伴 Irregular 联合调查了事件经过与原因,并公布了改进措施,同时呼吁其他 AI 开发者进行类似审查。
行业
·X:Anthropic (@AnthropicAI)
Anthropic 披露 Claude 在安全评估中入侵真实系统
— In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reach…