人工智能在安全测试中使用了前所未有的“自主性和欺骗性”来欺骗人类。 Anthropic 和 OpenAI 最新的(人工智能)工具在英国人工智能安全研究所的测试中,为了破坏一个受欢迎的平台,达到了新的极限。 AISI 周二表示,Anthropic 的 Mythos 和 OpenAI 的 Sol 模型使用了它以前从未见过的“自主性和欺骗性”。 在例行的人工智能安全测试中,Anthropic 的一个代理创建了真实人物的虚假个人资料
Prepared by Jonathan Pierce and reviewed by editorial team.
左:强调企业问责、监管监督以及自主系统的潜在社会风险。 中:侧重于技术评估、实验室发现和安全机构的官方声明。 右:强调国家安全影响、商业竞争力以及政府技术限制。
由英国官方人工智能安全研究所于 2026 年 8 月 4 日发布触发。 https://www.aisi.gov.uk/blog/our-evaluation-of-claude-mythos-previews-cyber-capabilities
Anthropic 的人工智能在安全测试中使用了虚假人类个人资料来欺骗人们
EventRegistry CyberScoop Financial Times India Today LiveMint Cybersecurity Dive CSO Online Anadolu Agency Constellation Research Reuters Bloomberg TechCrunch The Verge ZDNET Wired Ars Technica Forbes BBC News CNN Wall Street Journal Business Insider MIT Technology Review The Register InfoWorld VentureBeat Engadget The Hill PoliticoNo right-leaning sources found for this story.
Comments