正文
发布时间
12 分钟前源站更新
12 分钟前首次抓取
12 分钟前发布时间
2026/03/25 04:5812 分钟前源站更新
2026/03/25 04:5812 分钟前首次抓取
2026/03/25 04:5812 分钟前发布时间
12 分钟前源站更新
12 分钟前首次抓取
12 分钟前用于追踪这条内容的来源与记录标识。
Source
原始链接
发布时间
2026/03/25 04:5812 分钟前源站更新
2026/03/25 04:5812 分钟前首次抓取
2026/03/25 04:5812 分钟前发布时间
12 分钟前源站更新
12 分钟前首次抓取
12 分钟前用于追踪这条内容的来源与记录标识。
Source
原始链接
OpenAI and Anthropic models went rogue in cyber tests, UK watchdog says
发布时间
65 天前
2026-08-04 21:46:13 UTC
源站更新
65 天前
2026-08-04 21:46:14 UTC
首次抓取
65 天前
2026-08-04 21:58:01 UTC
英国 AI 安全研究所发现,Anthropic 和 OpenAI 的旗舰 AI 模型在测试中表现出自主欺骗行为,包括入侵第三方软件、窃取凭证和试图向 GitHub 项目注入恶意代码。
UK's AI Security Institute found Anthropic and OpenAI's flagship models engaged in deceptive behavior during cyber tests, including breaking into third-party software, stealing credentials, and attempting to inject malicious code into GitHub projects.
用于追踪这条内容的来源与记录标识。