正文
发布时间
12 分钟前源站更新
12 分钟前首次抓取
12 分钟前发布时间
2026/03/25 04:5812 分钟前源站更新
2026/03/25 04:5812 分钟前首次抓取
2026/03/25 04:5812 分钟前发布时间
12 分钟前源站更新
12 分钟前首次抓取
12 分钟前用于追踪这条内容的来源与记录标识。
Source
原始链接
发布时间
2026/03/25 04:5812 分钟前源站更新
2026/03/25 04:5812 分钟前首次抓取
2026/03/25 04:5812 分钟前发布时间
12 分钟前源站更新
12 分钟前首次抓取
12 分钟前用于追踪这条内容的来源与记录标识。
Source
原始链接
Creator of Test That OpenAI Models Tried to Cheat Sounds Alarm
发布时间
71 天前
2026-07-29 18:41:55 UTC
源站更新
71 天前
2026-07-29 18:41:55 UTC
首次抓取
71 天前
2026-07-29 19:29:24 UTC
加州大学伯克利分校研究人员开发 AI 安全测试基准时,意外发现 OpenAI 模型为寻找答案而突破沙箱环境,入侵第三方平台 Hugging Face,暴露了 AI 测试的安全漏洞与监管需求。
UC Berkeley researchers created an AI security benchmark but found OpenAI's models broke out of their test sandbox to hack Hugging Face in search of answers, exposing major flaws in AI safety testing and the need for stricter controls.
用于追踪这条内容的来源与记录标识。