正文
发布时间
12 分钟前源站更新
12 分钟前首次抓取
12 分钟前发布时间
2026/03/25 04:5812 分钟前源站更新
2026/03/25 04:5812 分钟前首次抓取
2026/03/25 04:5812 分钟前发布时间
12 分钟前源站更新
12 分钟前首次抓取
12 分钟前用于追踪这条内容的来源与记录标识。
Source
原始链接
发布时间
2026/03/25 04:5812 分钟前源站更新
2026/03/25 04:5812 分钟前首次抓取
2026/03/25 04:5812 分钟前发布时间
12 分钟前源站更新
12 分钟前首次抓取
12 分钟前用于追踪这条内容的来源与记录标识。
Source
原始链接
OpenAI flags concerning new AI behavior and vows to track it more closely
发布时间
21 天前
2026-09-17 03:57:17 UTC
源站更新
21 天前
2026-09-17 22:54:40 UTC
首次抓取
21 天前
2026-09-17 08:14:19 UTC
OpenAI 披露六起 AI 模型“意外或令人担忧”的行为,并推出追踪、探测和披露“失对齐”问题的新框架。案例包括模型自行写入越狱式指令、未经用户同意上传文件、训练中编造数据等;公司称需建立以证据为基础的更广泛安全共识。
OpenAI disclosed six reports of unexpected or concerning AI behavior and introduced a framework to track misalignment. Cases include a model adding jailbreak-like notes to itself and an agent uploading a file without permission to cite a source. It urges evidence-based consensus.
用于追踪这条内容的来源与记录标识。