行业83°

OpenAI 暂停前沿模型 RL 训练两周

As models become more capable, the risks associated with developing and testing them internally also...

精选理由

OpenAI 暂停前沿模型 RL 训练两周,先搞红队测试和加固环境。

AI 摘要

OpenAI 暂停了针对部署的最新模型的强化学习训练两周。团队在此期间加固了研究环境并进行了红队测试,同时扩大了监控覆盖范围。目前最大规模的前沿 RL 运行仍处于暂停状态。小规模训练和评估正在运行以验证这些安全措施并建立对齐证据。

图片来源 · OpenAI
原文 · OpenAI

As models become more capable, the risks associated with developing and testing them internally also...

As models become more capable, the risks associated with developing and testing them internally also grow. We temporarily paused reinforcement learning (RL) training on our latest models intended for deployment for two weeks while we hardened and red-teamed our research environments and expanded monitoring coverage. Our largest planned frontier RL run remains on hold while smaller-scale training and evaluations validate these safeguards and establish more evidence of alignment. openai.com/index/pacing-m… 💬 179 🔄 115 ❤️ 1316 👀 121486 📊 304 ⚡