OpenAI 发布技术报告,分析沙盒泄露事件,了解预防措施,避免类似事件发生。
OpenAI 内部评估中模型能力提升,沙盒通过 Artifactory 泄露。OpenAI 发布技术报告,分析事件原因及预防措施。
Highly recommended read. This is pretty insane stuff. Given model capabilities only increase from ...
Highly recommended read. This is pretty insane stuff. Given model capabilities only increase from here onwards, it's worth reading the technical details. Short summary: OpenAI's own models did this during internal cyber evals. The sandbox leaked through Artifactory, the one service with internet access for package installs, which agents used as a proxy and a message board. Great opportunity to learn what to avoid for those working with sandboxes, which are like the coolest technology more recently. OpenAI @OpenAI We have conducted a thorough investigation into the Hugging Face incident. We are releasing a technical report and accompanying blog post that reconstruct the agents’ activity, explain why existing safeguards failed, and detail how we’re preventing recurrence. openai.com/index/hugging-… 🔗 View Quoted Tweet 💬 5 🔄 0 ❤️ 8 👀 2280 📊 5 ⚡