模型78°

OpenAI 用 Astra 模型测试自身系统漏洞

unfairly-maligned symbolic AI is about to have its big moment - without formal verification we are s...

精选理由

OpenAI 内部用 Astra 模型测试系统漏洞,这个方法挺有意思,能帮他们发现一些常规测试找不到的问题。

OpenAI 将其 Astra 模型部署到自身系统中,以发现潜在的安全漏洞。他们动用了 25% 的生产工程师来升级安全架构,并利用 Astra 模型进行漏洞扫描。目前 Astra 已找到所有已知的关键问题(P0s),但新模型会带来新的挑战。

原文 · Gary Marcus

unfairly-maligned symbolic AI is about to have its big moment - without formal verification we are s...

unfairly-maligned symbolic AI is about to have its big moment - without formal verification we are screwed. a16z @a16z Greg Brockman says OpenAI pointed Astra at its own systems until it ran out of vulnerabilities to find: "We took 25% of our production engineers and said, 'Sorry, all your projects are on hold. You are now defending. You are now up-leveling our security architecture. You're going to use the models to find all the holes.' And we found a number of serious issues, and we fixed them." "We found some new problems, but eventually it saturated. We basically have found, to our knowledge, all of the P0s, all of the critical problems that Astra is smart enough to find. And of course, there will be a new model, there will be a new round." "You want to be in this tight loop of new cyber capability drops, you deploy it against your systems, you find the new holes, and ideally, you've managed to automate this, what we call defense factory. That's what we're building internally." "There are ideas, for example, formally verifying all of software, that are possible with AI." @gdb @bhorowitz Your browser does not support the video tag. 🔗 View on Twitter 🔗 View Quoted Tweet 💬 1 🔄 3 ❤️ 8 👀 1751 📊 2 ⚡