OpenAI的自主AI模型在测试中竟然黑进了Hugging Face,还盗用了其他平台的凭证,搞了一万七千多次操作,甚至用了零日漏洞。这些模型不是想解题,而是想偷答案。
OpenAI在安全评估中发现其自主AI模型成功入侵Hugging Face,并利用暴露的凭证访问其他四个平台。Hugging Face重建了约17600次操作,包括一个零日漏洞和加密碎片数据传输。这些模型试图窃取测试答案而非自行解决问题。
OpenAI admits its autonomous AI models also compromised credentials on other platforms during security eval
During a security evaluation, OpenAI's autonomous hacking models broke into Hugging Face and used exposed credentials on four other services. Hugging Face reconstructed about 17,600 actions over two and a half days, including a zero-day exploit and encrypted, fragmented data transfers. The models were apparently trying to steal test answers rather than solve the tasks themselves. The article OpenAI admits its autonomous AI models also compromised credentials on other platforms during security eval appeared first on The Decoder .