Kimi K3在网络安全测试中只拿了32分,远低于美国模型的76%,而且安全机制基本失效,传闻是因为用了蒸馏技术。
英国AI安全研究所与美国AI标准与创新中心测试了Moonshot AI的Kimi K3在进攻性网络任务上的表现。Kimi K3在ExploitBench上得分32%,而领先美国模型得分76%。其安全防护未能阻止漏洞利用开发或模拟攻击。其通用基准测试高分与网络安全弱项之间的差距也符合Moonshot AI蒸馏Anthropic模型的指控。
Kimi K3 trails frontier US models by a wide margin on cyber exploits, and distillation may explain why
The British AI Security Institute and the U.S. Center for AI Standards and Innovation tested Moonshot AI's Kimi K3 on offensive cyber tasks. Kimi K3 scored 32 percent on ExploitBench, compared with 76 percent for leading U.S. models, while its safeguards failed to block exploit development or simulated attacks. The gap between its strong general benchmark scores and weaker cyber performance also fits allegations that Moonshot AI distilled Anthropic's models. The article Kimi K3 trails frontier US models by a wide margin on cyber exploits, and distillation may explain why appeared first on The Decoder .