AI模型精选

GLM-5.2在三项物理模拟任务中击败Kimi K2.7 Code

New @Zai_org GLM-5.2 beats Kimi K2.7 Code on phys…

精选理由

智谱的GLM-5.2写物理模拟代码完胜Kimi K2.7,三个场景全部精准,Kimi翻车在弹簧穿透和球乱撞上。

AI 摘要

智谱GLM-5.2与月之暗面Kimi K2.7 Code在三个物理模拟HTML5编程任务中对比。GLM-5.2使用12,640 tokens完成全部任务,包括台球碰撞、弹簧上方方块弹跳和高尔顿板,粒子和动量表现正确。Kimi K2.7 Code仅用7,420 tokens,但三个场景均出现严重错误:方块穿透弹簧、台球碰撞不真实、高尔顿板珠子重叠。评测显示GLM-5.2在物理模拟细节和精度上显著优于Kimi K2.7 Code。

原文 · @atomic_chat_hq

New @Zai_org GLM-5.2 beats Kimi K2.7 Code on phys…

New @Zai_org GLM-5.2 beats Kimi K2.7 Code on physics contest!

We gave both models the same three prompts and asked them to build self contained HTML5 sims with real physics and no libraries:

1. Pool break 2. Block on a bed of springs 3. Galton board

Outputs: GLM-5.2: 12,640 tokens Kimi K2.7 Code: 7,420 tokens

GLM 5.2 nailed all three, and it did it with way more detail and polish. The break conserved momentum, the block bounced off the springs and the Galton beads spread into a clean bell curve. Kimi struggled on every scene: its block fell straight through the springs, its break didn't look realistic with the balls colliding all wrong, and on the Galton board its balls overlapped and piled into each other instead of spreading out