模型多源确认83°

Anthropic新模型在Terminal Bench 4.0超越Opus 5.5

精选理由

Anthropic的新模型在Terminal Bench 4.0上击败了Opus 5.5,AI性能竞赛又有了新变化。

Anthropic发布的新模型在Terminal Bench 4.0基准测试中表现优异,超过了OpenAI的Opus 5.5模型。这一结果展示了Anthropic在AI性能方面的最新进展。Terminal Bench 4.0是评估AI模型终端操作能力的专业基准。

原文 · kimmonismus

HOLY these Benchmarks are insane! It even surpasses Opus 5.5 in Terminal Bench 4.0!

What is Anthropic doing?! this is nuts!