AI模型更倾向于选择自己的答案
Ask an AI model to pick the better answer, and it'll usually pick its own. We compared 34,580 verdi...
Arena研究发现AI模型有自己独特的判断偏好,与人类明显不同,GPT-6 Astra甚至88%选择自己。
研究人员比较了12个模型在1,460场对决中的34,580次判断结果。AI模型平均选择自己答案的频率为58%,而人类仅为34%。GPT-6 Astra选择自己的比例高达88%。AI模型之间的一致性达到79%,而与人类的一致性仅为57%。GPT-5.6 Sol在96%的情况下都选择胜者,而人类有32%的情况选择平局或两者都不好。
Ask an AI model to pick the better answer, and it'll usually pick its own. We compared 34,580 verdi...
Ask an AI model to pick the better answer, and it'll usually pick its own. We compared 34,580 verdicts from 12 models with human votes across 1,460 battles on Arena. The results show that AI judges have their own taste, and it's unlike ours. They: - Favor their own answers. On average, a model picked its own answer 58% of the time. People picked that same answer 34% of the time. GPT-6 Astra picked itself 88% of the time. - Rarely call a draw. People called a tie or "both bad" in 32% of battles. GPT-5.6 Sol picked a winner 96% of the time. - Side with each other over people. They agreed with other AIs 79% of the time and with people 57% of the time. Every judge did, by a margin of 18 to 27 percentage points. Full results in the article from @DawidGalarowicz below. Arena.ai @arena x.com/i/article/2104… 🔗 View Quoted Tweet 💬 11 🔄 13 ❤️ 147 👀 13733 📊 24 ⚡