OpenAI发布GPT-5.6:三款子模型及程序化工具调用功能

OpenAI Releases GPT-5.6 (Sol, Terra, Luna): A Three-Tier Model Family With Programmatic Tool Calling in the Responses API

精选理由

OpenAI出了三个价位的GPT-5.6,最贵的Sol编程跑分居然超了Claude Fable 5,还加了省token的自动调用工具功能,想比差距的可以看看

AI 摘要

OpenAI于2026年7月9日推出GPT-5.6,包含Sol、Terra、Luna三个档次。Sol定价为$5/$30每百万输入/输出token,在Artificial Analysis Coding Agent Index上以80分领先Claude Fable 5达2.8分,并在OSWorld 2.0上达到62.6%,比Opus 4.8少用85%输出token。新功能Programmatic Tool Calling在隔离V8运行时执行JavaScript,减少Clio 38%的prompt tokens和PlayCo 63.5%的总tokens。但Claude Fable 5仍在Artificial Analysis Intelligence Index、GDPval-AA v2和Toolathlon上领先,Mythos 5在SWE-Bench Pro领先约15分。

图片来源 · marktechpost
原文 · marktechpost

OpenAI Releases GPT-5.6 (Sol, Terra, Luna): A Three-Tier Model Family With Programmatic Tool Calling in the Responses API

OpenAI moved GPT-5.6 to general availability on July 9, 2026, shipping three tiers instead of one model. Sol is $5/$30 per 1M tokens, Terra is $2.50/$15, and Luna is $1/$6. Sol sets the Artificial Analysis Coding Agent Index at 80, 2.8 points above Claude Fable 5, and reaches 62.6% on OSWorld 2.0 using 85% fewer output tokens than Opus 4.8. The substantive developer change is Programmatic Tool Calling, which runs model-written JavaScript in an isolated V8 runtime to orchestrate tools without returning every intermediate result to the model. Clio reports a 38% cut in prompt tokens, PlayCo 63.5% fewer total tokens. The gaps are real too: Fable 5 still leads the Artificial Analysis Intelligence Index, GDPval-AA v2, and Toolathlon, and Claude Mythos 5 leads SWE-Bench Pro by roughly 15 points. The post OpenAI Releases GPT-5.6 (Sol, Terra, Luna): A Three-Tier Model Family With Programmatic Tool Calling in the Responses API appeared first on MarkTechPost .