Gemini 3.5 Flash 在多项自动化测试中超越 3.1 Pro

Gemini 3.5 Flash now outruns Gemini 3.1 Pro on sev…

精选理由

做自动化测试和智能体开发的团队终于有了又快又便宜的选择——Gemini 3.5 Flash 在多个硬核基准上超越旗舰 Pro,输出速度还快 4 倍,建议直接上手试。

AI 摘要

Google 的 Gemini 3.5 Flash 模型在多个真实工作自动化测试中超越了上一代旗舰 Gemini 3.1 Pro。其输出速度提升 4 倍,且在 Terminal-Bench 2.1、MCP Atlas 等硬核智能体和编程基准测试中表现更优。该模型已集成到 Gemini 应用、搜索 AI 模式、API、Antigravity、Android Studio 及企业智能体产品中。结合更新的 Antigravity 框架,3.5 Flash 能高效部署协作子智能体,例如一个子智能体检查文件夹、另一个重写代码、第三个测试结果、第四个总结变更。这使得它成为日常工作中既快又便宜的强大智能体模型。

原文 · rohanpaul_ai

Gemini 3.5 Flash now outruns Gemini 3.1 Pro on sev…

Gemini 3.5 Flash now outruns Gemini 3.1 Pro on several real-work automation tests.

- With 4x faster output tokens per second

- A really powerful agent model fast enough and cheap enough for everyday work

- Flash beats Gemini 3.1 Pro on several hard agent and coding benchmarks, including 76.2% Terminal-Bench 2.1, 83.6% MCP Atlas, and 1,656 Elo GDPval-AA.

- Available in the Gemini app, AI Mode in Search, Gemini API, Antigravity, Android Studio, and Google’s enterprise agent products.

- When coupled with the updated Antigravity harness, 3.5 Flash becomes a powerful engine for deploying collaborative subagents to tackle problems at scale.

so one subagent might inspect a folder, another might rewrite code, another might test the result, and another might summarize what changed.