Gemini 3.5 Flash 的主动扩展行为展示了 AI 从被动执行到主动理解的转变,做创意生成或前端开发的团队值得关注这种新能力,建议试试看它能否提升你的工作流。
用户要求 Gemini 3.5 Flash 渲染佩特拉宝库,模型不仅生成了主体建筑,还自动构建了周围整个石峡谷,并添加了环境音效,这些并未在提示中指定。这种主动扩展场景的行为与其他前沿模型不同,展示了更强的智能体特性。在 Arena 评测中,Gemini 3.5 Flash 在文本和代码前端任务中排名第9,相比前代提升70分,并在同价位模型中达到最高分。该模型在内容创作、游戏、消费产品等子类别中表现突出。
Asked Gemini 3.5 Flash to render the Petra Treasury. It built the entire stone canyon around it - so...
Asked Gemini 3.5 Flash to render the Petra Treasury. It built the entire stone canyon around it - something other frontier models didn't do. Gemini also added ambient sound, which wasn’t in the prompt either. Whether you want this agentic behavior depends on what you're trying to do, but it's a notable departure from how other frontier models behave on the same prompts. More side-by-side prompts with @GoogleDeepMind 's latest release in the full video (link in thread) 👇 Your browser does not support the video tag. 🔗 View on Twitter Arena.ai @arena Gemini 3.5 Flash has landed #9 for Text and Code Arena: Frontend. Code Arena: Frontend evaluates models on agentic frontend coding tasks from real users building apps and websites (HTML and React). Scoring 1507, this is a significant +70 point improvement over Gemini-3 Flash. Sub-category highlights: - #7 Content Creation Tools - #8 Gaming - #8 Consumer Product - #9 Data & Analytics - #10 Reference-Based Design In Text Arena: #9 overall. Gemini 3.5 Flash also moves the price–performance frontier as the new top Arena score in its price tier. Congrats to the @GoogleDeepMind team on this launch! Click into the thread to see the rankings by each arena. 🔗 View Quoted Tweet 💬 1 🔄 1 ❤️ 9 👀 899 📊 2 ⚡