模型多源确认

四款图像编辑模型连续编辑 30 步对比:Ideogram 4.5 稳定性领先

精选理由

四个模型改同一张图 30 次,Ideogram 4.5 几乎不动画面,GPT Image 2.5 越改越走样,做图的朋友值得看看测试方法。

Artificial Analysis 用同一张照片和连续 30 步编辑测试了 GPT Image 2.5 Sunburst、Ideogram 4.5、FLUX 3 和 Nano Banana 2.1 四款模型。Ideogram 4.5 和 FLUX 3 采用局部编辑方式,小改动(如添加郁金香花瓶)下 95% 以上的画面保持不变。GPT Image 2.5 Sunburst 每次编辑会重渲染大部分画面,仅约五分之一内容不变,色彩和细节偏移随轮次累积最明显。Nano Banana 2.1 介于两者之间:编辑本身局部化,但画面整体会轻微变化并逐渐变暗。

原文 · Artificial Analysis

We gave four frontier image editing models the same photo and 30 edits in a row: Ideogram 4.5 keeps most of the room intact, while the others drift, GPT Image 2.5 Sunburst most visibly.

We've seen some interesting demos of multi-turn editing consistency from the latest image editing models, so we ran our own test: GPT Image 2.5 Sunburst, #1 on our Image Editing leaderboard, against Ideogram 4.5, FLUX 3 and Nano Banana 2.1. Our leaderboard scores single edits; this tests what happens when 30 consecutive changes stack up, with each model editing its own previous output through a 30-step real estate staging sequence: light the fire, add a sofa, repaint the walls, swap day for twilight, and more.

Why are the final results so different?

@ideogram_ai's Ideogram 4.5 and @bfl_ml's FLUX 3 edit locally. On small edits like adding a vase of tulips, we measured that they left 95% or more of the image essentially untouched. GPT Image 2.5 (Sunburst) re-renders most of the entire image on every edit, leaving only about a fifth of the image unchanged, so small shifts in colour and detail compound over turns. Nano Banana 2.1 sits in between: its edits stay local, but the rest of the image shifts slightly and gradually darkens.

  • Ideogram10-09 17:30原文
  • rohanpaul_ai10-08 17:30原文
  • OpenAI: 官网动态10-08 12:00原文