X:Arena (@arena)· @arena·· 2 天前精选AI 评分
Gemini 4 Argon 与前沿模型对比:单任务成本比 GPT-6.1 Sol 低 33%
Gemini 4 Argon by @GoogleDeepMind just went head-to-head with top frontier models across a set of generations. At 33% lower cost per task than GPT-6.1 Sol, and 70% lower than Claude Opus 5.5, see first impressions from @petergostev on how Gemini 4 Argon compares. https://www.youtube.com/watch?v=h5EL5zThKaI
Arena 分享 Google DeepMind 的 Gemini 4 Argon 与多款前沿模型在生成任务上的对比,单任务成本比 GPT-6.1 Sol 低 33%,比 Claude Opus 5.5 低 70%,并附上 @petergostev 的第一印象视频。评论区对评测方式提出质疑,认为用图形设计任务评估文本生成模型意义有限,也有人追问单任务成本是否包含思考 token。
原文给出 Gemini 4 Argon 与前沿模型的成本对比和实测片段,读者可据此判断其性价比定位。
来源:X:Arena (@arena) · x.lingyaoai.com