Artificial Analysis 在旧金山广告牌展示模型智能指数与成本对比图,指出 Fable 5 智能最高但成本高,Grok 4.5 和 DeepSeek v4 Flash 性价比更优
AI 摘要
Artificial Analysis 在旧金山广告牌展示模型智能指数与成本对比图,指出 Fable 5 智能最高但成本高,Grok 4.5 和 DeepSeek v4 Flash 性价比更优。
推荐理由常规快讯,保留列表
原文
Cost efficiency is becoming a primary factor in evaluating model performance. Fable 5 from @AnthropicAI leads the Index at 60 but averages $2.75 Cost per Task. Grok 4.5 (high) from @elonmusk's @SpaceXAI scores 54 on the Artificial Analysis Intelligence Index but at $0.31 is ~9x cheaper and also sits on the Pareto frontier. DeepSeek v4 Flash is again lower in intelligence at 44 but is ~69x cheaper than Fable at $0.04 Cost per Task.
The Pareto frontier is the set of models that no other model beats on both intelligence and cost at their respective positions. Everything else is paying more for less.
讨论
暂无评论。