Opus 5 在 ARC-AGI 3 基准上两个月内性能从不足 5% 提升至 30% 以上,且比 Fable 5 每任务成本低 26%,展现 AI 模型快速进步
AI 摘要
Opus 5 在 ARC-AGI 3 基准上两个月内性能从不足 5% 提升至 30% 以上,且比 Fable 5 每任务成本低 26%,展现 AI 模型快速进步。
推荐理由常规快讯,保留列表
原文
1) Opus 4.8 is only two months old. In the span of just two months, its performance on the ARC-AGI 3 benchmark has improved from under 5% with Opus 4.8 to over 30% with Opus 5. Two months.
2) Over that same two month period, Opus 5 has emerged as a model that is not only better than Fable 5, but also more efficient than its predecessor, Opus 4.8. According to the Artificial @ArtificialAnlys Benchmark, Opus 5 is “offering comparable intelligence to Fable 5 at 26% lower Cost per Task.”
Two months that have completely changed the game.
Opus 5 is the release many had been hoping for, and it proves just how quickly everything is improving. Now it’s OpenAI’s turn with GPT-6, and Fable 5.1 probably won’t be far behind. Absolutely insane!
讨论
暂无评论。