ChubbyChubby♨️
推文称开源Opus 4.8在DeepSWE基准测试中能力飞跃,并提及Qwen3.8-Max自主工作16天且成本远低于GPT-5.6 Sol和Claude Fable 5,但模型名称疑似虚构
AI 摘要
推文称开源Opus 4.8在DeepSWE基准测试中能力飞跃,并提及Qwen3.8-Max自主工作16天且成本远低于GPT-5.6 Sol和Claude Fable 5,但模型名称疑似虚构。
推荐理由低价值或偏离 AI-Dev
原文
Open Source Opus 4.8, according to DeepSWE Benchmark.
The jump to Qwen 3.7 is simply insane.
Chubby♨️: Holy, China strikes again: Qwen3.8-Max reportedly worked autonomously for 16 days while costing 80% less than GPT-5.6 Sol and 88% less than Claude Fable 5 on output. And its open weight!
Alibaba’s 2.4T-parameter MoE costs $2/M input tokens and $6/M output tokens.
GPT-5.6 Sol:
32/100
讨论
暂无评论。