# Kimi K3 在 nextjs.org/evals 基准测试中表现最佳，超越所有专有模型，成为首个在该基准上领先的开源模型，可能标志开源模型突破

- 来源：月之暗面
- 发布时间：2026-07-17 07:37
- AIWatch 分数：70
- AIWatch 标记：当日精选
- AIWatch 链接：https://aiwatch.icu/events/evt_01kxzhmt3r6zm9htdxer243pvw
- 原文链接：https://x.com/Kimi_Moonshot/status/2078058885064351839

## 精选理由

常规快讯，保留列表

## AI 摘要

Kimi K3 在 nextjs.org/evals 基准测试中表现最佳，超越所有专有模型，成为首个在该基准上领先的开源模型，可能标志开源模型突破。

## 正文

This is the first time that an open model is ahead of all proprietary ones for this comprehensive web engineering benchmark.

Notes:

▪️ Benchmarks don’t always tell the full story, although this is important signal, adding to mounting evidence that this could be a breakthrough moment for open models

▪️ No model as of yet has reached 100% completion on this set of evals. The top performer peaks at 92% and 96% “with help”
