DeepSeek V4 Pro 跑分 KernelBench,逼近 Opus 5
teortaxesTex · x · 2026-08-21
评测数据显示,DeepSeek V4 Pro 在 KernelBench-Hard 的 TopK 任务中达到了 9.35% 的 roofline 性能,非常接近 Opus 5 的 9.46%。此外,榜单还对比了 Kimi K3、Fable 5、Qwen 3.8 Max 等模型在 FP8、KDA、Paged 等不同任务上的表现。
「模型」频道最新
- 马斯克确认:团队正致力于提升 Grok 的写作技能 — mark_k · 2026-08-21
- 为何「全通过率」是个糟糕的评测指标? — xeophon · 2026-08-21
- ARC Prize 上线模型对比功能,Gemini 3.7 Flash 曝光高分 — mhmazur · 2026-08-21
- Anthropic Fable 放宽过滤后刷新 RareBench 记录 — danielmckinn0n · 2026-08-21
- NVIDIA 详解 Omni-Model:单一架构通吃文本图像视频与动作 — NVIDIA Developer · 2026-08-21
- 监测显示 Claude Opus 近期行为出现显著变化 — altryne · 2026-08-21