Claude Opus 5 dominates KernelBench, hinting at even stronger internal model
scaling01 · x · 2026-08-16
scaling01 tweets that Claude Opus 5 excels on KernelBench, hinting that its internal model (Model 2) is even more capable.
Related event: Claude Opus 5 Tops KernelBench Leaderboard(2 posts)→
More from Models
- Qwen3.8 Quantization Experiments: Which Weight Groups Matter Most? KLD Tests Reveal Surprises — enginetown · 2026-08-16
- Mini AGI benchmark exposes vision models' failure to spot pareidolic patterns — legit_api · 2026-08-16
- 0831 Version Shows Solid Improvement Over Preview — jeff_weinstein · 2026-08-16
- Discussion on Context Activating Weights and MoE Mechanisms — teortaxesTex · 2026-08-16
- Where do LoRAs fit into the H3 and official workflow? — james25679 · 2026-08-16
- Analysis on Chinese Models' Overfitting to Eval Framing vs. Capability — teortaxesTex · 2026-08-16