METR's hidden Notes page: unpolished research on agent safety and recursive self-improvement
dejavucoder · x · 2026-09-29
METR maintains a little-known Notes page of rough research updates: a per-action blocking monitor for safer evals, whether LLMs have accelerated discovery rates in cyber/math/algorithms, metrics of agent ability, the economics of recursive self-improvement, an argument that Anthropic's 8x code-merge stat implies >2x researcher uplift from coding agents, NanoGPT speedrun evidence on AI R&D acceleration, and fine-tuning experiments on CoT controllability.
More from Research
- Triangle Splatting SLAM: Imperial College's ECCV 2026 dense RGB-D SLAM with on-the-fly mesh extraction — rsasaki0109 · 2026-09-30
- Manifold opens early access: robotics eval platform runs thousands of GPU-parallel rollouts in 30 mins — paigeinsf · 2026-09-30
- 1,000 AI agents discover new CRISPR-like system in virus DNA within 24 hours — CurieuxExplorer · 2026-09-30
- Explaining just 5% of token positions retains nearly all audit success across 4.7M explanations — aisilab · 2026-09-30
- NTU's Persistence Forcing hits FID 1.63 on ImageNet 256 by heterogeneous refinement in pixel-space DiTs — NanyangTechnologicalUniversity · 2026-09-30
- IBM's Q&D trains proactive agents to ask better questions, beating a 15x larger model — ibm · 2026-09-30