AGI Discussions on Data and Defensive Capabilities
RyanGreenblatt · x · 2026-07-13
This in-depth post continues the discussion on AI training data, evaluation data, and whether "improving defensive capabilities is sufficient."
Core viewpoints:
- Previously, the general consensus was to oversample certain data while filtering out other data (like evals); today's views are more nuanced but directionally similar.
- The author argues that by the time many issues become "visible," it is often too late or requires drastic intervention.
- Merely improving defensive capabilities in a broad sense is usually insufficient to truly solve such problems, though it still helps to some extent.
Related event: AI Community's Shift in Data Strategy(3 posts)→
More from AGI Musings
- Daniel Lemire says Claude, Grok, and even Gemma beat his NLP pipeline — lemire · 2026-07-21
- AI Didn't Make Software Development Cheap, It Made Bad Ideas Cheap — Warm-Reaction-456 · 2026-07-21
- HarmonicMath says Lean autonomously solved eight previously studied open problems — MarioKrenn6240 · 2026-07-21
- AI hype turns noise into risk, the author argues, and proposes radical transparency for ethics — AryHHAry · 2026-07-21
- A 2015 ML veteran reflects on how the field changed and where it goes next — j_foerst · 2026-07-21
- After AI Accelerates Analysis, Real-World Feedback Becomes the New Bottleneck — claud_fuen · 2026-07-21