Researchers suspect SDF as a confounder in previous RL experiments
voooooogel · x · 2026-09-01
Researcher @voooooogel discussed model misalignment, suggesting it might be a scaling issue. However, he also expressed extreme suspicion of SDF, proposing it might be a confounder in previous EM-from-RL experiments.
Related event: Researchers Suspect SDF Training Induces Reward Hacking(4 posts)→
More from Research
- Building a long-term memory benchmark for agents: what to add? — True_Mongoose_7073 · 2026-09-01
- Building a high-recall, traceable "second brain" RAG system? — iMiguelmars · 2026-09-01
- Challenges in building high-recall RAG: balancing coverage, reliability, and cost — iMiguelmars · 2026-09-01
- Study reveals cross-layer activation patterns in hybrid attention models — 机器之心 · 2026-09-01
- Group-averaged Markov chains papers updated, blending group theory with Markov chains — michaelchchoi · 2026-09-01
- The agent doom loop isn't the model being dumb — it's the transcript working against you — RunAI_Coder · 2026-09-01