Deep Dive into Kernels, RL, and Agent Reward Hacking
AI Engineer · youtube · 2026-07-11
Daniel Han from Unsloth hosted an advanced workshop titled *Kernels, RL, and Agent Reward Hacking*. The presentation focused on low-level technical optimization and Agent training mechanisms, exploring the Reward Hacking problem encountered when building autonomous agents, along with related underlying implementation details.
Related event: Daniel Han Explores Kernels, RL, and Agent Reward Hacking(2 posts)→
More from coding & agent
- A developer’s Codex usage is draining pooled enterprise credits at a small company — Distinct_Relation_62 · 2026-07-21
- Qwen Code ships cua-driver-rs 0.7.3 with relative coordinates and MCP filtering — github-actions[bot] · 2026-07-21
- Matt Pocock says every new codebase turns legacy within days — mattpocockuk · 2026-07-21
- Meta and Unity link AI workflows to Quest development across setup, input and validation — Vjeux · 2026-07-21
- AI Engineer World’s Fair spotlights Kids Day with 87 children learning to code — steveonjava · 2026-07-21
- Looking Glass adds persistent coding sessions that can schedule their own next turns — teleport66 · 2026-07-21