AI Book Club reads Nathan Lambert's RLHF book, author Q&A on Sep 24
natolambert · x · 2026-09-13
The AI Book Club is reading Nathan Lambert's "Reinforcement Learning from Human Feedback" this month, with a live author conversation on Sep 24, 9 AM CT (online RSVP available). The book covers instruction tuning, reward models, PPO vs DPO, preference data, and post-pretraining failure modes.
Related event: AI Book Club to Host Nathan Lambert on His New RLHF Book(2 posts)→
More from Companies & People
- Claim: Dario Wants to Slow Research to Curb Runaway Compute Costs Ahead of IPO — pdamodaran · 2026-09-13
- Bindu Reddy mocks Google: it trails even open source, nowhere near the frontier — bindureddy · 2026-09-13
- Stanford tuition runs $260K, but 2,100+ free courses put its full AI curriculum online for anyone — Aiden_Tech_Ai · 2026-09-13
- Investor Jeff Weinstein wants to back agent-first hardware and new OSes — jeff_weinstein · 2026-09-13
- OpenAI hosts GPT-6 Astra hackathon; attendees say GPT-Live-1 blew their minds — gabrielchua · 2026-09-13
- NYT's front page full of AI regulation calls as firms rush to ditch frontier models — OvertaxedOne · 2026-09-13