AI Book Club reads Nathan Lambert's RLHF book, author Q&A on Sep 24

natolambert · x · 2026-09-13

The AI Book Club is reading Nathan Lambert's "Reinforcement Learning from Human Feedback" this month, with a live author conversation on Sep 24, 9 AM CT (online RSVP available). The book covers instruction tuning, reward models, PPO vs DPO, preference data, and post-pretraining failure modes.

Related event: AI Book Club to Host Nathan Lambert on His New RLHF Book(2 posts)→

Original post →

More from Companies & People

Companies & People channel →