FactArticle
My post-training textbook 'Reinforcement Learning from Human Feedback: Aligning and Post-training LLMs' is complete and now shipping via Manning, Amazon US, and (in October) Amazon UK.
The author announces the completion and release of his post-training book, published by Manning, and details its distribution across Manning, Amazon US now, and Amazon UK in October. ✦ AI generated
Nathan Lambert · Interconnects · 2026-08-10 · original ↗
After a few long years of finding time to document my lessons from training open models, my post-training book is done! It's published by Manning, under the title Reinforcement Learning from Human Feedback: Aligning and Post-training LLMs. ... The book is shipping from Manning and Amazon US now, and from Amazon UK in October.
Read full article ↗excerpt · fair-use quotation
Around this claim
Mechanism · 2
The book communicates the intuitions and history behind post-training, explaining in simple terms why post-training works, the trade-offs to get it right, and the misconceptions people get stuck on.Nathan Lambert · Interconnects · conf 80%The book existed because critical post-training methods, such as rejection sampling, outcome reward models, and character training, had no foundational online material explaining them and still lack such resources today.Nathan Lambert · Interconnects · conf 75%