Manning published Reinforcement Learning from Human Feedback, a new textbook by Nathan Lambert. The book documents specific post-training methods, including rejection sampling and outcome reward models, that previously lacked comprehensive online documentation. It provides a technical guide for practitioners aligning LLMs. This resource fills a gap in available educational materials for model alignment.