Manning published Reinforcement Learning from Human Feedback, a new textbook by Nathan Lambert. The book documents specific post-training methods, including rejection sampling and outcome reward models, that previously lacked comprehensive online guides. It fills a technical gap for developers. Practitioners gain a structured reference for aligning and refining large language models.