Manning published Reinforcement Learning from Human Feedback, a new textbook by Nathan Lambert. The text fills gaps in online documentation for post-training methods. It covers specific techniques like rejection sampling and outcome reward models. Practitioners gain a structured technical guide for aligning large language models beyond basic fine-tuning.