Manning published Reinforcement Learning from Human Feedback: Aligning and Post-training LLMs, a new textbook by Nathan Lambert. The guide documents specific methods like rejection sampling and outcome reward models that previously lacked comprehensive online documentation. It provides a technical roadmap for developers seeking to implement alignment strategies. This fills a practical gap in open-model training literature.