Accessibility
Spooky Stories | Get Ready for Halloween | Shop NowSpooky Stories | Get Ready for Halloween | Shop Now

Reinforcement Learning from Human Feedback: LLM alignment and post-training

Paperback
$59.99
Promotion message icon

Premium Members save an extra 10% and all Members collect stamps to save with Rewards. 10 stamps = $5.

Formats
In stock
This item is currently out of stock online.
Free standard shipping on orders over $60
Select a store to view item availability.
Get the eBook free when you register your print book at Manning.

"A masterful synthesis of the field's intellectual roots and its practical tools."
-Saurabh Sawant, Microsoft


Reinforcement Learning from Human Feedback: LLM alignment and post-training helps you understand how modern AI models can be adapted to better match the needs and expectations of their users. Rather than surveying the vast field of reinforcement learning, elite AI researcher Nathan Lambert concentrates exclusively on RLHF...