Reinforcement learning from human feedback

Auteur : Lambert, Nathan
Éditeur : Manning Publications
ISBN : 9781633434301
Date de publication : 7 oct. 2026
Langue : Anglais
Pays d'origine : USA

Aligning AI models to human preferences helps them become safer, smarter, easier to use and tuned to the exact style the creator desires. Reinforcement Learning from Human Feedback (RLHF) is the process of using human responses to a model’s output to shape its alignment and therefore its behaviour. 

67,99 €
Prix de vente belge indicatif
Disponibilité
À paraître

Pour commander, veuillez vous connecter à votre compte.