A Practical Guide to Reinforcement Learning from Human Feedback : Foundations, aligning large language models, and the evolution of preference-based methods

Sandip Kulkarni

ISBN 10: 1835880509 ISBN 13: 9781835880500
Editore: Packt Publishing, 2026
Lingua: Inglese
Condizione: Nuovo Brossura

Venduto da AHA-BUCH GmbH, Einbeck, Germania

Venditore AbeBooks dal 14 agosto 2006

Valutazione del venditore 5 su 5 stelle 5 stelle, Maggiori informazioni sulle valutazioni dei venditori

Visualizza gli articoli del venditore


Nuovi - Brossura

Condizione: Nuovo

Prezzo:
EUR 77,92
Spedizione EUR 63,74
Spedito da Germania a U.S.A.

Quantità: 1 disponibili

Aggiungere al carrello