A Practical Guide to Reinforcement Learning from Human Feedback. Foundations, aligning large language models, and the evolution of preference-based methods
Wydawnictwo: Packt Publishing (Z chęcią przeczytam książkę w języku polskim)
ISBN: 978-18-3588-051-7
ISBN bez myślników: 9781835880517
Typ okładki: Softcover
Liczba stron: 402
Język: polski (Polska)