A Practical Guide to Reinforcement Learning from Human Feedback. Foundations, aligning large language models, and the evolution of preference-based methods
Editorial: Packt Publishing (Z chęcią przeczytam książkę w języku polskim)
ISBN: 978-18-3588-051-7
ISBN sin guiones: 9781835880517
Tipo de cubierta: Softcover
Páginas: 402
Idioma: polaco (Polonia)