preface
I became interested in AI during graduate school at UC Berkeley, where I arrived in the summer of 2017 and learned how deep learning was taking the computer science world by storm. I had no idea it was the same summer that the famous “Attention Is All You Need” paper, which introduced the Transformer, was released. Years later, after earning my Ph.D., I was working at the machine learning platform Hugging Face when ChatGPT captured the world’s attention.
This brings us to the topic of the book: reinforcement learning from human feedback (RLHF). RLHF burst onto the scene following the release of ChatGPT, serving as the crucial added technique that transformed GPT-3.5 into the ChatGPT we fell in love with. Over the last few years, ...
Become an O’Reilly member and get unlimited access to this title plus top books and audiobooks from O’Reilly and nearly 200 top publishers, thousands of courses curated by job role, 150+ live events each month,
and much more.
Read now
Unlock full access