July 2024
Intermediate to advanced
650 pages
17h 23m
English
This is an exciting chapter. It starts from the foundations and builds toward one of the most exciting uses of RL. Have you have used ChatGPT or another Large Language Model (LLM) and found it amazing how these models seem to follow your prompts and complete a task that you describe in English? Apart from the machinery of generative AI and transformers-driven architecture, RL also plays a very important role. Proximal Policy Optimization (PPO) using human annotated (or machine annotated) ...
Read now
Unlock full access