January 2021
Intermediate to advanced
384 pages
8h 9m
English
In this chapter, we will build a RoBERTa model from scratch. The model will take the bricks of the Transformer construction kit we need for BERT models. Also, no pretrained tokenizers or models will be used. The RoBERTa model will be built following the fifteen-step process described in this chapter.
We will use the knowledge of transformers acquired in the previous chapters to build a model that can perform language modeling on masked tokens step by step. In Chapter 1, Getting Started with the Model Architecture of the Transformer, we went through the building blocks of the original Transformer. In Chapter 2, Fine-Tuning BERT Models, we fine-tuned a pretrained BERT model.
This chapter will focus on ...
Read now
Unlock full access