Training Objectives

Next-token prediction, masked LM, span corruption, fill-in-the-middle, and auxiliary losses.

14concepts
76flashcards
98minutes of reading