Training Objectives

Next-token prediction, masked LM, span corruption, fill-in-the-middle, and auxiliary losses.

13concepts
66flashcards
91minutes of reading