LM head / Unembedding
The final layer that converts the model's internal representation into a probability for every token in the vocabulary.
The final layer that converts the model's internal representation into a probability for every token in the vocabulary. (M01)