Question 2
To reduce the computational cost of the softmax layer during training.
To act as a regularizer similar to Dropout.
To reduce the computational cost of the softmax layer during training.
To act as a regularizer similar to Dropout.
Correct answer
Question 2 of 21 in the IIT Madras BS Large Language Models (LLM) Quiz 2 paper sat on 12 Apr 2026, in the January 2026 term (Large Language Models 07 Apr 26). It carries 2 marks.