Quiz Space

Large Language Models End Term: 10 May 2026 (January 2026 term)

Question 1

+2 marksOne correct option

In the standard Transformer Decoder, the Multi-Head Attention layer is "Masked". What is the specific purpose of this mask during training?

  1. A

    To filter out padding tokens to save computation.

  2. B

    To prevent the model from attending to the [CLS] and [SEP] special tokens.

  3. C
  4. D

    To force the model to focus on the Encoder output rather than the Decoder input.

Question 2

+2 marksOne correct option
  1. A
  2. B
  3. C
  4. D

Question 3

+2 marksOne correct option
  1. A

    It has a limited receptive field and cannot capture long-range dependencies directly.

  2. B

    It requires more memory than full attention.

  3. C

    It cannot be parallelized on GPUs.

  4. D

    It causes the gradients to explode.

17 more questions in this paper

Sign in with Google — it is free — to see every question with its answer and explanation, practise it in learning mode, or take it as a timed mock test.

More on the LLM End Term 10 May 2026 paper

The IIT Madras BS Large Language Models (LLM) End Term paper sat on 10 May 2026, in the January 2026 term: 20 questions for 50 marks in 180 minutes. The first 3 questions are below. Sign in with Google — it is free — to see the whole paper with its answers and explanations, in learning mode or as a timed mock test.

FeatureLLM End Term 10 May 2026 at a glance
TermJanuary 2026 term
SubjectLarge Language Models
Course codeBSDA5004
Questions20
Marks50
Duration180 min
MCQ10
MSQ5
Numerical5
Official paperLarge Language Models 10 May 26
Negative markingNo negative marking.
Updated

Same End Term, other subjects

More LLM