Quiz Space

Large Language Models Quiz 2: 12 April 2026 (January 2026 term)

Question 1

+2 marksOne correct option
  1. A
  2. B
  3. C
  4. D

Question 2

+2 marksOne correct option
  1. A

    To reduce the computational cost of the softmax layer during training.

  2. B
  3. C
  4. D

    To act as a regularizer similar to Dropout.

Question 3

+2 marksOne correct option

Why is the standard GPT architecture (Decoder-only with causal masking) generally unsuitable for the Masked Language Modeling (MLM) objective as implemented in BERT?

  1. A

    GPT models are too small to learn bidirectional contexts.

  2. B

    The causal mask in GPT prevents the model from attending to future tokens, making it impossible to use right-side context to predict a masked token.

  3. C

    GPT does not have positional embeddings, which are required for MLM.

  4. D

    GPT uses ReLU activation, while BERT uses GELU, which is required for MLM.

18 more questions in this paper

Sign in with Google — it is free — to see every question with its answer and explanation, practise it in learning mode, or take it as a timed mock test.

More on the LLM Quiz 2 12 Apr 2026 paper

The IIT Madras BS Large Language Models (LLM) Quiz 2 paper sat on 12 Apr 2026, in the January 2026 term: 21 questions for 50 marks in 120 minutes. The first 3 questions are below. Sign in with Google — it is free — to see the whole paper with its answers and explanations, in learning mode or as a timed mock test.

FeatureLLM Quiz 2 12 Apr 2026 at a glance
TermJanuary 2026 term
SubjectLarge Language Models
Course codeBSDA5004
Questions21
Marks50
Duration120 min
MCQ10
MSQ6
Numerical5
Official paperLarge Language Models 07 Apr 26
Negative markingNo negative marking.
Updated

Same Quiz 2, other subjects

More LLM