LLM End Term 31 Aug 2025 — Question 7
Show answer
Correct answer: 6
Question 7 of 17 in the IIT Madras BS Large Language Models (LLM) End Term paper sat on 31 Aug 2025, in the May 2025 term (IIT M DEGREE AN EXAM QDB3 31 Aug 2025). It carries 2 marks.
More questions from this paper
- Given the input string: sunshine And the following vocabulary of subword tokens with their corresponding log-probabilit…
- A GPT model is trained using causal language modeling. During training, for a sequence of T = 4 tokens, which of the fo…
- Figure question
- Which of the following preprocessing steps are commonly used when preparing data for transformer models like BERT?
- Which modifications are used in transformers to handle longer sequences efficiently?
- In Transformer models that use relative position embeddings (such as in Transformer-XL or T5), clipping is often applie…
- For the input “you enjoy tea often”, compute the final representation of the word “enjoy” after the attention layer (i.…
- Suppose the input sentence is “tea you enjoy often”. Using the same matrices and processing method, what is the attenti…
- Choose the correct representation of π for the given mask image.
- How many permutations of π are possible for n = 5?
- What percentage of entries in the attention matrix are zero (i.e., the sparsity)?
- Consider the short sentence: “small models sometimes beat big ones” A model processes this sequence using the naive rel…
- Consider the short sentence: “small models sometimes beat big ones” A model processes this sequence using the naive rel…
- Consider the short sentence: “small models sometimes beat big ones” A model processes this sequence using the naive rel…
- Consider the short sentence: “small models sometimes beat big ones” A model processes this sequence using the naive rel…
- Consider the short sentence: “small models sometimes beat big ones” A model processes this sequence using the naive rel…