LLM End Term: 22 December 2024, Set QDB4 (September 2024 term)
Question 1
+3 marksNumerical answer
The input embeddings for the words “learning”, “brings” and “joy” are h1=[1.0,0.5,1], h2=[1,0.25,0], and h3=[0.1,0.1,0.9], respectively. Note that the embeddings are row vectors. The projection matrices are as follows
WQ=1−10111WK=11−1101WV=0−110−11
The following quantities are computed as
Q=HWQK=HWKV=HWV
Let ej denote the unnormalized attention score, aj denote the normalized attention score (ignore the scaling by dk) and zj denote the linear combination of the value vectors for the j−th word.
Enter the value of first element i.e. with index (0,0) of ∂e3∂a3
Question 2
+3 marksOne correct option
Assume that we have a large corpus of text. The vocabulary constructed from the text contains 10000 words. Of these, 100 words occurred only once in the entire corpus of text. The parameters of the embedding layer and the output layer of the model are shared. Suppose we create a batch of 256 samples (each sample is a sentence from the corpus). None of these samples contains any of the 100 rare words. Suppose we pre-train the model for one iteration using the batch of samples, then:
A
it is certain that the embeddings of none of these 100 rare words will get updated.
B
there is a chance that the embeddings of all or some of these 100 rare words will get updated
C
the embeddings of all these 100 rare words will defintely get updated
D
None of these
Question 3
+3 marksOne correct option
Suppose we use a pre-trained model for text generation with the given prompt “I am going to”. Which of the following decoding strategies can be used such that the pre-trained model generates same text completion each time it is executed
A
Beam search with beam size 4
B
Greedy approach
C
Top-K with k = 2
D
None of these
18 more questions in this paper
Sign in with Google — it is free — to see every question with its answer and explanation, practise it in learning mode, or take it as a timed mock test.
More on the LLM End Term 22 Dec 2024 Set QDB4 paper
The IIT Madras BS Large Language Models (LLM) End Term paper sat on 22 Dec 2024, in the September 2024 term, set QDB4: 21 questions for 50 marks in 180 minutes. The first 3 questions are below. Sign in with Google — it is free — to see the whole paper with its answers and explanations, in learning mode or as a timed mock test.