Quiz Space

Large Language Models · Quiz 2 · 1 Dec 2024 · September 2024 term

Question 11: The strikeout words in the passage given below denote th…

Question 11

+4 marksOne or more correct options

The strikeout words in the passage given below denote the words to be dropped from the original sentence.

“Metacognition is an awareness of one's thought processes and an understanding of the patterns behind them. The term comes from the root word meta, meaning ”beyond”, or ”on top of”. Metacognition can take many forms, such as reflecting on one's ways of thinking, and knowing when and how oneself and others use particular strategies for problem-solving. There are generally two components of metacognition: cognitive conceptions and cognitive regulation system.”

Which of the following represents the correct target sequence, with sentinel tokens, to the baseline model (according to T5 study) that uses the pre-training denoising objective? The characters inside the square brackets are the sentinel tokens and [z] represents the end of the sentinel token in a sentence

Select all that apply.

  1. A

    [v] is an awareness [w] can take many forms [x] particular strategies [y] and cognitive regulation

  2. B

    [v] is an awareness [w] can take many forms [x] particular strategies [y] and cognitive regulation [z]

  3. C

    [v] is an awareness [w] can take many forms [x] and cognitive regulation [y] particular strategies

  4. D

    [v] is an awareness [w] particular strategies [x] can take many forms [y] and cognitive regulation

  5. E

    [a] is an awareness [b] can take many forms [c] particular strategies [d] and cognitive regulation

  6. F

    [a] is an awareness [b] can take many forms [c] particular strategies [d] and cognitive regulation[z]

  7. G

    [a] is an awareness [b] can take many forms [c] and cognitive regulation

  8. H

    None of these.

Show answer

Correct answers

  • B

    [v] is an awareness [w] can take many forms [x] particular strategies [y] and cognitive regulation [z]

  • F

    [a] is an awareness [b] can take many forms [c] particular strategies [d] and cognitive regulation[z]

Question 11 of 16 in the IIT Madras BS Large Language Models (LLM) Quiz 2 paper sat on 1 Dec 2024, in the September 2024 term (IIT M DEGREE AN EXAM QDB2 01 Dec 2024). It carries 4 marks.

More questions from this paper

  1. Q1Consider the following dictionary with the number of word occurrences in a corpus: Note: Append identifier/special symb…
  2. Q2Consider the following dictionary with the number of word occurrences in a corpus: Note: Append identifier/special symb…
  3. Q3Consider the following dictionary with the number of word occurrences in a corpus: Note: Append identifier/special symb…
  4. Q4Consider the following dictionary with the number of word occurrences in a corpus: Note: Append identifier/special symb…
  5. Q5Consider the following dictionary with the number of word occurrences in a corpus: Note: Append identifier/special symb…
  6. Q6Consider the following dictionary with the number of word occurrences in a corpus: Note: Append identifier/special symb…
  7. Q7Consider the following dictionary with the number of word occurrences in a corpus: Note: Identifier/special symbol <…
  8. Q8Consider the following dictionary with the number of word occurrences in a corpus: Note: Identifier/special symbol <…
  9. Q9Which of the following represent(s) the normalization step(s) in building a tokenizer for the English language?
  10. Q10Choose all the aspects of the pre-training datasets that impact the model’s performance.
  11. Q12Suppose you are working on prefix language modeling. The sequence length is 32 and the first two tokens represent the t…
  12. Q13Which of the following components in the data pre-processing pipeline removes pages that contain bad words?
  13. Q14Which of the following is an important implication of the scaling law:
  14. Q15Which of the following are the design choices for building a large language model?
  15. Q16Consider an ideal data preprocessing pipeline to prepare a dataset. Which of the following sentences or sets of paragra…