Quiz Space

Large Language Models · Quiz 2 · 16 Mar 2025 · January 2025 term

Question 5: The strikeout words in the passage given below denote the…

Question 5

+3 marksOne or more correct options

The strikeout words in the passage given below denote the words to be dropped from the original sentence.

“These models can't read your mind. If outputs are too long, ask for brief replies. If outputs are too simple, ask for expert-level writing. If you dislike the format, demonstrate the format you'd like to see. The less the model has to guess at what you want, the more likely you'll get it.”

Which of the following represents the correct input sequence, with sentinel tokens, to the baseline model that uses pre-training denoising objectives? [z] represents the end of the sentinel token in a sentence. The characters inside the square brackets are the sentinel tokens.

Select all that apply.

  1. A

    [a] your mind [b] ask for brief replies [c] demonstrate the format [d] less the model has to guess [e]

  2. B

    [a] your mind [b] ask for brief replies [c] demonstrate the format [d] less the model has to guess

  3. C

    [v] your mind [w] ask for brief replies [x] demonstrate the format [y] less the model has to guess

  4. D

    [v] your mind [w] ask for brief replies [x] demonstrate the format [y] less the model has to guess [z]

  5. E

    [a] your mind [b] demonstrate the format [c] ask for brief replies [d] less the model has to guess

  6. F

    None of these

Show answer

Correct answers

  • B

    [a] your mind [b] ask for brief replies [c] demonstrate the format [d] less the model has to guess

  • C

    [v] your mind [w] ask for brief replies [x] demonstrate the format [y] less the model has to guess

Question 5 of 18 in the IIT Madras BS Large Language Models (LLM) Quiz 2 paper sat on 16 Mar 2025, in the January 2025 term (IIT M DEGREE AN EXAM QDB2 16 Mar 2025). It carries 3 marks.

More questions from this paper

  1. Q1Consider following assertion and reason pair:\ Assertion: A small vocabulary is desirable, as it reduces the size of th…
  2. Q2Consider following statement and mark if it is true or false:\ Repeating examples in the pre-training datasets are comp…
  3. Q3What is the order of the language modeling pipeline?
  4. Q4Consider fine-tuning the T5 model on a sentiment classification task. The T5 model was pre- trained using the C4 datase…
  5. Q6Suppose the T5 base model has to be fine tuned with gradual unfreezing. First layer of the encoder takes the input in a…
  6. Q7Which of the following models uses a form of denoising objective for pre-training?
  7. Q8Consider the C4 pipeline. Which of the following will NOT pass through it as it is?
  8. Q9Which of the following are correct about GPT-3 model with 175B parameters:
  9. Q10Consider a vocabulary \mathcal{V}, \mathcal{V} =([start], deep, hot, is, learning, research, topic, very, [end]). Assum…
  10. Q11Figure question
  11. Q12Consider the following dictionary with the number of word occurrences in a corpus: Note: Append </w> to each word…
  12. Q13Consider the following dictionary with the number of word occurrences in a corpus: Note: Append </w> to each word…
  13. Q14Consider the following dictionary with the number of word occurrences in a corpus: Note: Append </w> to each word…
  14. Q15Consider the following dictionary with the number of word occurrences in a corpus: Note: Append </w> to each word…
  15. Q16Consider the following dictionary with the number of word occurrences in a corpus: Note: Append </w> to each word…
  16. Q17Consider the following dictionary with the number of word occurrences in a corpus: Note: Append </w> to each word…
  17. Q18Consider the following dictionary with the number of word occurrences in a corpus: Note: Append </w> to each word…