Quiz Space

Reinforcement Learning End Term: 1 September 2024, Set QDB1 (May 2024 term)

Question 1

+2 marksOne correct option

Consider the following assertion reason pair:
Assertion: Reinforcement learning is a type of unsupervised learning algorithm as both don’t have correct labels.
Reason: In unsupervised learning, a reward like quantity is not maximized.

  1. A

    Assertion and Reason are both true and Reason is a correct explanation of Assertion.

  2. B

    Assertion and Reason are both true and Reason is not a correct explanation of Assertion.

  3. C

    Assertion is true but Reason is false.

  4. D

    Assertion is false but Reason is true.

Question 2

+2 marksOne correct option
  1. A
  2. B
  3. C
  4. D
  5. E

Question 3

+2 marksOne correct option

Consider a reinforcement learning agent navigating a grid world environment. The agent receives rewards of +1 for reaching the goal state and 0 otherwise. Which of the following statements accurately describes the differences between Monte Carlo and TD learning in this scenario?

  1. A

    Monte Carlo updates are unbiased estimators of the true value function, while TD updates may introduce bias.

  2. B

    TD updates are guaranteed to converge to the optimal value function, while Monte Carlo updates may not converge.

  3. C

    Monte Carlo updates require less memory and computational resources compared to TD updates.

  4. D

    TD updates are more robust to noise and stochasticity in the environment compared to Monte Carlo updates.

  5. E

    None of these

19 more questions in this paper

Sign in with Google — it is free — to see every question with its answer and explanation, practise it in learning mode, or take it as a timed mock test.

More on the Reinforcement Learning End Term 1 Sept 2024 Set QDB1 paper

The IIT Madras BS Reinforcement Learning (Reinforcement Learning) End Term paper sat on 1 Sept 2024, in the May 2024 term, set QDB1: 22 questions for 50 marks in 180 minutes. The first 3 questions are below. Sign in with Google — it is free — to see the whole paper with its answers and explanations, in learning mode or as a timed mock test.

FeatureReinforcement Learning End Term 1 Sept 2024 Set QDB1 at a glance
TermMay 2024 term
SubjectReinforcement Learning
Course codeBSDA5007
Questions22
Marks50
Duration180 min
MCQ14
MSQ3
Numerical5
Official paperIIT M DEGREE AN EXAM QDB3 01 Sep 2024
Negative markingNo negative marking.
Updated

Other sets that day

Same End Term, other subjects

More Reinforcement Learning