Quiz Space

Reinforcement Learning Quiz 1: 16 July 2023 (May 2023 term)

Question 1

+3 marksOne correct option
  1. A
  2. B
  3. C
  4. D

Question 2

+3 marksOne correct option
  1. A
  2. B
  3. C
  4. D

Question 3

+3 marksOne correct option

In the context of a multi-armed bandit problem with stationary reward distributions, consider the following:
Assertion: UCB minimizes the regret better than the softmax approach.
Reason: Softmax approach assigns a low probability of picking a sub-optimal arm that has a very low expected reward.

  1. A

    Assertion and Reason are both true and Reason is a correct explanation of the Assertion.

  2. B

    Assertion and Reason are both true and Reason is not a correct explanation of the Assertion.

  3. C

    Assertion is true and Reason is false

  4. D

    Both Assertion and Reason are false.

16 more questions in this paper

Sign in with Google — it is free — to see every question with its answer and explanation, practise it in learning mode, or take it as a timed mock test.

More on the Reinforcement Learning Quiz 1 16 Jul 2023 paper

The IIT Madras BS Reinforcement Learning (Reinforcement Learning) Quiz 1 paper sat on 16 Jul 2023, in the May 2023 term: 19 questions for 50 marks in 120 minutes. The first 3 questions are below. Sign in with Google — it is free — to see the whole paper with its answers and explanations, in learning mode or as a timed mock test.

FeatureReinforcement Learning Quiz 1 16 Jul 2023 at a glance
TermMay 2023 term
SubjectReinforcement Learning
Course codeBSDA5007
Questions19
Marks50
Duration120 min
MCQ5
MSQ3
Numerical11
Official paperIIT M DEGREE AN2 EXAM QPE2 16 JULY 2023
Negative markingNo negative marking.
Updated

Same Quiz 1, other subjects

More Reinforcement Learning