Quiz Space

Reinforcement Learning End Term: 10 May 2026, Set S2 (January 2026 term)

Question 1

+1 markOne correct option
  1. A

    The discount factor used during option execution

  2. B
  3. C
  4. D

Question 2

+1 markOne correct option

In DDPG, exploration is typically achieved by:

  1. A
  2. B

    Randomly sampling from the replay buffer with higher priority

  3. C

    Injecting noise into the critic network parameters

  4. D

    Adding time-correlated noise (e.g., Ornstein-Uhlenbeck) or Gaussian noise to the deterministic policy output

Question 3

+1 markOne correct option
  1. A

    Reduces variance without introducing bias

  2. B

    Eliminates all bias in the gradient estimate

  3. C

    Makes the algorithm off-policy

  4. D

    Removes the need for a value function approximator

18 more questions in this paper

Sign in with Google — it is free — to see every question with its answer and explanation, practise it in learning mode, or take it as a timed mock test.

More on the Reinforcement Learning End Term 10 May 2026 Set S2 paper

The IIT Madras BS Reinforcement Learning (Reinforcement Learning) End Term paper sat on 10 May 2026, in the January 2026 term, set S2: 21 questions for 30 marks in 180 minutes. The first 3 questions are below. Sign in with Google — it is free — to see the whole paper with its answers and explanations, in learning mode or as a timed mock test.

FeatureReinforcement Learning End Term 10 May 2026 Set S2 at a glance
TermJanuary 2026 term
SubjectReinforcement Learning
Course codeBSDA5007
Questions21
Marks30
Duration180 min
MCQ8
MSQ7
Numerical6
Official paperReinforcement Learning 10 May 26 (Session 2)
Negative markingNo negative marking.
Updated

Other sets that day

Same End Term, other subjects

More Reinforcement Learning