Quiz Space

Deep Learning Quiz 2: 6 August 2023 (May 2023 term)

Question 1

+3 marksOne or more correct options

Suppose that a neural network has millions of parameters (weights and biases). A team decides to use an optimization algorithm with a learning rate scheme that is local to each individual parameters in the network. Moreover, the learning rate changes in each iteration based on the magnitude of gradients pertaining to a parameter in the past. Which of the following optimization algorithms satisfy the team’s requirements?

Select all that apply.

  1. A

    GD with an exponentially decaying learning rate scheduler

  2. B

    AdaGrad

  3. C

    AdaM

  4. D

    NADAM

  5. E

    RMSProp

  6. F

    SGD with line search

Question 2

+4 marksOne or more correct options

Suppose a data set has NN samples and each sample has two features f1∈{0,1}f_1 \in \{0, 1\} and f2∈{0,1}f_2 \in \{0, 1\}. Assume that f1f_1 is a sparse feature and f2f_2 is a dense feature (that is, f1f_1 value for most of the samples is zero). Further, we apply Stochastic Gradient Descent on this data and plot a contour plot of the loss surface and observe the movement of parameters (i.e., trajectory) across iterations. Let w1w_1 and w2w_2 be the parameters corresponding to f1f_1 and f2f_2 respectively.

Assume bias to be zero for this question and the parameters are initialized to zero.

In the contour plot of a loss surface, what is the initial movement expected to look like if w1w_1 is plotted on the horizontal axis and w2w_2 is plotted on the vertical axis?

Hint: Think of the possible configurations of inputs for the first few iterations

Select all that apply.

  1. A

    Zig-zag movement along the horizontal axis

  2. B

    Zig-zag movement along the vertical axis

  3. C

    Mostly straight along the horizontal axis

  4. D

    Mostly straight along the vertical axis

  5. E

    Insufficient data

Question 3

+4 marksOne or more correct options

Select all that apply.

  1. A

    ADAM takes more memory than AdaGrad

  2. B

    ADAM takes lesser memory than AdaGrad

  3. C

    AdaGrad takes more memory than RMSprop

  4. D

    Vanilla GD takes lesser memory than AdaGrad

12 more questions in this paper

Sign in with Google — it is free — to see every question with its answer and explanation, practise it in learning mode, or take it as a timed mock test.

More on the Deep Learning Quiz 2 6 Aug 2023 paper

The IIT Madras BS Deep Learning (Deep Learning) Quiz 2 paper sat on 6 Aug 2023, in the May 2023 term: 15 questions for 48 marks in 120 minutes. The first 3 questions are below. Sign in with Google — it is free — to see the whole paper with its answers and explanations, in learning mode or as a timed mock test.

FeatureDeep Learning Quiz 2 6 Aug 2023 at a glance
TermMay 2023 term
SubjectDeep Learning
Course codeBSCS3004
Questions15
Marks48
Duration120 min
MSQ7
Numerical8
Official paperIIT M DEGREE AN3 EXAM QPE3 06 Aug 2023
Negative markingNo negative marking.
Updated

Same Quiz 2, other subjects

More Deep Learning