Deep Learning for Computer Vision End Term 13 Apr 2025 — Question 20
Show answer
Correct answer: -0.21 (accepted within ±0.005)
Question 20 of 47 in the IIT Madras BS Deep Learning for Computer Vision (Deep Learning for Computer Vision) End Term paper sat on 13 Apr 2025, in the January 2025 term (IIT M IMPROVEMENT FN EXAM QIM2 13 Apr). It carries 2 marks.
More questions from this paper
- Why does DETR typically exhibit poor performance in detecting small objects compared to larger ones? Choose the best an…
- Which one of the following statements is false?
- What is the correct order of operations for processing an image through a Vision Transformer (ViT)?
- Vector Quantized Variational Autoencoder (VQ-VAE) utilizes a discrete latent representation as opposed to continuous la…
- Given two normalized embeddings from CLIP, I = [0.2,−0.5, 0.3, 0.4] for an image and T = [−0.1, 0.6, −0.3,−0.7] for a t…
- Which one of the following statements regarding hyperparameter tuning is false?
- Which one of the following statements is false? (Pick the most appropriate one.)
- Figure question
- Figure question
- Figure question
- In which one of the following applications would you use a one- to-many RNN architecture? Choose the most appropriate a…
- Why might Segment Anything (SAM) be particularly useful in data annotation tasks compared to traditional segmentation m…
- Which of the following techniques help control the exploding or vanishing gradient problem in recurrent neural networks?
- Which of the following statements are true? (Select all possible correct options)
- Which of the following statements are true ? (Select all possible correct options)
- Which of the following statements about CLIP are TRUE? (Select ALL that apply)
- Which of the following statements are true?(Select ALL that apply)
- Figure question
- Figure question
- Consider a reverse process in a diffusion model where the goal is to reconstruct the original data from the noise. If t…
- Figure question
- Figure question
- Figure question
- Based on the above data, answer the given subquestions.
- Based on the above data, answer the given subquestions.
- Element 1:
- Element 2:
- Element 3:
- Element 4:
- Consider the BLIP model architecture with an input (image-text pair) as shown below:\ For each of the given entities, e…
- Consider the BLIP model architecture with an input (image-text pair) as shown below:\ For each of the given entities, e…
- Consider the BLIP model architecture with an input (image-text pair) as shown below:\ For each of the given entities, e…
- Consider the BLIP model architecture with an input (image-text pair) as shown below:\ For each of the given entities, e…
- Consider the BLIP model architecture with an input (image-text pair) as shown below:\ For each of the given entities, e…
- Consider the BLIP model architecture with an input (image-text pair) as shown below:\ For each of the given entities, e…
- Consider the BLIP model architecture with an input (image-text pair) as shown below:\ For each of the given entities, e…
- Consider the BLIP model architecture with an input (image-text pair) as shown below:\ For each of the given entities, e…
- Consider the BLIP model architecture with an input (image-text pair) as shown below:\ For each of the given entities, e…
- Consider the BLIP model architecture with an input (image-text pair) as shown below:\ For each of the given entities, e…
- Consider the BLIP model architecture with an input (image-text pair) as shown below:\ For each of the given entities, e…
- Sigmoid
- Linear
- Indicator Function _
- Softplus _
- Relu _
- Leaky-Relu