Opening the paper…
Deep Learning for Computer Vision, End Term
An RGB image of size 128×96 is stored with 32-bit floating-point values for every channel. Ignoring metadata, how many bytes are required?
An RGB image of size 128×96 is stored with 32-bit floating-point values for every channel. Ignoring metadata, how many bytes are required? Let a 1-D filter be h=\[2,-1,3\]. A library routine performs cross-correlation. Which filter should be supplied to the routine to reproduce mathematical convolution with h, ignoring boundary effects? At a point on an ideal straight intensity edge, the local image-gradient vector is most naturally interpreted as pointing: