Quiz Space

Introduction to Big Data End Term: 30 April 2023 (January 2023 term)

Question 1

+3 marksOne or more correct options

What differentiates “streaming processing” from “batch processing” in the context of big data?

Select all that apply.

  1. A

    Batch operates on a set of data elements taken together while streamingoperates on data individually as it streams in

  2. B

    Batch operates on data that is static while streaming operates on data that isdynamically changing

  3. C

    Batch processing can assume data as fully specified and complete whilestreaming cannot make that assumption

  4. D

    Batch processing is typically high latency while streaming processing isnecessary for real-time latencies

  5. E

    Batch processing can operate on massively larger data sets than streamingcan

Question 2

+3 marksOne or more correct options

What are the core components for a Big Data Streaming application?

Select all that apply.

  1. A

    Data needs to be retrievable from a persistent store that supports messagereplay. Replay refers to being able to fetch a specific set of data from that persistent store on-demand., where the data is chosen based on filters usually defined on sequence numbers ortimestamps

  2. B

    Data processing needs to be splittable across machines using divide-and-conquer, and composable into steps that execute very quickly

  3. C

    Hadoop needs to be setup for its big data capabilities

  4. D

    The application needs to have the cloud-native properties of being resilient,manageable and observable

  5. E

    Application needs to be deployable on public cloud natively using PaaScomponents

Question 3

+3 marksOne or more correct options

A big data streaming application that reads from using Kafka as source is observed to be really slow. The Kafka cluster has 2 broker nodes and this application is reading from 1 topic that has 10 partitions. On closer investigation, it was found that Kafka is not scaling to the velocity of input data coming in. How will you scale Kafka further?

Select all that apply.

  1. A

    Increase memory in each of the brokers in the cluster

  2. B

    Add disks to each broker in the cluster

  3. C

    Add new brokers to the cluster

  4. D

    Create more topics and change input application to reroute data to all topics tobe able to spread input data better

  5. E

    Double the number of partitions for this single topic to be able to spread inputdata better

16 more questions in this paper

Sign in with Google — it is free — to see every question with its answer and explanation, practise it in learning mode, or take it as a timed mock test.

More on the Intro to Big Data End Term 30 Apr 2023 paper

The IIT Madras BS Introduction to Big Data (Intro to Big Data) End Term paper sat on 30 Apr 2023, in the January 2023 term: 19 questions for 45 marks in 180 minutes. The first 3 questions are below. Sign in with Google — it is free — to see the whole paper with its answers and explanations, in learning mode or as a timed mock test.

FeatureIntro to Big Data End Term 30 Apr 2023 at a glance
TermJanuary 2023 term
SubjectIntroduction to Big Data
Course codeBSDA5001
Questions19
Marks45
Duration180 min
MSQ5
MCQ14
Official paperIIT M DEGREE ET1 EXAM QPE2 S1 30 Apr 2023
Negative markingNo negative marking.
Updated

Same End Term, other subjects

More Intro to Big Data