Opening the paper…
Introduction to Big Data, End Term
What differentiates “streaming processing” from “batch processing” in the context of big data?
What differentiates “streaming processing” from “batch processing” in the context of big data? What are the core components for a Big Data Streaming application? A big data streaming application that reads from using Kafka as source is observed to be really slow. The Kafka cluster has 2 broker nodes and this application is reading from 1 topic that has 10 partitions. On closer investigation, it was found that Kafka is not scaling to the velocity of input data coming in. How will you scale Kafka further?