Streaming Analytics and Spark
摘要
This chapter discusses the generic streaming data processing architecture and Apache Hadoop stack components Kafka, Flume and Spark that support main stages of the streaming data processing. Simple programming examples are presented. The chapter also refers to the popular Spark libraries and platforms such as Spark MLlib and Databricks Spark.