Skip to content

Exploring Spark Structured Streaming features by making use of Jupiter notebooks, Pyspark and interacting with a Kafka cluster.

Notifications You must be signed in to change notification settings

RosarioB/spark-streaming-kafka

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 

History

3 Commits
 
 

Repository files navigation

spark-streaming-kafka

In this repository we will read the data stored in a Kafka topic and apply some transformations on them by making use of Spark Structured Streaming. Each branch has different feature:

  • python makes use of Jupyter notebooks and pyspark
  • scala makes use of a Maven project in Scala.

About

Exploring Spark Structured Streaming features by making use of Jupiter notebooks, Pyspark and interacting with a Kafka cluster.

Topics

Resources

Stars

Watchers

Forks

Releases

No releases published

Packages

No packages published