This project implements a scalable ETL pipeline that processes streaming data from Kafka and performs analytics using Spark.