First lab for Data-Intensive Computing course at KTH where we are introduced to Apache Spark MLlib and Spark SQL, Hadoop, and HBase.
-
Updated
Feb 10, 2021 - Jupyter Notebook
First lab for Data-Intensive Computing course at KTH where we are introduced to Apache Spark MLlib and Spark SQL, Hadoop, and HBase.
Second lab for Data-Intensive Computing course at KTH where we use Apache Kafka, Spark, and Cassandra to practice stream processing.
Data Intensive Computing ID2221 - Project, Essay, Review Questions
To associate your repository with the id2221 topic, visit your repo's landing page and select "manage topics."