Pinned Loading
-
USImmigrationDataLakeETL
USImmigrationDataLakeETL PublicThe project aims to create a data lake for US immigration data and developing an ETL pipeline to build this data lake using data from various sources. The project was completed as a part of Udacity…
Jupyter Notebook 1
-
DataPipelinesWithAirflow_Udacity
DataPipelinesWithAirflow_Udacity PublicCreated the DAGs and designed an ETL pipeline using custom operators to perform tasks such as staging the data, filling the data warehouse, and running checks on the data as the final step. This pr…
Python
-
FrequentPatternMining
FrequentPatternMining PublicComparative study of frequent pattern mining algorithm on Adult Census Data
Jupyter Notebook
-
DataWarehousingOnAWS_Udacity
DataWarehousingOnAWS_Udacity PublicThis project builds a data warehouse using Amazon Redshift and S3, and was completed as a part of Udacity's Data Engineering Nanodegree program.
Python
-
DataLakeWithSpark_Udacity
DataLakeWithSpark_Udacity PublicThis project creates a data lake using AWS S3, EMR and Spark to build an ETL pipeline for a music database.
Python
-
DataModelingWithApacheCassandra_Udacity
DataModelingWithApacheCassandra_Udacity PublicModeled song data using Apache Cassandra and designed an ETL pipeline.
Jupyter Notebook
If the problem persists, check the GitHub status page or contact support.
