Skip to content

Repository files navigation

s3-data-lake-example

Creating a S3 Data lake with pyspark ETL.

First step involves using pandas to only extract the columns that are required and to create files in data lake using parquet format.

Queries to be supported The Lines where the expected vs actual arrival time is long.

About

No description or website provided.

Topics

Resources

Stars

0 stars

Watchers

0 watching

Forks

Releases

Packages

Contributors

Languages