Description
You will manage large-scale data sets within a data lake environment and optimize processing workflows.
Responsibilities
- Manage large data sets within a data lake environment.
- Optimize data processing workflows using Spark.
- Perform performance tuning on Spark applications.
- Execute complex data management tasks using SQL and Python.
Required Skills
- 10+ years of experience in the data management space.
- 5+ years of experience working with large data sets in a data lake environment.
- Highly proficient in SQL.
- Solid understanding of Spark, including performance tuning.
- Solid understanding of the AWS platform.
- Professional experience with Python.
- Bachelor's degree or equivalent graduate qualification.