Description
You will own data pipeline development and deep-dive investigations into data quality and governance issues.
Responsibilities
- Build and maintain ETL data pipelines using Hive/HDFS and Hive SQL.
- Perform data evaluation and root-cause analysis to ensure data accuracy.
- Manage customer data processes while adhering to strict compliance and governance standards.
- Implement data modeling, dimensional modeling, and metadata management practices.
- Execute statistical analysis to present actionable insights and build OLAP cubes.
Required Skills
- 5+ years of experience in Data Analysis.
- 5+ years of hands-on experience using ETL tools to create data pipelines.
- Strong background in Data Warehousing and dimensional modeling.
- Proficiency in Hive, HDFS, and Hive SQL.
- Experience with data governance, data standards, and data quality frameworks.
- Hands-on experience with business data modeling and metadata management.
- Understanding of OLAP tools and cube building.
- Basic understanding of AWS environments.
- Knowledge of encryption and decryption processes for customer PII data.
Preferred Skills
- Experience with customer data management processes and compliance.