Description
You will design and architect data monitoring and quality assurance systems to ensure high data standards across downstream products.
Responsibilities
- Design and implement data monitoring pipelines to proactively identify and resolve data quality issues.
- Develop metrics for data pipeline quality and create monitoring solutions using Python, Spark, and Airflow.
- Serve as a technical lead for data observability and federated query systems.
- Collaborate with stakeholders to define requirements and negotiate data quality SLAs.
Required Skills
- 3+ years of experience in Data Engineering or a similar role working with big data pipelines and analytics.
- 2+ years of hands-on experience with Apache Spark.
- 2+ years of coding experience in Python, Java, or an equivalent programming language.
- 2+ years of experience with SQL in scalable data warehouses such as BigQuery or Snowflake.
- Proficiency with cloud technologies, specifically GCP or AWS.
- Experience with Apache Airflow.
- Bachelor’s Degree in Computer Science, Engineering, or a related STEM field.