Description
You will own the Core Data platform pipelines, tools, and governance standards to ensure reliability and accuracy for engineering, data science, and analytics stakeholders.
This role is on-site.
Responsibilities
- Maintain, update, and expand existing data pipelines using Airflow, Spark, and Databricks.
- Build tools and services to support data discovery, lineage, governance, and privacy.
- Collaborate with product managers, architects, and engineers to drive platform success and define best practices.
- Ensure high operational efficiency and quality of datasets to meet SLAs and project reliability.
- Maintain detailed documentation for pipeline configurations, naming conventions, and data quality requirements.
Required Skills
- 5+ years of experience in distributed data processing or software engineering of data services.
- Strong proficiency with Databricks, Spark, Delta Lake, and Kubernetes.
- Deep understanding of AWS and infrastructure as code.
- Familiarity with Airflow for workflow orchestration.
- Strong background in data modeling, data warehousing methodologies, and OLTP vs OLAP environments.
- Bachelor’s degree in Computer Science, Information Systems, or equivalent industry experience.
- Excellent algorithmic problem-solving and analytical skills.
Preferred Skills
- Experience with Snowflake.
- Familiarity with Agile/Scrum methodologies.