Description
You will design and build scalable data architecture and ETL pipelines to support feature engineering and data analysis.
Responsibilities
- Design and develop scalable ETL architectures and performant data pipeline solutions.
- Write well-structured, efficient, and maintainable code for data processing.
- Implement data quality checks, data mapping, and preprocessing workflows.
- Develop and maintain ETL routines using orchestration tools like Airflow.
- Provide technical guidance and mentorship to junior developers.
Required Skills
- 15+ years of experience in data engineering roles.
- Expertise in Python and SQL.
- Hands-on experience with PySpark.
- Proficiency in AWS environments.
- Strong knowledge of ETL development and data architecture.
- Experience with Airflow or similar orchestration tools.
- Ability to perform data analysis and identify gaps for feature engineering.
- Proven track record of developing and reusing data models.