Description
Build and maintain scalable data architectures and ETL pipelines on GCP.
Responsibilities
- Develop and maintain code-based ETL pipelines using Python and SQL.
- Manage end-to-end data engineering projects from design to deployment.
- Write maintainable code for Big Data domain applications.
- Execute plans to protect and enhance the value of data assets.
Required Skills
- 4+ years of experience with Big Data solutions on GCP.
- 3+ years of Python development experience.
- 3+ years of code-based ETL development using Python and SQL.
- 3+ years of experience writing complex SQL queries.
- 2+ years of experience with GCP services including BigQuery, Kubernetes, and Composer.
- 2+ years of experience working with Apache Airflow.
- Hands-on experience with GitHub and development IDEs like VS Code.
- B.S. or M.S. in Computer Science, Computer Engineering, or a related Engineering field.
Preferred Skills
- Experience with Spark and Dataproc.