Description
You will build and maintain data engineering pipelines within the GCP ecosystem.
This role is on-site.
Responsibilities
- Develop ETL processes using strong SQL and scripting languages.
- Build reusable frameworks in Python to enhance existing data infrastructure.
- Implement and manage data workflows using Dataflow or Airflow.
- Set up and manage GCP IAM configurations.
- Own infrastructure definition and deployment using Terraform within CI/CD pipelines.
Required Skills
- 4+ years of Information Technology experience.
- Hands-on experience with GCP for data engineering tasks.
- Proficiency in Python, Scala, Java, or R.
- Experience with BigQuery, Hadoop, Hive, Spark, or Kafka.
- Strong SQL background for data transformation and querying.
- Experience with Git and GitHub for version control.
- Knowledge of core GCP services including Dataproc and Composer.
- Familiarity with CI/CD pipeline concepts.
Preferred Skills
- Experience in Relational Modeling or Dimensional Modeling.
- Knowledge of Airflow DAG creation and monitoring.