Description
Build and maintain data pipelines on Google Cloud Platform, focusing on BigQuery, Fivetran, and dbt to ensure data quality and accessibility.
This role is on-site.
Responsibilities
- Design and maintain ELT/ETL pipelines using BigQuery, configuring and extending Fivetran connectors, including custom development.
- Develop and manage dbt models, tests, and documentation to guarantee data lineage and quality.
- Optimize BigQuery tables for cost and performance through partitioning, clustering, and materialization.
- Monitor pipeline health, implement alerting, and resolve incidents to maintain data freshness.
- Partner with the Senior Data Architect on data governance, including cataloguing, certification, and access controls.
Required Skills
- 4+ years of experience in data engineering or related roles.
- Strong proficiency in SQL and Python for pipeline development and debugging.
- Hands-on experience with BigQuery or other cloud data warehouses.
- Experience with ELT/ETL tools such as Fivetran, Airbyte, or custom frameworks.
- Proficiency with dbt for transformations and testing.
- Experience with workflow orchestration tools like Airflow, Cloud Composer, or Prefect.
- Familiarity with GCP services (Cloud Storage, Pub/Sub, Cloud Functions) or equivalent AWS/Azure services.
- Strong understanding of data modeling concepts including star schema, SCDs, and normalization.
Preferred Skills
- Experience with Terraform or other IaC tools for infrastructure management.
- Experience integrating Salesforce, HubSpot, or Stripe data sources.