You will lead the architecture, design, and deployment of core data platforms, data warehouses, and data modeling frameworks.
Responsibilities
Lead the delivery and deployment of core data platforms and data warehouse architectures.
Design and execute data governance and data quality frameworks.
Build and manage data pipelines covering ingestion, validation, transformation, and storage.
Decompose high-level requirements into clear engineering objectives and detailed specifications.
Interface with product, laboratory, web services, and data science teams to drive solutions.
Required Skills
B.S./M.S. in a quantitative field (Computer Science, Engineering, Mathematics, Physics, or Computational Biology) with 6+ years of experience, or a Ph.D. with 4+ years of experience.
Substantial experience architecting scalable cloud-based data warehouses or data lakes on AWS, Azure, or GCP.
Expertise in data modeling principles, including facts, dimensions, and SCDs.
Proficiency in Python and Go.
Experience with data pipelining and workflow engines such as Apache Airflow and Spark.
Experience provisioning AWS infrastructure using Terraform or CloudFormation.
Strong knowledge of relational databases, query authoring, and performance tuning.
Experience implementing CI/CD for automation in cloud environments.
Preferred Skills
Experience building microservices, web applications, and supporting machine learning pipelines.
Proven track record managing data models for hundreds to thousands of tables.
Experience with DevOps, containerized deployment, and infrastructure as code.