Description
You will design and lead scalable data architectures and pipelines to support analytics and business intelligence.
Responsibilities
- Design and maintain scalable data pipelines and architectures.
- Lead data projects and enforce engineering best practices.
- Optimize data systems for analytics and reporting.
- Ensure data quality and system reliability in production environments.
- Focus on data optimization, workflow automation, and cloud-based data operations.
Required Skills
- 8+ years of IT experience with 5+ years focused on Python, PySpark, and SQL.
- Experience with data lakes using Iceberg format.
- Proficiency in ETL processes using Informatica.
- Hands-on experience with AWS services: S3, Glue, Redshift, Lambda, EMR, and Airflow.
- Working knowledge of Postgres and BASH/Shell scripting.
- Experience managing data quality and reliability.
- Background working with healthcare data.
- Experience leading data teams in an Agile development environment.