Description
You will build and maintain data pipelines and architecture within healthcare data environments.
Responsibilities
- Develop and optimize ETL processes using Python and PySpark.
- Design and implement data lake architectures and AWS Glue workflows.
- Manage healthcare data integration using HL7/FHIR standards and APIs.
- Maintain data security protocols for healthcare eligibility and claims data.
- Automate infrastructure and deployments using Terraform and CI/CD processes.
Required Skills
- 3+ years of experience in data engineering.
- Mastery of Python, PySpark, and SQL.
- Experience with AWS products, specifically AWS Glue.
- Knowledge of data lakes and ETL scheduling solutions.
- Hands-on experience with HL7/FHIR standards and healthcare data security.
- Practical use of GIT and CI/CD processes.
- Experience with Infrastructure as Code using Terraform or Cloudformation.
- Background in implementing APIs for data movement.
- Any Graduate degree.
Preferred Skills
- Experience working across both on-prem and cloud environments.
- Familiarity with Business Intelligence and Analysis tools.