Description
You will build healthcare applications and end-to-end business intelligence solutions within an Agile team.
Responsibilities
- Design and develop ETL pipelines and data warehousing solutions using Databricks and AWS.
- Build and maintain scalable data models and data processing workflows using Apache Spark and Delta Lake.
- Integrate data from relational databases, APIs, and cloud storage into centralized warehouses.
- Develop interactive dashboards and automated reports using Power BI to provide stakeholder insights.
- Implement AWS services including S3 and AWS DMS to support CDC and analytics workloads.
- Ensure data accuracy, security, and compliance with governance policies through RBAC and encryption.
Required Skills
- 5+ years of experience in data engineering roles.
- Strong proficiency in SQL, Python, and PySpark.
- Hands-on experience with Apache Spark and Databricks Delta Lake.
- Expertise in ETL development, data modeling, and data warehousing.
- Experience with AWS services including S3, Lambda, AWS DMS, and API Gateway.
- Proficiency in building and maintaining Power BI dashboards.
- Experience with Terraform for infrastructure management.
- Experience using AzureDevOps for CI/CD pipelines.
- Knowledge of on-prem to AWS data connectivity.
Preferred Skills
- Understanding of machine learning concepts and their application in BI.
- Familiarity with data governance and data quality frameworks.