Description
You will build and maintain scalable data pipelines and automate data workflows using Databricks and AWS.
This role is on-site.
Responsibilities
- Develop scalable data solutions using Databricks and Python.
- Build and maintain data pipelines in collaboration with data analysts, data scientists, and DevOps teams.
- Automate data workflows and manage CI/CD pipelines.
- Utilize AWS services to support data architecture and processing.
Required Skills
- 5+ years of experience in data engineering.
- Strong hands-on experience with Databricks (notebooks, jobs, clusters, workflows).
- Proficiency in Python for data manipulation, APIs, and scripting.
- Deep knowledge of Apache Spark and PySpark.
- Hands-on experience with the AWS ecosystem: S3, Glue, Lambda, Redshift, and CloudWatch.
- Experience with CI/CD tools such as Git, Jenkins, GitHub Actions, or Azure DevOps.
Preferred Skills
- Familiarity with Terraform, Databricks CLI, or REST APIs.