Description
You will build and maintain scalable data pipelines and automate data workflows.
Responsibilities
- Develop scalable data solutions using Databricks and Python.
- Build and maintain data pipelines in collaboration with data analysts, data scientists, and DevOps teams.
- Automate data workflows and manage CI/CD pipelines.
- Utilize AWS services to support data architecture and processing.
Required Skills
- 5+ years of experience in data engineering.
- Strong hands-on experience with Databricks (notebooks, jobs, clusters, workflows).
- Proficiency in Python for data manipulation, APIs, and scripting.
- Deep knowledge of Apache Spark and PySpark.
- Hands-on experience with the AWS ecosystem: S3, Glue, Lambda, Redshift, and CloudWatch.
- Experience with CI/CD tools such as Git, Jenkins, GitHub Actions, or Azure DevOps.
- Degree in any field (Any Graduate).
Preferred Skills
- Familiarity with Terraform, Databricks CLI, or REST APIs.