You will build and maintain scalable data pipelines using distributed frameworks.
Responsibilities
- Implement batch and streaming ETL pipelines using Apache Spark.
- Develop microservices and APIs using Java Spring Boot.
- Design and deploy scalable data architectures using parallel distributed frameworks.
- Apply microservice architecture and design patterns to application development.
Required Skills
- 5+ years of experience in data engineering or software development.
- In-depth programming knowledge of Scala and Java.
- Hands-on experience with Apache Spark for ETL processes.
- Experience developing APIs with Java Spring Boot.
- Proficiency with SQL and NoSQL databases.
- Practical exposure to AWS cloud services.
- Solid grasp of microservice architecture.
- Strong understanding of data structures and algorithms.
Preferred Skills
- Experience with Hive.
- Containerization and orchestration using Docker and Kubernetes (Spark on Kubernetes).