Description

You will build and maintain scalable data pipelines using distributed frameworks.

Responsibilities

  • Implement batch and streaming ETL pipelines using Apache Spark.
  • Develop microservices and APIs using Java Spring Boot.
  • Design and deploy scalable data architectures using parallel distributed frameworks.
  • Apply microservice architecture and design patterns to application development.

Required Skills

  • 5+ years of experience in data engineering or software development.
  • In-depth programming knowledge of Scala and Java.
  • Hands-on experience with Apache Spark for ETL processes.
  • Experience developing APIs with Java Spring Boot.
  • Proficiency with SQL and NoSQL databases.
  • Practical exposure to AWS cloud services.
  • Solid grasp of microservice architecture.
  • Strong understanding of data structures and algorithms.

Preferred Skills

  • Experience with Hive.
  • Containerization and orchestration using Docker and Kubernetes (Spark on Kubernetes).

Education

Any Graduate