← Back to jobs

Saransh Inc Logo
PySpark Developer

Saransh Inc

 

Wilmington, DE, USA

Posted On: 13 days ago
Experience: 5+ years
Availability: Onsite
Openings: 1
Category: PySpark Developer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will build and maintain distributed data processing pipelines using Spark and AWS services.

This role is on-site.

Responsibilities

  • Build Spark Streaming processes and manage Hadoop clusters and their associated services.
  • Design, install, configure, and support Hadoop environments.
  • Develop High-Level Design (HLD) and Low-Level Design (LLD) mapping specifications.
  • Integrate with internal Archival Service Platforms for data purging and lifecycle management.
  • Collaborate with data engineering teams to improve integration pipelines and adapt to business needs.

Required Skills

  • 5+ years of experience in data engineering or related roles.
  • Proficiency in Spark and Spark Streaming.
  • Hands-on experience with Hadoop cluster management.
  • Experience with Kafka and messaging systems.
  • Knowledge of NoSQL databases.
  • Experience using AWS services including S3, Athena, and Glue.
  • Strong understanding of distributed computing principles.

Preferred Skills

  • Understanding of cloud technology and data lifecycle management.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs