← Back to jobs

Applab Systems Logo
Python PySpark Developer

Applab Systems

 

Pittsburgh, PA, USA

Posted On: 30+ days ago
Experience: 5+ years
Availability: Hybrid
Openings: 2
Category: Python Pyspark Developer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will build and maintain data integration pipelines within an enterprise data platform.

This role is on-site.

Responsibilities

  • Design, develop, test, and deploy data integration solutions to connect enterprise systems.
  • Build and enhance Apache Spark-based integration capabilities.
  • Maintain and improve existing data integration pipelines and standard operating procedures.
  • Author technical design documents, testing plans, and scripts.
  • Facilitate requirements gathering, process mapping workshops, and technical review sessions.

Required Skills

  • 5+ years of experience in data integration and pipeline development.
  • Extensive Python development experience, specifically using PySpark.
  • 2+ years of experience with AWS Cloud data integration.
  • Proficiency with Apache Spark, EMR, Glue, Kafka, Kinesis, and Lambda.
  • Hands-on experience with S3, Redshift, RDS, and MongoDB/DynamoDB ecosystems.
  • Strong database skills including complex query writing, optimization, debugging, UDFs, views, and indexes.
  • Experience with Git, Bitbucket, and Jenkins for continuous integration.

Preferred Skills

  • Experience with Databricks or Apache Spark.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs