← Back to jobs

Interon IT Logo
PySpark Data Engineer with Python

Interon IT

 

United States

Posted On: 10 days ago
Experience: 8+ years
Availability: Remote
Openings: 1
Category: PySpark Data Engineer
Tenure: Contract - Corp-to-Corp
Related Jobs

No related jobs found

Description

You will design, build, and maintain scalable data pipelines within an AWS ecosystem.

This role is remote.

Responsibilities

  • Develop and maintain data pipelines using Python and PySpark.
  • Design and implement efficient SQL queries for data extraction and manipulation.
  • Implement and manage AWS cloud services including Athena, S3, Lambda, and EventBridge.
  • Ensure data quality and integrity through rigorous testing and validation.
  • Own architectural decisions and strategic vision for data workflows.

Required Skills

  • 8+ years of experience in data engineering roles.
  • Proficiency in Python and PySpark.
  • Strong SQL skills for data transformation and analysis.
  • Hands-on experience with AWS Glue and AWS Glue Data Catalog.
  • Experience managing data in Aurora Postgres and Redshift.
  • Working knowledge of Airflow or MWAA DAGs for orchestration.
  • Experience with ECS, Docker, or distributed computing concepts like EMR.

Preferred Skills

  • Familiarity with Informatica or similar ETL tools for data integration.
  • Understanding of Lakehouse concepts and database consumption patterns.

Education

Any Gradute

Related Jobs

No related jobs found

← Back to jobs