← Back to jobs

Oblige IT Solutions Logo
Data Engineer - Python/PySpark

Oblige IT Solutions

 

Hyderabad, Telangana, India

Posted On: 30+ days ago
Experience: 5+ years
Availability: Onsite
Openings: 2
Category: Data Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

Design, develop, and optimize large-scale data pipelines and ETL workflows within an AWS environment.

This role is on-site.

Responsibilities

  • Build and optimize scalable ETL workflows using PySpark and Apache Spark.
  • Leverage AWS services including EMR, S3, Lambda, and Glue to manage data pipelines.
  • Orchestrate and schedule complex workflows using Apache Airflow.
  • Write efficient SQL queries for data processing, validation, and troubleshooting production issues.
  • Integrate data from diverse sources while ensuring high quality, consistency, and reliability.

Required Skills

  • 5+ years of experience in data engineering.
  • Strong proficiency in Python and PySpark.
  • Hands-on experience with Apache Spark and SQL.
  • Deep knowledge of AWS services: EMR, S3, Lambda, and Glue.
  • Experience orchestrating pipelines with Apache Airflow.
  • Proven ability to work with large datasets and optimize Spark jobs.
  • Experience building data lakes and data warehouses on AWS.
  • Working knowledge of DynamoDB.

Preferred Skills

  • Experience with Kafka or other real-time streaming platforms.
  • Familiarity with DevOps tools like Terraform or CloudFormation.
  • Exposure to NoSQL databases such as MongoDB.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs