← Back to jobs

Prophecy Technologies Logo
Data Engineer

Prophecy Technologies

 

Irvine, CA, USA

Posted On: 10 days ago
Experience: 5+ years
Availability: Hybrid
Openings: 1
Category: Data Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will build and maintain data pipelines, optimize large-scale processing jobs, and support data migration efforts.

This role is on-site.

Responsibilities

  • Develop Hive queries using joins and partitions to process large datasets and load filtered data.
  • Build Spark jobs using Scala and Python (PySpark) APIs, utilizing Spark SQL for structured data creation.
  • Implement shell scripts to schedule full and incremental loads and automate metadata synchronization.
  • Perform performance tuning for Spark applications, focusing on parallelism and memory optimization.
  • Troubleshoot connectivity and performance issues for production-critical jobs and validate data.

Required Skills

  • 5+ years of experience in data engineering.
  • Proficiency in Python and Scala.
  • Hands-on experience with HiveQL, Hive, and HDFS.
  • Experience developing Spark applications using PySpark and Scala.
  • Knowledge of GCP services including Google Cloud Storage (GCS) and GCP Hive.
  • Ability to develop shell scripts for automation and scheduling.
  • Experience with Druid ingestion and data flattening.

Preferred Skills

  • Experience with BigQuery.

Education

Master’s degree in Computer Science, Computer Engineering, Data Analytics,

Related Jobs

No related jobs found

← Back to jobs