← Back to jobs
Charlotte, NC, USA
No related jobs found
Key Responsibilities:
Design, develop, and optimise large-scale data pipelines using PySpark and Python.
Implement and adhere to best practices in object-oriented programming to build reusable, maintainable code.
Write advanced SQL queries for data extraction, transformation, and loading (ETL).
Collaborate closely with data scientists, analysts, and stakeholders to gather requirements and translate them into technical solutions.
Troubleshoot data-related issues and resolve them in a timely and accurate manner.
Leverage AWS cloud services (e.g., S3, EMR, Lambda, Glue) to build and manage cloud-native data workflows (preferred).
Participate in code reviews, data quality checks, and performance tuning of data jobs.
Required Skills & Qualifications:
Relevant experience in a data engineering or backend development role.
Experience working with AWS cloud ecosystem (S3, Glue, EMR, Redshift, Lambda, etc.). Kafka
Strong hands-on experience with PySpark and Python, especially in designing and implementing scalable data transformations.
Solid understanding of Object-Oriented Programming (OOP) principles and design patterns.
Proficient in SQL, with the ability to write complex queries and optimise performance.
Strong problem-solving skills and the ability to troubleshoot complex data issues independently.
Excellent communication and collaboration skills.
Hands-on experience with AI Tools
Bachelor's degree
No related jobs found
← Back to jobs