Description
You will build and maintain ETL pipelines to transform raw data into usable formats for analysis.
This role is hybrid.
Responsibilities
- Transform raw data through reshaping, aggregating, and normalizing datasets.
- Clean and preprocess data using Python, SQL, and Excel to resolve quality issues.
- Execute data profiling to identify patterns, anomalies, duplicates, and missing values.
- Ensure data integrity and consistency across APIs, logs, and databases.
Required Skills
- 5+ years of experience in data engineering and ETL development.
- Proficiency with ETL tools such as IICS, Talend, or Hadoop.
- Strong SQL, PL/SQL, and Oracle database skills.
- Hands-on experience with PySpark and Shell scripting.
- Experience with Azure environments and basic understanding of GCP and AWS.
- Ability to clean and preprocess data using Python and Excel.
- Experience with Power BI for data visualization.
- Proficiency using JIRA, Confluence, and MS Visio.
Preferred Skills
- Knowledge of R programming.