Description
You will build and maintain scalable data infrastructure and ETL pipelines to support large-scale data processing.
Responsibilities
- Design, implement, and optimize ETL processes to ingest and transform large datasets into data warehouses.
- Develop and maintain data models, schemas, and databases for reporting and visualization.
- Perform data cleansing, validation, and quality assurance to ensure accuracy and consistency.
- Monitor and optimize the performance, scalability, and reliability of data systems.
- Implement data governance policies to ensure security, privacy, and regulatory compliance.
Required Skills
- 5+ years of experience in a Data Engineer or similar role.
- Bachelor's degree in Computer Science, Engineering, or a related field.
- Strong programming proficiency in Python and SQL.
- Deep understanding of data warehousing technologies including AWS Redshift, Google BigQuery, Snowflake, or Azure.
- Hands-on experience with ETL tools such as SSIS, Azure Data Factory, Apache Spark, Talend, or Informatica.
- Expertise in data modeling, database design, and SQL query optimization.
- Proven ability to collaborate with cross-functional teams to define data requirements.
Preferred Skills
- Familiarity with data visualization tools like Tableau or Power BI.