Design, develop, and manage scalable data integration solutions across hybrid and cloud environments.
This role is remote.
Responsibilities
Design and implement data integration workflows between internal and external systems, including APIs, databases, SaaS applications, and cloud platforms.
Build and maintain scalable ETL/ELT pipelines for structured and unstructured data using Informatica, Talend, SSIS, Apache NiFi, or custom Python/SQL scripts.
Develop real-time and batch data pipelines leveraging technologies like Kafka and Spark Streaming.
Implement data validation, cleansing, deduplication, and monitoring mechanisms to ensure high data quality and consistency.
Troubleshoot and resolve data integration and pipeline issues, providing documentation and knowledge transfer.
Required Skills
5+ years of experience in data integration and data engineering with strong ETL and SQL proficiency.
Strong experience with integration tools such as Informatica, Talend, SSIS, or Boomi.
Proficient in SQL and Python for data manipulation and automation.
Experience with cloud data platforms (GCP) and services such as Google Cloud Dataflow.
Familiarity with REST/SOAP APIs, JSON, XML, and flat file integrations.
Preferred Skills
Understanding of data warehousing concepts and tools (e.g., Snowflake, Redshift, BigQuery).
Knowledge of data security, privacy, and compliance best practices (HIPAA, GDPR, etc.).
Experience with message queues or data streaming platforms (Kafka, RabbitMQ, Kinesis).