You will design and build scalable data pipelines and ETL applications to support advertising, content, and finance operations in Plano, TX.
Responsibilities
Design database architectures and ETL pipelines using modern technologies.
Build highly scalable data pipelines using AWS, Snowflake, Spark, and Kafka to ingest data from multiple systems.
Translate complex business requirements into technical solutions that meet data warehousing standards.
Define methodologies and standards for the data warehousing environment.
Collaborate with product management and operations to improve system performance and resolve migration issues.
Required Skills
2+ years of software or data engineering experience.
Proficiency in Python and experience building REST APIs for back-end services.
Hands-on experience with ETL processes and managing large data sets.
Experience with AWS, Snowflake, Spark, Kafka, Presto, or Athena.
Proficiency in SQL database administration, including PostgreSQL or MS SQL.
Experience implementing, testing, and deploying pipelines using tools such as Prefect, Airflow, Glue, or Serverless technologies like Lambda and Kinesis.
Degree in Computer Science or a relevant technical field.
Preferred Skills
Understanding of complex, distributed microservice web architectures.
Experience with Python back-end development and database-to-database ETL.
Ability to build generic solutions to improve analytical efficiency.