Description
You will lead the development of observability solutions and drive system reliability, performance, and scalability across distributed systems.
This role is on-site.
Responsibilities
- Lead a team of Site Reliability Engineers through direct mentorship.
- Develop and implement observability strategies using New Relic.
- Drive system reliability and performance across e-commerce and distributed environments.
- Optimize performance for Java, Kafka, and SQL workloads.
- Manage and improve CI/CD pipelines and automation workflows.
Required Skills
- 15+ years of professional experience in technical leadership roles.
- Expertise in New Relic development and observability practices.
- Proficiency in automation scripting using Python or Java.
- Hands-on experience with Linux environments.
- Experience with Terraform and Ansible for infrastructure automation.
- Strong knowledge of CI/CD pipelines.
- Experience managing Kafka and SQL systems.
- Background in managing distributed systems.
Preferred Skills
- Bachelor's degree or equivalent experience.