Description
You will lead a team of SREs focused on system reliability, performance, and scalability.
Responsibilities
- Build and mentor a team of SREs through goal setting and performance reviews.
- Oversee the design and maintenance of high-availability systems.
- Drive infrastructure automation and enhance CI/CD pipelines.
- Lead incident management, performance monitoring, and issue resolution.
- Implement SRE best practices to foster continuous improvement.
Required Skills
- 5+ years of experience in Site Reliability Engineering or related roles.
- Expertise in System Reliability and high-availability system design.
- Hands-on experience with Terraform for infrastructure automation.
- Proficiency with Ansible for configuration management.
- Experience developing and managing CI/CD pipelines.
- Proven ability to lead technical teams and mentor engineers.
- Strong background in automation scripting and performance optimization.