You will own infrastructure reliability and automation at scale.
Responsibilities
Define and create Customer User Journeys (CUJ), Service Level Objectives (SLO), Service Level Indicators (SLI), and Error Budgets based on Non-Functional Requirements (NFR).
Design and implement automated workflows within the Software Development Lifecycle (SDLC) or IT operations environment.
Reduce toil through automation across operational processes.
Manage and maintain infrastructure using Infrastructure as Code (IaC) principles.
Required Skills
5+ years of professional experience in Site Reliability Engineering or a related operations role.
Strong hands-on experience with Terraform and managing IaC pipelines.
Proficiency in scripting languages including Python, Bash, PowerShell, and Ansible.
Hands-on experience with container orchestration using Kubernetes.
Working knowledge of GCP.
Experience with CI/CD tooling, specifically GitHub and Docker.
Familiarity with version control systems like Git and SonarQube.