← Back to jobs

Pure Storage Logo
Senior Site Reliability Engineer

Pure Storage

 

Prague, Czechia

Posted On: Just posted
Experience: 5+ years
Availability: Onsite
Openings: 2
Category: Senior Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You own the reliability and availability of core cloud services by building operational frameworks, monitoring, and automation.

This role is on-site.

Responsibilities

  • Lead incident response and root cause analysis to drive rapid recovery and long-term fixes.
  • Architect and implement Infrastructure-as-Code to streamline deployments and service operations.
  • Partner with product and engineering teams to embed SRE best practices into service architecture.
  • Develop and evolve observability systems, including metrics, logging, and actionable alerting.
  • Design and operate highly available cloud-native systems based on SLOs and uptime requirements.

Required Skills

  • 5+ years of experience designing and operating highly available cloud services.
  • Expertise with public cloud platforms: AWS, Azure, or GCP.
  • Proficiency in Infrastructure-as-Code using Terraform, Ansible, CloudFormation, or Puppet.
  • Hands-on experience with containerized environments and Kubernetes.
  • Strong programming skills in Python, Go, Java, or Ruby.
  • Ability to build observability stacks (e.g., ELK, Prometheus, OpenTelemetry).
  • Deep understanding of Linux systems and networking fundamentals.
  • Experience managing on-call processes using tools like PagerDuty.

Preferred Skills

  • Experience with OpenTelemetry for distributed tracing.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs