← Back to jobs

Experis Logo
Senior Site Reliability Engineer

Experis

 

San Francisco, CA, USA

Posted On: 5 days ago
Experience: 7+ years
Availability: Onsite
Openings: 1
Category: Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

Own the reliability, performance, and availability of production systems running on Azure.

This role is hybrid.

Responsibilities

  • Design and operate highly reliable production systems within the Azure cloud environment.
  • Lead on-call rotations and drive high-severity incident response and resolution.
  • Perform deep root cause analysis and facilitate blameless post-incident reviews.
  • Evaluate observability standards by defining SLIs, SLOs, and maintaining critical dashboards.
  • Resolve production issues across Java services, Kubernetes clusters, and underlying cloud infrastructure.

Required Skills

  • 7+ years of Site Reliability Engineering or Production Engineering experience.
  • Hands-on expertise with Azure cloud infrastructure.
  • Proficiency in container orchestration using Kubernetes and Docker.
  • Strong background with Java-based systems and microservices architecture.
  • Experience building and maintaining CI/CD pipelines, specifically GitHub Actions.
  • Proficiency with observability platforms, particularly Dynatrace.
  • Strong automation skills using Python, Bash, or Ansible.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs