← Back to jobs

SAP Logo
Site Reliability Engineer

SAP

 

Montreal, QC, Canada

Posted On: 30+ days ago
Experience: 5+ years
Availability: Onsite
Openings: 1
Category: Site Reliability Engineer.
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will operate and support business-critical cloud services, focusing on service reliability and stability.

Responsibilities

  • Act as a technical expert during live site incidents, performing deep technical investigations and troubleshooting to resolve issues.
  • Drive root cause analysis and implement follow-up improvements to prevent incident recurrence.
  • Build software-based solutions and automation to improve service reliability and infrastructure monitoring.
  • Enhance platform monitoring by gathering system metrics and implementing recovery tools.
  • Collaborate with development teams on postmortem outputs and product improvements.

Required Skills

  • 5+ years of professional experience in SRE or related roles.
  • BSc degree in Computer Science or a related technical field.
  • Experience with Kubernetes and container technologies.
  • Understanding of modern cloud architectures.
  • Proficiency in scripting and CI/CD workflows.
  • Knowledge of AWS, Azure, or GCP.
  • Ability to work efficiently in emergency situations and solve problems within a global team.
  • Fluency in English; basic French is a plus.

Preferred Skills

  • Coding experience with Go, Terraform, or Python.
  • Experience with Unix/Linux operating systems.
  • Hands-on experience with monitoring, logging, and alerting tools like Grafana, Prometheus, Kibana, Loki, Splunk, or Dynatrace.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs