← Back to jobs

Tekgence Logo
Site Reliability Engineer

Tekgence

 

Toronto, ON, Canada

Posted On: 30+ days ago
Experience: 5+ years
Availability: Hybrid
Openings: 2
Category: Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

Position Summary:

We are seeking an experienced Site Reliability Engineer (SRE) with strong expertise in Observability, DevOps, and Reliability Engineering to design, implement, and maintain highly available, scalable, and resilient cloud-native platforms. The ideal candidate will have hands-on experience with modern monitoring and incident management platforms, Kubernetes-based infrastructure, Infrastructure as Code (IaC), and reliability frameworks such as SLI/SLOs.


Required Qualifications:

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or related field.
  • 5+ years of experience in SRE, DevOps, Platform Engineering, or Infrastructure Operations roles.
  • Strong experience with observability and monitoring platforms:
  • Dynatrace
  • ELK Stack
  • Splunk
  • PagerDuty
  • Proven expertise in implementing and managing SLI/SLO-based reliability programs.
  • Hands-on experience with Azure Cloud and Azure-managed services.
  • Strong experience with Azure Kubernetes Service (AKS) and container orchestration.
  • Expertise in Terraform and Infrastructure as Code (IaC).
  • Experience with Linux systems administration and cloud-native architectures.
  • Strong scripting and automation skills using Python, Bash, PowerShell, or similar languages.
  • Experience with CI/CD pipelines and DevOps tooling.
  • Excellent troubleshooting, analytical, and problem-solving skills

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs