← Back to jobs

Akkodis Logo
Senior Site Reliability Engineer

Akkodis

 

Carrollton, TX, USA

Posted On: 15+ days ago
Experience: 8+ years
Availability: Hybrid
Openings: 1
Category: Senior Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You own reliability and operational excellence for our systems.

This role is hybrid.

Responsibilities

  • Lead L2/L3 incident response, driving deep-dive troubleshooting and executing RCAs with permanent corrective actions.
  • Develop and optimize monitoring, alerting, and observability systems to ensure adherence to SLOs/SLIs.
  • Collaborate with backend engineering teams through Production Readiness Reviews and influence architectural decisions.
  • Analyze incident patterns and system metrics to reduce operational toil through automation and long-term engineering solutions.
  • Create and maintain runbooks, documentation, and operational guidelines for fast incident recovery.

Required Skills

  • 8–10 years of overall experience, including 5–7 years in SRE, DevOps, or System Engineering.
  • Bachelor’s degree in computer science, Information Technology, or a related technical field.
  • Strong hands-on AWS expertise across core services (EC2, VPC, CloudWatch, S3, RDS, IAM, Lambda).
  • Proven L2/L3 production support experience with advanced troubleshooting and incident resolution capabilities.
  • Expertise in Cloud Infrastructure management.
  • Experience with IaC Pipelines.
  • Proficiency in troubleshooting complex production issues.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs