← Back to jobs

LanceSoft Inc Logo
Site Reliability Engineer

LanceSoft Inc

 

Charlotte, NC, USA

Posted On: 30+ days ago
Experience: 10+ years
Availability: Hybrid
Openings: 1
Category: Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will own production stability, availability, and observability for the engineering platform.

Responsibilities

  • Manage production availability and system health through holistic monitoring and On-Call rotation.
  • Implement and automate SLO adoption, production readiness scoring, and error budget tracking.
  • Automate manual processes to increase system stability and facilitate rapid feature development.
  • Lead incident response, conduct blameless postmortems, and maintain runbook standards.
  • Drive observability maturity by identifying performance bottlenecks and providing developer feedback.

Required Skills

  • 10+ years of experience in site reliability or related engineering roles.
  • Proficiency with GitLab and Infrastructure as Code tools like Terraform.
  • Hands-on experience with monitoring tools including Dynatrace and Splunk.
  • Experience managing public cloud platforms, specifically AWS and API gateways.
  • Expertise in source version control using Git.
  • Strong troubleshooting skills and experience enhancing observability.
  • Proven ability to manage incident response and support application teams.
  • Degree in any field (Any Graduate).

Preferred Skills

  • Experience developing APIs, microservices, or frontend components.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs