← Back to jobs

VDart Logo
Site Reliability Engineering Lead / L3 Support

VDart

 

Boston, MA, USA

Posted On: 5 days ago
Experience: 8+ years
Availability: Hybrid
Openings: 1
Category: Site reliability engineering
Tenure: Contract - Corp-to-Corp
Related Jobs

No related jobs found

Description

Lead operations for production systems, ensuring reliability and scalability of applications, infrastructure, and databases.

None of the work mode flags (remote, hybrid, on-site) are set to 1.

Responsibilities

  • Maintain and configure VMs, network servers, and applications across on-prem, cloud, and hybrid environments.
  • Implement logging, monitoring, alerting, and service level observability solutions.
  • Troubleshoot hardware and software incidents by running diagnostics and assessing impact alongside application teams.
  • Perform capacity planning and execute hardware and software upgrades.
  • Develop scalable solutions for applications and databases using Infrastructure as Code practices.

Required Skills

  • 8+ years of experience in production systems operations.
  • Proficiency in scripting with Bash, Python, Groovy, or GoLang.
  • Hands-on experience with SRE tools, accelerators, and CI/CD pipelines.
  • Strong understanding of SLO, SLI, Error Budgets, and metrics like MTTR, MTBR, and MTD.
  • Experience with containerization and container platforms.
  • Ability to create and modify runbooks and SOPs.
  • Knowledge of Chaos Engineering tools and practices.
  • Core technical skills in Java, Jenkins, API, Database, and Microservices.
  • Experience providing L3 support.

Preferred Skills

  • 3+ years of application architecture and development experience with Java or .NET.
  • Internal or external SRE certifications.

Education

Any Gradute

Related Jobs

No related jobs found

← Back to jobs