← Back to jobs

VARITE INC Logo
Site Reliability Engineer

VARITE INC

 

Bangalore, Karnataka, India

Posted On: 3 days ago
Experience: 5+ years
Availability: Onsite
Openings: 2
Category: Site Reliability Engineer
Tenure: Full-time Only
Related Jobs

No related jobs found

Description

You will own the reliability, performance, and scalability of the Managed Services infrastructure.

Responsibilities

  • Maintain operational excellence through root cause analysis and post-incident reviews.
  • Build and maintain observability frameworks to monitor distributed systems and microservices.
  • Define and track SLIs, SLOs, and Error Budgets to balance reliability with innovation.
  • Automate operational tasks and optimize processes using Python and Ansible.
  • Manage Kubernetes clusters, containerized environments, and service-to-service communication.

Required Skills

  • 5+ years of experience in reliability engineering or systems administration.
  • Proficiency in Linux administration, including system tuning, SSH, and log analysis.
  • Hands-on experience with Kubernetes and container orchestration.
  • Strong automation skills using Ansible and Python scripting.
  • Expertise in monitoring and observability tools including Prometheus, Grafana, and Dynatrace.
  • Deep understanding of networking protocols and Layer 1-3 troubleshooting.
  • Experience managing PKI certificates and HashiCorp Vault.
  • Proven ability to instrument and analyze performance metrics in distributed architectures.

Preferred Skills

  • Experience with OpenTelemetry observability frameworks.
  • Familiarity with ITIL frameworks and incident management best practices.

Education

Any Gradute

Related Jobs

No related jobs found

← Back to jobs