← Back to jobs

Persistent Systems Logo
Site Reliability Engineer

Persistent Systems

 

Hyderabad, Telangana, India

Posted On: 11 days ago
Experience: 7+ years
Availability: Onsite
Openings: 3
Category: Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will manage system observability, infrastructure automation, and cloud reliability.

This role is on-site.

Responsibilities

  • Implement and manage observability stacks using Prometheus, Grafana, Datadog, New Relic, or AppDynamics.
  • Maintain logging platforms including Elasticsearch, Splunk, Loki, or Fluentd.
  • Configure distributed tracing using OpenTelemetry, Tempo, Jaeger, or Zipkin.
  • Manage infrastructure via Terraform and Ansible within Azure and Kubernetes environments.
  • Troubleshoot system performance, networking, and complex distributed systems issues.

Required Skills

  • 7+ years of experience in SRE, Observability, DevOps, or a related field.
  • Hands-on experience with Azure cloud platforms and Kubernetes.
  • Proficiency in Prometheus, Grafana, Datadog, New Relic, or AppDynamics.
  • Experience with logging tools such as Elasticsearch, Splunk, Loki, or Fluentd.
  • Strong knowledge of distributed tracing (OpenTelemetry, Tempo, Jaeger, Zipkin).
  • Proficiency in scripting with Python, Bash, or PowerShell.
  • Experience with Infrastructure as Code using Terraform and Ansible.
  • Any Graduate degree.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs