← Back to jobs

People Prime Worldwide Logo
Site Reliability Engineer

People Prime Worldwide

 

Hyderabad, TS, India

Posted On: Just posted
Experience: 6+ years
Availability: Hybrid
Openings: 1
Category: Site Reliability Engineer
Tenure: Contract - Corp-to-Corp
Related Jobs

No related jobs found

Description

You will own the reliability and operational health of critical systems.

This role is hybrid.

Responsibilities

  • Define and implement SRE strategies and best practices aligned with organizational objectives.
  • Monitor service level agreements (SLAs), service level objectives (SLOs), and service level indicators (SLIs).
  • Lead initiatives to improve system reliability, availability, scalability, and performance.
  • Implement and improve incident management processes to minimize downtime and ensure timely resolutions.
  • Drive observability practices by implementing monitoring, logging, and alerting systems.

Required Skills

  • 6+ years of professional experience.
  • Proficiency in writing Splunk Queries and Alerts.
  • Hands-on experience with at least one APM tool (NewRelic, AppDynamics, Honeycomb, DataDog).
  • Expertise in automation tools and scripting languages (Python or JavaScript/Node.js).
  • Proficiency in any cloud platforms (AWS, GCP, Azure).
  • Strong understanding of distributed systems and microservices architecture.
  • Experience with container orchestration tools like Kubernetes.
  • Experience with monitoring tools including Prometheus and Grafana.

Education

Bachelor's degree

Related Jobs

No related jobs found

← Back to jobs