← Back to jobs

Nisum Technologies Logo
DevOps/Site Reliability Engineer

Nisum Technologies

 

Hyderabad, Telangana, India

Posted On: 30+ days ago
Experience: 6+ years
Availability: Hybrid
Openings: 2
Category: Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will own production stability and incident resolution for high-availability systems. This role requires active participation in a 24x7 on-call rotation, taking primary or secondary responsibility to ensure rapid incident resolution and system health.

Responsibilities

  • Triage incidents efficiently, lead resolution efforts, and ensure closure within defined SLAs during on-call rotations.
  • Monitor system health and performance proactively to identify and mitigate issues before they escalate.
  • Build automation, enhance monitoring and alerting systems, and create clear runbooks to streamline operations.
  • Collaborate with engineering and infrastructure teams to support deployments, troubleshoot complex issues, and drive system improvements.
  • Conduct post-incident reviews to identify root causes and implement preventive measures to minimize downtime.

Required Skills

  • 6+ years of experience in DevOps, SRE, or production support environments.
  • Hands-on proficiency with cloud platforms: AWS, Azure, and GCP.
  • Strong scripting skills in Python and Bash for automation and operational efficiency.
  • Experience with Linux/Unix systems administration and networking fundamentals.
  • Proficiency in monitoring and incident management tools: PagerDuty, ServiceNow, Datadog, Splunk, and New Relic.
  • Demonstrated ability to maintain structured incident response processes and conduct root cause analysis (RCA).
  • Bachelor’s degree in Computer Science, Information Systems, Engineering, Computer Applications, or related field.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs