← Back to jobs

ACL Digital Logo
Site Reliability Engineer

ACL Digital

 

Ahmedabad, Gujarat, India

Posted On: 5 days ago
Experience: 5+ years
Availability: Hybrid
Openings: 1
Category: Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will own the reliability, monitoring, and incident response for production systems.

This role is on-site.

Responsibilities

  • Monitor system performance and identify potential issues before they impact users.
  • Respond to incidents, troubleshoot Level 1 issues, and resolve problems promptly.
  • Analyze monitoring data to identify trends, anomalies, and potential risks.
  • Collaborate with development and platform teams to enhance application performance.
  • Automate monitoring processes to reduce manual overhead and improve efficiency.

Required Skills

  • 5+ years of experience in Site Reliability Engineering or DevOps roles.
  • Hands-on experience with AWS cloud platforms.
  • Proficiency with containerization technologies: Docker and Kubernetes.
  • Experience with monitoring and logging tools: Prometheus, Grafana, and ELK Stack.
  • Familiarity with Infrastructure as Code using Terraform.
  • Strong understanding of Linux operating systems and networking.
  • Experience with database management and distributed systems.
  • Ability to participate in on-call rotations and support major incidents.

Preferred Skills

  • Experience with incident management tools like PagerDuty.
  • Knowledge of best practices in observability and alerting.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs