← Back to jobs

Infosys Limited Logo
Site Reliability Engineer

Infosys Limited

 

Bengaluru East, Bommenahalli, Karnataka, India

Posted On: 15+ days ago
Experience: 5+ years
Availability: Hybrid
Openings: 1
Category: site reliability engineering
Tenure: Full-time Only
Related Jobs

No related jobs found

Description

You will design, build, and operate large-scale distributed systems handling millions of transactions. You own the service lifecycle from conception to deployment, ensuring high availability and performance across public and private clouds.

This role is on-site.

Responsibilities

  • Define and implement system architecture, deployment standards, and operational best practices.
  • Manage capacity and performance to scale infrastructure globally.
  • Monitor system health, manage incidents, and conduct postmortems to improve reliability.
  • Integrate GenAI and AIOps tools for automated incident detection, root cause analysis, and self-healing workflows.
  • Develop ML models using operational telemetry to predict failures and enable proactive remediation.

Required Skills

  • 5+ years of experience in SRE or distributed systems engineering.
  • Proficiency in Python, Ruby, or GoLang with Object-Oriented Programming knowledge.
  • Hands-on experience with AIOps platforms: Moogsoft, Dynatrace, Splunk, BigPanda, Datadog, or Elastic.
  • Experience integrating Generative AI (GenAI) and Prompt Engineering for observability and automation.
  • Familiarity with cloud platforms: AWS, GCP, or Azure (e.g., Bedrock, Vertex AI).
  • Strong background in system design, monitoring, and incident response.

Education

Any Gradute

Related Jobs

No related jobs found

← Back to jobs