← Back to jobs

InfiCare Technologies Logo
Site Reliability Engineer

InfiCare Technologies

 

Austin, TX, USA

Posted On: Just posted
Experience: 10+ years
Availability: Onsite
Openings: 1
Category: Site Reliability Engineer
Tenure: Full-time Only
Related Jobs

No related jobs found

Description

You will own the reliability and operational health of large-scale production systems.

This role is on-site.

Responsibilities

  • Troubleshoot and fix application failures, performance degradation, and infrastructure issues across cloud, batch, database, and network layers.
  • Initiate and drive incident response for outages and major incidents, ensuring rapid service restoration.
  • Manage Incident, Problem, Release, and Change management processes effectively.
  • Debug, analyze issues, write code fixes, and document SOPs for assigned user stories.
  • Build monitoring solutions using APM tools and automate day-to-day operational tasks.

Required Skills

  • 10+ years of experience working with distributed systems, microservices, and configuration management.
  • Hands-on experience troubleshooting Linux/Unix environments.
  • Strong programming experience in at least one language: Java, C#, or .NET.
  • Expertise in cloud platforms: AWS, GCP, Azure Cloud, or PCF.
  • Proficiency in scripting using Shell, PowerShell, or Python.
  • Experience with monitoring tools such as Splunk, AppDynamics, ThousandEyes, and Grafana.
  • Solid understanding of networking concepts (TCP/IP, SSL/TLS, Load Balancers).
  • Experience performing production deployments using CI/CD pipelines.

Preferred Skills

  • Bachelor's degree in Engineering or related field.
  • Experience with ITRS monitoring solutions.

Education

Bachelor's degree

Related Jobs

No related jobs found

← Back to jobs