← Back to jobs

Galaxy i technologies Inc Logo
Senior Site Reliability Engineer

Galaxy i technologies Inc

 

Owings Mills, MD, USA

Posted On: 2 days ago
Experience: 6+ years
Availability: Onsite
Openings: 2
Category: Senior Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will ensure the availability, performance, and reliability of large-scale distributed systems and critical infrastructure.

This role is on-site.

Responsibilities

  • Monitor system availability and implement automated alerting mechanisms to identify outages or performance degradation.
  • Analyze performance metrics to resolve latency bottlenecks and implement optimization techniques.
  • Develop and maintain metrics dashboards to track key performance indicators and identify system trends.
  • Optimize resource utilization to minimize infrastructure expenditure and improve allocation for new services.
  • Manage release processes by developing automated deployment and rollback procedures to mitigate update risks.

Required Skills

  • 6+ years of experience as a Site Reliability Engineer or in a similar role.
  • Proven experience operating and supporting Kubernetes in production at scale, preferably EKS.
  • Expertise in Linux systems administration, including server, operating system, and network configuration management.
  • Proficiency in scripting and automation using Bash or Python.
  • Hands-on experience with AWS.
  • Experience with Docker and GitLab CI/CD.
  • Strong troubleshooting skills for resolving complex technical issues in distributed systems.
  • Ability to communicate technical concepts to both technical and non-technical stakeholders.

Preferred Skills

  • Bachelor's degree or equivalent graduate education.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs