← Back to jobs

Hirexa Solutions Logo
Site Reliability Engineer

Hirexa Solutions

 

Sweden, ME, USA

Posted On: 30+ days ago
Experience: 5+ years
Availability: Onsite
Openings: 1
Category: Site reliability engineering
Tenure: Full-time Only
Related Jobs

No related jobs found

Description

You will develop platform core services and solve complex cloud-infrastructure automation and multi-region networking challenges.

Responsibilities

  • Build and lead an Observability team, driving the roadmap for observability platforms and mentoring engineers.
  • Design and implement tooling for transaction tracing, performance analysis, and large-scale logging and metrics collection.
  • Contribute to technical design discussions to ensure the reliability of distributed systems at scale.
  • Drive adoption of best practices in monitoring and alerting across development teams.
  • Participate in a 24/7 on-call rotation for monitoring and observability services.

Required Skills

  • 5+ years of professional experience in software engineering or SRE roles.
  • Proven experience delivering observability at scale and working with distributed tracing systems like Jaeger or Open Zipkin.
  • Proficiency in at least one of the following: Go, Java, Kotlin, Scala, Clojure, Python, or Ruby.
  • Hands-on experience with Kubernetes and Docker.
  • Experience with Azure cloud infrastructure and automation.
  • Strong knowledge of Infrastructure as Code tools such as Terraform, CDK, Pulumi, or CloudFormation.
  • Deep understanding of distributed systems development, including asynchronous communication patterns and consensus algorithms.
  • Experience managing and working within Linux systems.
  • Bachelor's Degree in Computer Science or a related field.

Education

Bachelor's Degree

Related Jobs

No related jobs found

← Back to jobs