← Back to jobs

Hirexa Solutions Logo
Site Reliability Engineer

Hirexa Solutions

 

Sweden, ME, USA

Posted On: 1 day ago
Experience: 5+ years
Availability: Onsite
Openings: 1
Category: Site reliability engineering
Tenure: Full-time Only
Related Jobs

No related jobs found

Description

You will build core platform services and lead the observability strategy for distributed cloud infrastructure.

This role is on-site.

Responsibilities

  • Build and lead the Observability team, defining the roadmap for monitoring platforms.
  • Design and implement tooling for transaction tracing, performance analysis, and large-scale logging.
  • Drive adoption of monitoring and alerting best practices across development teams.
  • Participate in technical design discussions to ensure reliability of distributed systems at scale.
  • Participate in a 24/7 on-call rotation for monitoring services.

Required Skills

  • 5+ years of professional experience in software engineering or SRE roles.
  • Hands-on experience with Kubernetes and Docker.
  • Proficiency in at least one language: Go, Java, Kotlin, Scala, Clojure, Python, or Ruby.
  • Experience with Azure cloud infrastructure and automation.
  • Deep understanding of Infrastructure as Code tools: Terraform, CDK, Pulumi, or CloudFormation.
  • Proven experience delivering observability at scale with distributed tracing systems (Jaeger, Open Zipkin).
  • Strong knowledge of distributed systems, including asynchronous communication and consensus algorithms.
  • Experience managing Linux systems.
  • Bachelor's Degree in Computer Science or a related field.

Education

Bachelor's Degree

Related Jobs

No related jobs found

← Back to jobs