← Back to jobs

ClifyX Logo
Site Reliability/Observability Engineer (SRE)

ClifyX

 

Johnston, RI, USA

Posted On: 30+ days ago
Experience: 5+ years
Availability: Hybrid
Openings: 2
Category: Site Reliability/Observability Engineer
Tenure: Full-time Only
Related Jobs

No related jobs found

Description

You will manage the production environment by monitoring availability and system health across large-scale distributed software applications.

Responsibilities

  • Monitor system availability and gather metrics from operating systems and applications to assist in performance tuning and fault finding.
  • Build software and automated systems to manage platform infrastructure, applications, and capacity planning.
  • Partner with development teams, Data Scientists, and MLOps engineers to improve services through rigorous testing and release procedures.
  • Troubleshoot production issues and manage API-related problems.
  • Balance feature development speed with reliability through well-defined service-level objectives.

Required Skills

  • 5+ years of experience in site reliability or systems engineering.
  • Proficiency in programming using Python, Java, C/C++, Ruby, or JavaScript.
  • Strong knowledge of Object-Oriented Programming (OOP) and structured programming.
  • Experience with AWS services including Amazon S3, SageMaker, and Amazon Bedrock.
  • Hands-on experience with cloud-native infrastructure such as AWS Lambda and OpenShift.
  • Experience managing cloud infrastructure using AWS CloudFormation or Terraform.
  • Ability to manage and troubleshoot API management systems.
  • Any graduate degree.

Education

ANY GRADUATE

Related Jobs

No related jobs found

← Back to jobs