← Back to jobs

InnoMethods Corporation Logo
Senior Site Reliability Engineer

InnoMethods Corporation

 

Jersey City, NJ, USA

Posted On: 30+ days ago
Experience: 10+ years
Availability: Hybrid
Openings: 1
Category: Senior Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will manage incident response and enhance system reliability through scalable infrastructure design.

Responsibilities

  • Manage incident response and participate in an on-call rotation.
  • Build monitoring solutions including dashboards and alerting using Dynatrace and Grafana.
  • Automate infrastructure and CI/CD pipelines using Python, Ansible, and Terraform.
  • Design and maintain scalable core infrastructure to improve reliability.
  • Apply SRE principles including SLIs, SLOs, observability, and toil reduction.

Required Skills

  • 10+ years of experience in systems or reliability engineering.
  • Strong programming skills with Python.
  • Hands-on experience with Azure and AWS.
  • Proficiency with Terraform, Ansible, and Infrastructure as Code.
  • Experience with Databricks and OpenAI.
  • Knowledge of monitoring tools such as Dynatrace and Grafana.
  • Experience with Jenkins and Bitbucket.
  • Understanding of Agile or Agile SAFe processes.
  • Ability to work in a blameless, data-driven environment.

Preferred Skills

  • Experience with OpenAI integration in production environments.
  • Deep expertise in Azure cloud services.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs