← Back to jobs

Hallmark Global Technologies Inc Logo
Site Reliability Engineer

Hallmark Global Technologies Inc

 

Columbus, Ohio, USA

Posted On: 11 days ago
Experience: 7+ years
Availability: Onsite
Openings: 1
Category: Site Reliability Engineer
Tenure: Contract - Corp-to-Corp
Related Jobs

No related jobs found

Description

Lead reliability efforts for critical infrastructure, owning complex incident resolution, automation, and system design.

This role is on-site.

Responsibilities

  • Lead troubleshooting for high-impact production incidents, providing detailed root cause analysis and preventative measures.
  • Develop and maintain infrastructure automation using Terraform, Ansible, or CloudFormation to eliminate manual tasks.
  • Build and optimize monitoring, alerting, and observability solutions using Prometheus, Grafana, Datadog, or Dynatrace.
  • Mentor junior engineers and drive architectural decisions for scalability and resilience.
  • Analyze system performance metrics to recommend optimizations and support capacity planning efforts.

Required Skills

  • 7+ years of experience in site reliability engineering, DevOps, or systems administration.
  • Strong proficiency in Linux/Unix administration and scripting with Python, Bash, or Go.
  • Hands-on experience with containerization and orchestration technologies, specifically Docker and Kubernetes.
  • Proven track record of managing complex infrastructure and troubleshooting production issues in high-scale environments.
  • Experience with Infrastructure as Code tools such as Terraform, Ansible, or CloudFormation.
  • Bachelor's degree in computer science, Information Technology, or equivalent work experience.

Education

Bachelor's degree

Related Jobs

No related jobs found

← Back to jobs