← Back to jobs

Openkyber Logo
Senior Site Reliability Engineer

Openkyber

 

Atlanta, GA, USA

Posted On: Just posted
Experience: 10+ years
Availability: Hybrid
Openings: 1
Category: Senior Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

Architect and maintain resilient, scalable infrastructure on Google Cloud Platform while leading incident response and automation efforts.

This role is on-site.

Responsibilities

  • Design and implement GCP infrastructure, including VPC, load balancing, GKE, Cloud Run, and BigQuery.
  • Define and manage SLIs, SLOs, and error budgets to drive system reliability and performance.
  • Lead major incident restoration, conduct root cause analysis, and facilitate blameless postmortems.
  • Build and maintain infrastructure as code using Terraform and automate operations with Python or Bash.
  • Enhance observability by implementing metrics, logs, and tracing using Prometheus, Grafana, or GCP Operations Suite.

Required Skills

  • 10+ years of experience in SRE, DevOps, or cloud engineering.
  • Deep expertise in Google Cloud Platform architecture, networking, and managed services.
  • Strong proficiency with Kubernetes and container orchestration platforms.
  • Hands-on experience implementing Infrastructure as Code using Terraform.
  • Proficiency in Python, Bash, or PowerShell for automation tasks.
  • Experience with modern observability stacks and log analysis (VPC Flow Logs, Cloud Audit Logs).
  • Familiarity with orchestration platforms such as Harness.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs