← Back to jobs

Kaizen Technologies, Inc Logo
Site Reliability Engineer

Kaizen Technologies, Inc

 

Sunnyvale, CA, USA

Posted On: 15+ days ago
Experience: 5+ years
Availability: Onsite
Openings: 1
Category: Site Reliability Engineer
Tenure: Contract - Corp-to-Corp
Related Jobs

No related jobs found

Description

You will own reliability and operational health for production systems.

This role is on-site.

Responsibilities

  • Maintain and operate production systems running on AWS and/or GCP.
  • Manage and debug containerized applications within Kubernetes clusters.
  • Implement and maintain continuous integration pipelines.
  • Monitor system health using established monitoring tools and logging services.

Required Skills

  • 5+ years of professional experience.
  • Proficiency in Linux and Python.
  • Experience with Shell Scripting.
  • Hands-on experience maintaining production environments on AWS or GCP.
  • Experience managing Kubernetes clusters and containerized apps (Golang, Java, Python).
  • Knowledge of Kafka, Spark, Storm, Cassandra, or ElasticSearch.
  • Familiarity with Infrastructure as Code tools like Terraform or CloudFormation.
  • Experience with monitoring stacks such as Prometheus, Grafana, or CloudWatch.

Preferred Skills

  • Understanding of specific data stores like PostgreSQL, Redis, or Zookeeper.
  • Experience with CI tools like Jenkins or CircleCI.

Education

Any Gradute

Related Jobs

No related jobs found

← Back to jobs