← Back to jobs

eTeam Logo
Senior Site Reliability Engineer

eTeam

 

Plano, TX, USA

Posted On: 2 days ago
Experience: 15+ years
Availability: Onsite
Openings: 1
Category: Sr. Site Reliability Engineer
Tenure: Full-time Only
Related Jobs

No related jobs found

Description

You will own the reliability, availability, and performance of production services in cloud environments.

This role is on-site.

Responsibilities

  • Maintain and optimize production systems on AWS or GCP.
  • Develop automation scripts using Python or Shell.
  • Manage and troubleshoot Kubernetes clusters and containerized applications.
  • Implement infrastructure as code using Terraform or CloudFormation.
  • Participate in incident response and post-mortem analysis to improve system stability.

Required Skills

  • 8 to 15 years of relevant experience in Site Reliability Engineering or a similar role.
  • Strong experience with Linux/Unix operating systems.
  • Proficiency with AWS or GCP cloud platforms.
  • Hands-on experience with Kubernetes for managing containerized applications.
  • Proficiency in Python and Linux/Unix Shell scripting.
  • Competency with Terraform, CloudFormation, Kafka, Spark, Storm, Cassandra.
  • Bachelor's degree in Computer Science, Engineering, or a related field.

Preferred Skills

  • Familiarity with monitoring solutions such as Prometheus, Grafana, AWS CloudWatch, or Stackdriver.
  • Experience with CI tools like Jenkins, Travis CI, or CircleCI.
  • Knowledge of PostgreSQL, Redis, Nginx, and AWS Elasticsearch Service.

Education

Bachelor's degree

Related Jobs

No related jobs found

← Back to jobs