← Back to jobs

Extend Information Systems Inc Logo
Site Reliability Engineer

Extend Information Systems Inc

 

Cupertino, CA, USA

Posted On: 8 days ago
Experience: 5+ years
Availability: Hybrid
Openings: 1
Category: site reliability engineering
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will own cloud deployment and monitoring systems within an AWS environment.

This role is on-site.

Responsibilities

  • Build and maintain automation scripts from scratch using Python.
  • Manage cloud deployments and infrastructure orchestration.
  • Monitor system health and performance using observability platforms.
  • Collaborate with engineering teams to support large-scale systems.

Required Skills

  • 3-5 years of Software Engineering or Development experience.
  • 3-5 years of Python experience, specifically writing scripts from scratch.
  • 3-5 years of AWS experience.
  • 2 years of Docker experience.
  • Experience with CloudWatch for monitoring and deployment.
  • Proficiency with Prometheus, Grafana, and Telegraf.
  • Experience using Splunk for log analysis and observability.
  • Bachelor Degree in a computer-related field.

Preferred Skills

  • Exposure to Kubernetes.
  • Experience with Terraform or Pulumi.
  • Operational experience supporting large-scale Banking or Payments systems.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs