← Back to jobs

Kanak IT Services Logo
Site Reliability Engineer

Kanak IT Services

 

Malvern, PA, USA

Posted On: 8 days ago
Experience: 14-16 years
Availability: Onsite
Openings: 1
Category: site reliability engineering
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will manage large-scale distributed systems using holistic engineering principles.

This role is on-site.

Responsibilities

  • Design and implement fault-tolerant, loosely coupled distributed systems.
  • Apply observability engineering principles and OpenTelemetry to monitor system health.
  • Execute chaos engineering, performance profiling, and reliability testing practices.
  • Conduct research and use data-driven, scientific methods to solve complex reliability problems.
  • Manage the full SDLC, including testing and operations.

Required Skills

  • 14-16 years of professional experience in reliability architecture and engineering.
  • Proficiency in Node.js, TypeScript, Angular, Java, Python, and Shell scripting.
  • Deep expertise in AWS services including Cloudfront, RDS, Dynamo, ECS, EKS, Lambda, FIS, and Kinesis.
  • Hands-on experience with DevOps tooling such as Terraform, CloudFormation, and GitHub Actions.
  • Strong understanding of networking and large-scale system scaling.
  • Ability to apply systematic thinking to improve resiliency and system architecture.

Preferred Skills

  • Any Graduate degree.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs