← Back to jobs

Quess Corp Limited Logo
Site Reliability Engineer (Node.js, TypeScript, IaC & Cloud Platforms)

Quess Corp Limited

 

Toronto, ON, Canada

Posted On: 15+ days ago
Experience: 5+ years
Availability: Remote
Openings: 1
Category: Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will own reliability and performance for high-traffic gaming systems across on-prem and hybrid infrastructure.

This role is remote.

Responsibilities

  • Design and operate resilient on-prem and hybrid infrastructure, integrating with cloud platforms.
  • Automate deployments and infrastructure provisioning using IaC and CI/CD workflows.
  • Implement observability practices including logging, metrics, tracing, and alerting to ensure system uptime.
  • Manage incident response, conduct postmortems, and drive continuous stability improvements.
  • Optimize system performance and reliability by partnering with backend and platform teams.

Required Skills

  • 5+ years of experience in Site Reliability Engineering or infrastructure engineering.
  • Hands-on expertise with on-prem, virtualization, and data center environments.
  • Strong knowledge of physical infrastructure, networking, and distributed systems.
  • Proficiency in Node.js and TypeScript backend environments.
  • Experience with automation scripting and deployment workflows.
  • Deep understanding of system design for redundancy, failover, and disaster recovery.
  • Familiarity with monitoring tools such as Prometheus, Grafana, or Datadog.
  • Exposure to container orchestration (Docker, Kubernetes) and cloud platforms (AWS, GCP, Azure).

Preferred Skills

  • Experience supporting real-time, high-concurrency systems in gaming or fintech sectors.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs