← Back to jobs

Bosch Global Software Technologies Private Limited Logo
Site Reliability Engineer
Posted On: 30+ days ago
Experience: 6+ years
Availability: Onsite
Openings: 2
Category: Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will ensure the reliability, scalability, and performance of systems supporting Data Engineering Projects.

This role is on-site.

Responsibilities

  • Design and engineer highly scalable, high-availability systems for high-throughput workloads.
  • Develop, deploy, and manage monitoring systems with proactive alerting to identify and resolve issues.
  • Automate routine tasks including deployments, monitoring, and policy enforcement using suitable frameworks.
  • Optimize system performance by identifying bottlenecks and implementing appropriate solutions.
  • Respond to system outages, perform root cause analysis, and implement fixes to prevent future incidents.

Required Skills

  • 6+ years of hands-on experience maintaining large-scale, high-availability Data Engineering solutions.
  • Expertise with cloud platforms, specifically Microsoft Azure.
  • Proficiency in Cloud infrastructure and CI/CD frameworks for Infrastructure as Code (IaC) using Terraform, ARM, and YAML.
  • Hands-on experience with cloud-native containerization and deployment using Docker and Kubernetes (k8s).
  • Experience with large-scale Azure DevOps and Azure PaaS components.
  • Tool knowledge including Argo, Terraform (CLI), Azure-CLI, Kubectl, Flux, Helm, Istio, and Grafana.
  • Strong Kubernetes administration skill set.
  • Programming experience in Python or Go.
  • Deep understanding of Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgeting.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs