← Back to jobs

Global Applications Solution Logo
Site Reliability Engineer (SRE) for Kubernetes
Posted On: 3 days ago
Experience: 10+ years
Availability: Hybrid
Openings: 1
Category: Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will support a machine learning platform running on an in-house Kubernetes cluster.

This role is on-site.

Responsibilities

  • Support the machine learning platform hosted on internal Kubernetes clusters.
  • Manage and maintain Kubernetes infrastructure to ensure platform stability.
  • Implement MLOps practices to support machine learning workflows.

Required Skills

  • 5+ years of professional work experience with a bachelor's degree.
  • 2+ years of hands-on experience with Kubernetes (K8s).
  • Direct experience with Machine Learning Operations (MLOps).
  • Experience managing or supporting Kubernetes-based environments.
  • Bachelor's degree in any field.

Preferred Skills

  • Experience with Kubeflow frameworks on Kubernetes.
  • Proficiency with Apache Spark.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs