← Back to jobs

Lorven Technologies Logo
Senior Site Reliability and Operations Engineer

Lorven Technologies

 

Irving, TX, USA

Posted On: 12 days ago
Experience: 10+ years
Availability: Hybrid
Openings: 1
Category: Senior Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will manage and optimize distributed caching and compute grid solutions within Kubernetes and OpenShift environments.

This role is hybrid.

Responsibilities

  • Design and optimize distributed caching and compute grid solutions on Kubernetes and OpenShift.
  • Implement high-throughput compute grids using technologies such as Apache Ignite, GridGain, or Coherence.
  • Ensure high availability, scalability, and reliability of distributed systems.
  • Automate infrastructure provisioning and deployments using Ansible and Helm charts.
  • Troubleshoot and resolve incidents related to platforms, infrastructure, and distributed caching to minimize downtime.

Required Skills

  • 10+ years of professional experience in site reliability or infrastructure engineering.
  • Extensive experience with Kubernetes, including OpenShift and on-prem/cloud clusters.
  • Proficiency in Java, Go, or Python.
  • Hands-on experience with Docker and Helm.
  • Strong knowledge of CI/CD pipelines using Jenkins, ArgoCD, or GitHub Actions.
  • Experience with observability tools including Prometheus, Grafana, Loki, Jaeger, Splunk, ELK, or OpenTelemetry.
  • Understanding of networking, service meshes like Istio or Linkerd, and Kubernetes security best practices.
  • Experience managing multi-cluster and hybrid cloud Kubernetes deployments.
  • Background in implementing microservices and containerized workloads.

Preferred Skills

  • Experience with Apache Ignite, GridGain, or Coherence.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs