← Back to jobs

Largeton Logo
Infrastructure Research Engineer

Largeton

 

United States

Posted On: 2 days ago
Experience: 10+ years
Availability: Remote
Openings: 1
Category: Infrastructure Engineers
Tenure: Full-time Only
Related Jobs

No related jobs found

Description

You will lead research and optimization of GPU computing performance and AI infrastructure.

This role is remote.

Responsibilities

  • Design and fine-tune large-scale distributed systems for AI workloads.
  • Conduct benchmarking and performance analysis for compute, storage, and networking components.
  • Optimize servers, GPUs, networks, and databases for efficiency and scalability.
  • Build, deploy, and manage containerized environments using Kubernetes, Rancher, and Kubeflow.
  • Develop automation and debugging scripts in Python.

Required Skills

  • 10+ years of experience with Kubernetes, containers, and distributed systems.
  • 5–7 years in GPU computing performance and infrastructure research.
  • Strong Python scripting skills.
  • Deep knowledge of AI infrastructure optimization and debugging.
  • Proficiency in benchmarking tools and performance tuning.
  • Experience troubleshooting infrastructure issues across servers, GPUs, networks, and storage.

Preferred Skills

  • Experience with Rancher and Kubeflow for container orchestration.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs