← Back to jobs

Volto Logo
LLM Engineer

Volto

 

Bangalore, Karnataka, India

Posted On: Just posted
Experience: 5+ years
Availability: Remote
Openings: 2
Category: LLM Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will engineer solutions to optimize LLM performance at scale.

This role is remote.

Responsibilities

  • Analyze tracing logs from LLM inference and training runs to identify performance issues and inefficiencies.
  • Develop tools and scripts to parse, visualize, and monitor LLM tracing data.
  • Investigate and resolve model latency and throughput issues related to runtime behavior.
  • Collaborate with ML and infra teams to recommend and implement performance optimizations.
  • Contribute to best practices for performance tracing, benchmarking, and logging across model deployments.

Required Skills

  • 5+ years of experience working with large-scale ML models, preferably LLMs (GPT, BERT).
  • Proficiency in Python and ML frameworks (PyTorch, TensorFlow).
  • Familiarity with model tracing tools such as PyTorch Profiler, TensorBoard, or DeepSpeed.
  • Experience with MLOps and strong problem-solving skills for analyzing complex logs and metrics.

Preferred Skills

  • Experience with distributed training/inference and GPU performance optimization.
  • Knowledge of systems profiling tools (e.g., NVIDIA Nsight, perf, Flamegraphs).

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs