Description
You will lead technical efforts to maintain system reliability, optimize cloud infrastructure, and drive automation initiatives.
This role is on-site.
Responsibilities
- Identify and resolve complex system issues to meet service level agreements.
- Develop and maintain runbooks to standardize operational procedures.
- Optimize cloud platform infrastructure for availability and scalability.
- Create automation tools using shell scripting or Python to improve team efficiency.
- Troubleshoot and resolve issues in collaboration with the engineering team.
Required Skills
- 8+ years of relevant experience in system administration or engineering.
- Bachelor's degree in Computer Science, Engineering, or equivalent experience.
- Strong Linux System Admin experience with advanced troubleshooting skills.
- Demonstrable scripting and automation skills in Bash and Python.
- Experience with Service Defined Storage (SDS).
- Familiarity with cloud services including compute, storage, and network.
- Experience with virtualization technologies.
- Familiarity with Istio and Envoy service mesh technologies.