← Back to jobs

Persistent Systems Logo
Senior Site Reliability Engineer

Persistent Systems

 

Mexico City, Mexico

Posted On: 11 days ago
Experience: 5+ years
Availability: Remote
Openings: 2
Category: Senior Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will own the reliability, performance, and scalability of Radarlive platforms and infrastructure.

This role is remote.

Responsibilities

  • Design and manage scalable infrastructure components for Radarlive platforms.
  • Develop monitoring solutions using metrics and logs to proactively resolve performance issues.
  • Lead incident troubleshooting and conduct post-mortem analyses to improve service stability.
  • Build automation scripts for deployment, configuration management, and repetitive operational tasks.
  • Analyze traffic patterns and resource usage to drive capacity planning and resource optimization.

Required Skills

  • 5+ years of experience in systems engineering or site reliability engineering.
  • Proficiency with Azure cloud platforms and services.
  • Hands-on experience with Docker and Kubernetes for containerization and orchestration.
  • Strong programming or scripting skills in Python, Java, or PowerShell.
  • Experience with configuration management using Ansible.
  • Working knowledge of CI/CD pipelines and version control via GitHub and Jenkins.
  • Practical application of Application SRE practices.
  • Experience managing Rating platforms, specifically Radarlive.

Preferred Skills

  • Experience with observability tools such as Prometheus, Grafana, or ELK Stack.
  • Knowledge of networking, load balancing, and database management.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs