← Back to jobs

MARVEL Infotech Inc Logo
Senior Site Reliability Engineer

MARVEL Infotech Inc

 

San Leandro, CA, USA

Posted On: 30+ days ago
Experience: 10+ years
Availability: Onsite
Openings: 1
Category: Site Reliability Engineer
Tenure: Contract - W2
Related Jobs

No related jobs found

Description

Maintain platform health and resiliency within a 12/7 support organization.

Responsibilities

  • Improve platform health by proactively identifying performance bottlenecks and areas for improvement.
  • Build dashboards and configure alerts using APM tools to monitor distributed systems.
  • Automate processes and testing to support rapid application development.
  • Collaborate with Security, Networking, and Infrastructure teams to resolve ecosystem challenges.
  • Manage distributed storage and dynamic resource management frameworks.

Required Skills

  • 10+ years of software engineering or production support/SRE experience.
  • Hands-on experience with Java/J2EE (Spring, Spring Boot), Python, and Shell Scripting.
  • Proficiency with Oracle and MongoDB databases.
  • Experience with messaging tools like Kafka or MQ.
  • Working knowledge of APM tools including Splunk, GCL, ELK, Grafana, or Prometheus.
  • Experience with distributed storage (NFS) and resource management (PCF, Kubernetes, OpenShift, AWS, or Azure).
  • Ability to write Ansible playbooks using YAML.
  • Familiarity with caching tools like Redis or memcache.
  • Bachelor's degree in Computer Science or related field.

Preferred Skills

  • Working knowledge of CI/CD pipelines using Git, Jenkins, or UCD Release.

Education

Bachelor's degree

Related Jobs

No related jobs found

← Back to jobs