← Back to jobs

Resourcesys Logo
Site Reliability Engineer

Resourcesys

 

Houston, TX, USA

Posted On: 6 days ago
Experience: 5+ years
Availability: Hybrid
Openings: 1
Category: site reliability engineering
Tenure: Contract - Corp-to-Corp
Related Jobs

No related jobs found

Description

You will own the stability, functionality, and reliability of applications within the Controls, Operational Risk, Compliance and Practices Technology domain.

Responsibilities

  • Troubleshoot incidents and conduct blameless post-mortems to ensure permanent closure of issues.
  • Automate manual operational work by designing, developing, and testing software.
  • Engage with development teams to implement self-healing and resiliency patterns.
  • Design and conduct performance tests to identify bottlenecks and optimization opportunities.
  • Drive adoption of monitoring frameworks to achieve end-to-end flow monitoring and noise-free alerting.

Required Skills

  • 5+ years of experience in site reliability or related engineering roles.
  • Expertise in Java, J2EE, or Python.
  • Hands-on experience with Cloud Foundry, Kubernetes, or AWS.
  • Proficiency with big data services including Hadoop, HDFS, Hive, Yarn, HBase, Kafka, and ZooKeeper.
  • Experience with deployment automation, CI/CD, DevOps, Jenkins, GIT, and BitBucket.
  • Strong debugging skills and expertise in performance monitoring and capacity management.
  • Experience developing and debugging distributed systems in Linux and Hadoop environments.

Preferred Skills

  • Experience with monitoring tools such as AppD, Splunk, ELK, or Geneos.
  • Familiarity with analytical tools like Tableau or Alteryx.

Education

Bachelor’s Degree

Related Jobs

No related jobs found

← Back to jobs