← Back to jobs

Saransh Inc Logo
Genesis Operations Support

Saransh Inc

 

Houston, TX, USA

Posted On: 3 days ago
Experience: 8+ years
Availability: Hybrid
Openings: 1
Category: Operations Support
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

Genesis Operations Support – As a Genesis Operations, you will operate and improve highly available production systems across bare-metal and cloud environments. You will focus on reliability, scalability, observability, incident management, infrastructure automation, and TOIL reduction, while collaborating with Application, Infrastructure, Network, Security, Database, and DevOps teams.

 

Primary Skill –

  • Genesis / Bare-Metal Infrastructure, Linux Administration
  • Bare-metal provisioning: PXE/iPXE, DHCP, BMC, IPMI, Redfish, iDRAC/iLO
  • Terraform, Ansible, Infrastructure as Code
  • Python, Bash, Go
  • Jenkins, GitHub Actions, GitLab CI/CD
  • Prometheus, Grafana, Datadog, Splunk, ELK, OpenTelemetry
  • Networking: TCP/IP, DNS, HTTP/HTTPS, TLS, proxies, load balancing

 

Secondary Skill – Git, security hardening, vulnerability management, certificates/secrets, storage, database troubleshooting, disaster recovery and configuration management.

 

  1. 8+ years of experience in DevOps, Production, Infrastructure or Systems Support.
  2. Strong hands-on Linux administration and production troubleshooting experience.
  3. Need a technical support professional with strong experience in end-user support and excellent clear communication skills.
  4. Operate and maintain Genesis/bare-metal infrastructure and automate server lifecycle management.
  5. Define and maintain SLIs, SLOs, SLAs and error budgets for critical services.
  6. Build monitoring, alerting, logging, dashboards and distributed tracing solutions.
  7. Automate infrastructure provisioning and configuration using Terraform/Ansible and scripting.
  8. Manage/troubleshoot Kubernetes, containerized workloads, networking, storage and deployments.
  9. Handle P1/P2 incidents, service restoration, RCA, postmortems and preventive actions.
  10. Troubleshoot across Linux, networking, DNS, load balancers, storage, Kubernetes, databases and applications.
  11. Identify repetitive operational activities and automate them to reduce TOIL.
  12. Perform capacity planning, performance tuning and infrastructure optimization.
  13. Maintain runbooks, operational procedures, production-readiness and disaster-recovery documentation.
  14. Participate in on-call support and collaborate across engineering teams to improve reliability and resilience.
  15. Perform application repaves and support end users in scheduled datacenter activities (including weekends when needed)


 

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs