← Back to jobs

Brillius Inc Logo
Site Reliability Engineer

Brillius Inc

 

San Jose, CA, USA

Posted On: Just posted
Experience: 8+ years
Availability: Hybrid
Openings: 1
Category: Site Reliability Engineer
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will own core infrastructure reliability and automation tasks.

This role is on-site.

Responsibilities

  • Create and support automation scripts using shell, Ansible, and Python for infrastructure deployment, validation, and monitoring.
  • Manage incident handling and problem management processes.
  • Schedule monitoring scripts using cron and Airflow.
  • Maintain Linux environments, managing shells, filesystems, and utilities on RHEL/CentOS.
  • Operate and manage distributed computing resources using container orchestration frameworks like Kubernetes.

Required Skills

  • 8+ years of professional experience.
  • Deep experience with Linux flavors (RHEL/CentOS), shells, and filesystem utilities.
  • Required knowledge of AWS cloud platform.
  • Experience with container orchestration, specifically Kubernetes objects (on-prem and Rancher).
  • Proficiency in creating and supporting automation scripts (Shell, Ansible, Python).
  • Experience with monitoring tools such as Grafana, Dynatrace, or Apica.
  • Database knowledge, including SQL and NoSQL databases.
  • Experience with storage concepts, including volume management, backups, and DR planning.
  • Familiarity with CI/CD pipeline development.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs