← Back to jobs

TalentOla Logo
Site Reliability Engineer

TalentOla

 

Plano, TX, USA

Posted On: 7 days ago
Experience: 5+ years
Availability: Onsite
Openings: 1
Category: site reliability engineering
Tenure: No Preference/Any
Related Jobs

No related jobs found

Description

You will own the migration of infrastructure, applications, and databases from on-premises environments to AWS.

This role is on-site.

Responsibilities

  • Execute on-prem to AWS migrations involving EC2, EKS, S3, RDS, and DynamoDB.
  • Build and maintain CI/CD pipelines using Jenkins, CodeDeploy, and CodePipeline.
  • Automate infrastructure provisioning, patching, backups, and configuration management using IaC.
  • Support Spark data pipelines and debug PySpark issues in production environments.
  • Implement disaster recovery and systems resilience strategies for data pipelines and web applications.

Required Skills

  • 9 years of DevOps or SRE experience.
  • 5 years of coding experience in Python, Bash, or Java.
  • 5 years of experience with Terraform and CI/CD infrastructure.
  • 5 years of experience in centralized monitoring and logging using CloudWatch, Grafana, or Datadog.
  • 5 years of provisioning and configuration management across Dev, QA, and Prod environments.
  • 3 years of experience supporting Spark data pipelines and PySpark debugging.
  • 3 years of security implementation including IAM roles, OAuth, SSO, and CloudTrail.
  • Hands-on expertise with Docker, Kubernetes, or ECS.

Preferred Skills

  • Databricks SRE experience.

Education

Any Graduate

Related Jobs

No related jobs found

← Back to jobs