← Back to jobs

StaffXpert LLC Logo
Senior CockroachDB Database Engineer

StaffXpert LLC

 

Sunnyvale, CA, USA

Posted On: 15+ days ago
Experience: 7+ years
Availability: Onsite
Openings: 1
Category: CockroachDB Database Engineer
Tenure: Contract - Corp-to-Corp
Related Jobs

No related jobs found

Description

Key Responsibilities

Design, deploy, administer, and maintain production-grade CockroachDB clusters in cloud and hybrid environments.

Monitor database health, performance, availability, and resource utilization to ensure reliable operations.

Perform database performance tuning, query optimization, indexing strategies, and capacity planning.

Implement and manage backup, recovery, disaster recovery, and business continuity solutions.

Develop automation and Infrastructure-as-Code (IaC) solutions to streamline provisioning, upgrades, and operational tasks.

Establish and maintain Site Reliability Engineering (SRE) practices, including SLIs, SLOs, and Error Budgets.

Lead incident response, troubleshooting, root cause analysis (RCA), and post-incident remediation activities.

Build and maintain monitoring, logging, and alerting solutions using industry-standard observability tools.

Collaborate with engineering, DevOps, and infrastructure teams to improve platform reliability, scalability, security, and performance.

Support database migrations, production releases, version upgrades, and modernization initiatives.

Participate in on-call support for critical production environments.

Implement database security, access controls, auditing, and compliance best practices.

Required Qualifications

Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field, or equivalent practical experience.

7+ years of experience in database engineering, administration, or site reliability engineering.

Strong hands-on experience administering and supporting CockroachDB or similar distributed SQL database platforms.

Deep understanding of distributed systems, cluster management, replication, and high-availability architectures.

Proven expertise in database performance tuning, query optimization, and capacity planning.

Experience designing and implementing backup, recovery, and disaster recovery strategies.

Strong knowledge of Site Reliability Engineering (SRE) principles and operational best practices.

Experience with incident management, reliability engineering, and production support.

Hands-on experience with AWS, Azure, or Google Cloud Platform (GCP).

Proficiency with Infrastructure-as-Code tools such as Terraform or Ansible.

Strong Linux/Unix administration skills.

Scripting experience using Python, Shell, Go, or similar languages.

Experience with CI/CD pipelines and automation frameworks.

Excellent analytical, troubleshooting, communication, and collaboration skills.

Preferred Qualifications

Experience supporting large-scale, mission-critical distributed systems.

Knowledge of Kubernetes and containerized platforms.

Experience with observability tools such as Prometheus, Grafana, Datadog, ELK, Splunk, or similar solutions.

Understanding of database security, governance, compliance, and auditing requirements.

CockroachDB certification or equivalent expertise in distributed database technologies.

Experience with PostgreSQL internals and PostgreSQL-compatible ecosystems.

Knowledge of multi-region architectures, distributed consensus mechanisms, and cloud-native platforms.

Experience in FinTech, Retail, E-Commerce, SaaS, or other high-scale environments

Education

Bachelor's degree

Related Jobs

No related jobs found

← Back to jobs