×
Register Here to Apply for Jobs or Post Jobs. X

Senior CockroachDB Database Engineer

Job in Sunnyvale, Santa Clara County, California, 94087, USA
Listing for: Programmers.io
Full Time position
Listed on 2026-08-03
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cloud Computing: Infrastructure & Operations, Disaster Recovery IT
Salary/Wage Range or Industry Benchmark: 140000 - 190000 USD Yearly USD 140000.00 190000.00 YEAR
Job Description & How to Apply Below

Location:
Sunnyvale, CA

Duration :

Full Time

Experience: 5-10 Years

Job Summary

We are seeking a highly skilled Cockroach DB Database Engineer with strong Site Reliability Engineering (SRE) experience to design, implement, manage, and optimize large-scale distributed database platforms. The ideal candidate will have hands-on expertise in Cockroach DB administration, performance tuning, high availability, disaster recovery, automation, observability, and operational reliability. The role requires close collaboration with development, infrastructure, and platform engineering teams to ensure highly available, resilient, and scalable database services.

Key Responsibilities
  • Design, deploy, administer, and maintain production-grade Cockroach DB clusters across cloud and on-premises environments.
  • Monitor database health, performance, latency, throughput, and resource utilization to ensure service reliability and availability.
  • Implement and manage backup, restore, disaster recovery, and business continuity strategies.
  • Perform database capacity planning, performance tuning, indexing, and query optimization.
  • Develop automation scripts and Infrastructure-as-Code (IaC) solutions to streamline provisioning, upgrades, and operational tasks.
  • Establish and manage SRE practices including Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets.
  • Drive incident management, root cause analysis (RCA), postmortems, and preventive remediation activities.
  • Build and maintain monitoring, logging, and alerting solutions using tools such as Prometheus, Grafana, ELK, Datadog, or similar platforms.
  • Collaborate with Dev Ops and Engineering teams to improve platform reliability, scalability, security, and operational excellence.
  • Support production releases, database migrations, version upgrades, and platform modernization initiatives.
  • Participate in on-call rotation and provide support for critical production incidents.
  • Implement database security controls, access governance, auditing, and compliance best practices.
Required Skills
  • Strong hands-on experience with Cockroach DB Administration
  • Expertise in distributed SQL databases and cluster management
  • Database performance tuning and query optimization
  • Backup, recovery, replication, and data protection strategies
  • High Availability and Disaster Recovery architecture
Site Reliability Engineering (SRE)
  • Strong understanding of SRE principles and operational excellence
  • Experience defining and tracking SLIs, SLOs, and Error Budgets
  • Incident response, RCA, and reliability engineering practices
  • Production monitoring, observability, and capacity management
  • Reliability automation and operational process improvement
  • Experience with AWS, Azure, or GCP
  • Infrastructure as Code (Terraform, Ansible, etc.)
  • Scripting using Python, Shell, or Go
  • CI/CD pipeline integration and automation
Preferred Qualifications
  • Experience supporting large-scale, mission-critical distributed systems.
  • Knowledge of Kubernetes and containerized deployments.
  • Experience with observability platforms such as Prometheus, Grafana, ELK, Datadog, or Splunk.
  • Understanding of security, compliance, and governance requirements for database platforms.
  • Cockroach DB certification or equivalent distributed database expertise is highly desirable.
Soft Skills
  • Strong analytical and problem-solving capabilities.
  • Excellent stakeholder communication and collaboration skills.
  • Ability to work independently in a fast-paced production environment.
  • Strong ownership mindset with a focus on reliability and customer experience.
#J-18808-Ljbffr
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary