CockroachDB Database Administrator

• Posted 7 hours ago • Updated 2 minutes ago
Full Time
Part Time
Fitment

Dice Job Match Score™

👾 Reticulating splines...

Job Details

Skills

  • Machine Vision
  • Scalability
  • Operational Excellence
  • Amazon S3
  • VPI
  • Network
  • CPU
  • Testing
  • Provisioning
  • Scripting
  • Operational Efficiency
  • Quality Control
  • Dashboard
  • Storage
  • D3.js
  • UG
  • Standard Operating Procedure
  • SAP WM
  • Replication
  • High Availability
  • Clustering
  • Docker
  • Computer Networking
  • Grafana
  • Performance Monitoring
  • Bash
  • Python
  • Shell Scripting
  • Backup
  • Backup & Restore
  • Recovery
  • Disaster Recovery
  • Failover
  • Business Continuity Planning
  • Incident Management
  • Root Cause Analysis
  • Capacity Management
  • Performance Tuning
  • Computer Science
  • Information Technology
  • Database Administration
  • SQL
  • Management
  • SQL Tuning
  • Database Performance Tuning
  • Linux Administration
  • Cloud Computing
  • Amazon Web Services
  • Microsoft Azure
  • Google Cloud
  • Google Cloud Platform
  • Analytical Skill
  • Communication
  • Documentation
  • Kubernetes
  • ZK
  • Terraform
  • Ansible
  • PostgreSQL
  • Database
  • Continuous Integration
  • Continuous Delivery
  • DevOps
  • Production Support

Summary

Position: Senior CockroachDB Database Administrator (CockroachDB DBA)

Experience Required

8-10 Years (with strong experience in distributed databases and production database administration)

Job Summary

We are seeking an experienced CockroachDB Database Administrator (DBA) to design, deploy, administer, and optimize CockroachDB clusters in large-scale production environments. The ideal candidate should have hands-on experience managing distributed SQL databases, ensuring high availability, disaster recovery, performance optimization, monitoring, automation, and production support.

The candidate will be responsible for maintaining highly available, fault-tolerant, multi-region CockroachDB clusters while collaborating with infrastructure, cloud, and application teams to ensure database reliability, scalability, and operational excellence.



Key Responsibilities

CockroachDB Administration

  • Design, deploy, configure, and manage CockroachDB clusters in production environments.
  • Build and maintain multi-region distributed database clusters.
  • Ensure high availability, fault tolerance, and data consistency across geographically distributed environments.
  • Monitor cluster health, node status, replication, latency, and resource utilization.
  • Perform database capacity planning and cluster scaling.

Performance Optimization

  • Monitor and optimize SQL query performance.
  • Analyze execution plans and identify bottlenecks.
  • Optimize database schema design and indexing strategies.
  • Resolve high-latency and throughput issues.
  • Address hotspotting, leaseholder imbalance, and replication lag.

Troubleshooting & Incident Management

  • Troubleshoot complex database and infrastructure issues, including:
    • Node failures
    • Network partitions
    • Replication issues
    • Leaseholder imbalance
    • Range imbalance
    • High CPU or memory utilization
    • Storage issues
    • Performance bottlenecks
  • Participate in incident management and on-call support.

Backup & Disaster Recovery

  • Design and implement disaster recovery strategies.
  • Configure backup and restore processes.
  • Implement Point-in-Time Recovery (PITR).
  • Validate backup integrity through regular recovery testing.
  • Manage failover and failback procedures.

Automation & Database Operations

  • Automate provisioning, deployment, upgrades, scaling, and maintenance of CockroachDB clusters.
  • Perform rolling upgrades with minimal or zero downtime.
  • Develop automation scripts using Bash, Python, or similar scripting languages.
  • Improve operational efficiency through Infrastructure as Code (IaC).

Monitoring & Observability

  • Monitor database health using observability tools.
  • Configure alerts and dashboards.
  • Track:
    • Cluster health
    • Replication status
    • Latency
    • Resource utilization
    • Storage growth
    • Query performance
  • Work with monitoring tools such as Grafana, Prometheus, and Cloud monitoring platforms.

Documentation

  • Create operational runbooks.
  • Develop Standard Operating Procedures (SOPs).
  • Maintain architecture diagrams and recovery procedures.
  • Document production support processes and troubleshooting guides.



Required Technical Skills

Database Technologies

  • CockroachDB
  • Distributed SQL Databases
  • SQL Performance Tuning
  • Database Replication
  • High Availability (HA)
  • Multi-region Database Deployment
  • Database Clustering

Cloud & Infrastructure

  • AWS / Azure / Google Cloud Platform
  • Linux Administration
  • Kubernetes (Preferred)
  • Docker
  • Networking Fundamentals

Monitoring & Observability

  • Prometheus
  • Grafana
  • Cloud Monitoring Tools
  • Performance Monitoring
  • Alert Management

Automation

  • Bash
  • Python
  • Shell Scripting
  • Terraform (Preferred)
  • Ansible (Preferred)

Backup & Recovery

  • Backup & Restore
  • Point-in-Time Recovery (PITR)
  • Disaster Recovery
  • Failover / Failback
  • Business Continuity Planning

Operations

  • Production Support
  • Incident Management
  • Root Cause Analysis (RCA)
  • Capacity Planning
  • Performance Tuning



Required Qualifications

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
  • 8-10 years of database administration experience.
  • Strong hands-on experience with CockroachDB or other distributed SQL databases.
  • Experience managing production database environments.
  • Expertise in SQL performance tuning and database optimization.
  • Experience with Linux system administration.
  • Knowledge of cloud platforms (AWS, Azure, or Google Cloud Platform).
  • Strong troubleshooting and analytical skills.
  • Excellent communication and documentation abilities.



Preferred Qualifications

  • Experience with Kubernetes-based database deployments.
  • Knowledge of Infrastructure as Code (Terraform, Ansible).
  • Experience with PostgreSQL or other distributed databases.
  • Experience with CI/CD pipelines and DevOps practices.
  • CockroachDB certification (if available).
  • Experience in mission-critical production support environments.

Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91018020
  • Position Id: PDT - 11711-12846-1785158905
  • Posted 7 hours ago

Company Info

About Purple Drive Technologies LLC

Founded in 2007, Purple Drive started as a tech solutions firm and has grown into a full-service consulting and talent partner. We help businesses navigate complex technology challenges while connecting top professionals with career-defining opportunities.

We believe in transforming businesses through smart IT solutions and empowering technologists to grow their expertise through challenging projects and meaningful partnerships. Built on over 20 years of trusted relationships, we create success stories for both our clients and the talented professionals who drive innovation forward.

Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Sunnyvale, California

Today

Easy Apply

Full-time, Part-time, Third Party, Contract

Pittsburgh, Pennsylvania

Today

Easy Apply

Full-time, Part-time, Contract, Third Party

Search all similar jobs