Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior CockroachDB Database Engineer / Site Reliability Engineer (SRE)

Staffxpert LLC

a { text-decoration: none; color: #464feb; } tr th, tr td { border: 1px solid #e6e6e6; } tr th { background-color: #f5f5f5; }

Senior CockroachDB Database Engineer / Site Reliability Engineer (SRE)

Location: Austin, TX or Sunnyvale, CA (Onsite)

About Us

STAFFXPERT LLC is a trusted talent solutions partner connecting highly skilled professionals with leading organizations across technology, engineering, cloud infrastructure, and digital transformation initiatives. We are committed to matching exceptional talent with innovative opportunities that drive business success and career growth.

Job Summary

STAFFXPERT LLC is seeking a Senior CockroachDB Database Engineer / Site Reliability Engineer (SRE) on behalf of our client in Austin, TX or Sunnyvale, CA.

This role is ideal for an experienced database and reliability engineering professional who excels in managing large-scale distributed database platforms. The successful candidate will be responsible for designing, administering, optimizing, and supporting highly available CockroachDB environments while driving automation, observability, performance, and operational excellence across mission-critical systems.

Key Responsibilities

  • Design, deploy, administer, and maintain production-grade CockroachDB clusters in cloud and hybrid environments.

  • Monitor database health, performance, availability, and resource utilization to ensure reliable operations.

  • Perform database performance tuning, query optimization, indexing strategies, and capacity planning.

  • Implement and manage backup, recovery, disaster recovery, and business continuity solutions.

  • Develop automation and Infrastructure-as-Code (IaC) solutions to streamline provisioning, upgrades, and operational tasks.

  • Establish and maintain Site Reliability Engineering (SRE) practices, including SLIs, SLOs, and Error Budgets.

  • Lead incident response, troubleshooting, root cause analysis (RCA), and post-incident remediation activities.

  • Build and maintain monitoring, logging, and alerting solutions using industry-standard observability tools.

  • Collaborate with engineering, DevOps, and infrastructure teams to improve platform reliability, scalability, security, and performance.

  • Support database migrations, production releases, version upgrades, and modernization initiatives.

  • Participate in on-call support for critical production environments.

  • Implement database security, access controls, auditing, and compliance best practices.

Required Qualifications

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field, or equivalent practical experience.

  • 7+ years of experience in database engineering, administration, or site reliability engineering.

  • Strong hands-on experience administering and supporting CockroachDB or similar distributed SQL database platforms.

  • Deep understanding of distributed systems, cluster management, replication, and high-availability architectures.

  • Proven expertise in database performance tuning, query optimization, and capacity planning.

  • Experience designing and implementing backup, recovery, and disaster recovery strategies.

  • Strong knowledge of Site Reliability Engineering (SRE) principles and operational best practices.

  • Experience with incident management, reliability engineering, and production support.

  • Hands-on experience with AWS, Azure, or Google Cloud Platform (GCP).

  • Proficiency with Infrastructure-as-Code tools such as Terraform or Ansible.

  • Strong Linux/Unix administration skills.

  • Scripting experience using Python, Shell, Go, or similar languages.

  • Experience with CI/CD pipelines and automation frameworks.

  • Excellent analytical, troubleshooting, communication, and collaboration skills.

Preferred Qualifications

  • Experience supporting large-scale, mission-critical distributed systems.

  • Knowledge of Kubernetes and containerized platforms.

  • Experience with observability tools such as Prometheus, Grafana, Datadog, ELK, Splunk, or similar solutions.

  • Understanding of database security, governance, compliance, and auditing requirements.

  • CockroachDB certification or equivalent expertise in distributed database technologies.

  • Experience with PostgreSQL internals and PostgreSQL-compatible ecosystems.

  • Knowledge of multi-region architectures, distributed consensus mechanisms, and cloud-native platforms.

  • Experience in FinTech, Retail, E-Commerce, SaaS, or other high-scale environments.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior CockroachDB Database Engineer / Site Reliability Engineer (SRE) in Sunnyvale, CA vacancy
  • $101k - $161k

     ...awards, such as Best Engineering Team, Best Company for...  ...WithWe’re looking for Site Reliability Engineers to join our...  ...Service (CVaaS) global SRE team. SREs at Arista combine...  ...different types of databases, both directly on...  ...EngineeringExperience level: Mid-Senior LevelIndustry:... 
    Senior

    Arista Networks

    Santa Clara, CA
    3 days ago
  • $167.7k - $245.2k

     ...approximately 2 days per week on-site at Cisco offices in...  ...as intended, improving reliability and reducing risks. This...  ...observability and control.As a Senior Site Reliability Engineer (SRE), you will build, operate...  ...components—including databases and services—to improve... 
    Senior
    Full time
    Temporary work
    Local area
    Flexible hours
    2 days per week

    CISCO Systems

    Sunnyvale, CA
    4 days ago
  • $186.9k - $267.7k

     ...approximately 2 days per week on-site at Cisco offices in either...  ...as intended, improving reliability and reducing risks. This unified...  ...As a Staff Site Reliability Engineer (SRE), you will provide technical...  ...deployment infrastructure, databases, and networking.Drive... 
    Suggested
    Full time
    Temporary work
    Local area
    Flexible hours
    2 days per week

    CISCO Systems

    Sunnyvale, CA
    7 hours ago
  • $168k - $270.25k

    NVIDIA is looking for a Senior Site Reliability Engineer (SRE) to join its GeForce Now (GFN) team. SRE at NVIDIA ensures that our internal and external-facing GPU cloud gaming services have reliability and uptime as promised to the users and at the same time enables developers... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  •  ...automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud...  ...Who You AreExperienced Architect: 5+ years of experience in SRE, DevOps, or Systems Engineering, with a proven track record... 
    Senior
    Full time
    Work at office
    2 days per week

    LeanData

    Santa Clara, CA
    7 hours ago
  • $90k - $180k

     ...serve people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale,...  ...and mission-driven Senior Site Reliability Engineer (SRE) to join our DevOps team. In this critical role, you will... 
    Senior
    Remote work

    Abbott

    Sunnyvale, CA
    2 days ago
  • $148k - $235.75k

     ...on the world.Join our team of innovative engineers who are building an AI Data Center AIOps...  ...that turns raw, high-volume telemetry into reliable, job-centric insights and automation for...  ...operating production distributed systems as SRE/DevOps/Platform Ops.Proven ownership of... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $152k - $241.5k

     ...artificial intelligence.We’re looking for a Senior SRE to join our Compute Farm team and help...  ...host lifecycle management, fleet reliability/auto-healing, E2E observability or data...  ...Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    7 hours ago
  • $210.6k - $305.1k

     ...soil. Lead, inspire, and develop a talented SRE team, fostering a culture of innovation,...  ...:  You have led a distributed team of 5+ engineers, can demonstrate strong technical vision...  ...insurance. Please see the Cisco careers site to discover more benefits and perks. Employees... 
    Senior
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    Los Altos, CA
    2 days ago
  • $145k - $165k

     ...A technology solutions firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key... 
    Senior

    Bolt Graphics, Inc.

    Sunnyvale, CA
    5 days ago
  •  ...Job Description About the Role Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking...  ...services, Kubernetes platforms, databases, messaging systems, and cloud infrastructure...  .... ~ Strong knowledge of SRE principles, including SLOs, SLIs, error... 
    Senior

    Kody

    Sunnyvale, CA
    a month ago
  • $174k - $253k

     ...Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent...  ...Science or Engineering. ABOUT THE JOB: Site Reliability Engineering (SRE) is what you get when you treat...  ...systems ranging from planet-spanning databases [ to near real-time scalable data... 
    Senior

    Socket

    Sunnyvale, CA
    4 days ago
  •  ...of your updated resumes Title: Sr. SRE / DevOps Engineer Location: Sunnyvale, CA (Only...  ...Sunnyvale, California location. As Site Reliability Engineer, the individual will work closely...  ...using Java or Python CI/CD Database Vertica, Snowflake. Behavioral Skills... 
    Senior
    Local area
    Immediate start

    Donato Technologies Inc

    Sunnyvale, CA
    more than 2 months ago
  • $120k - $180k

     ...cybersecurity starts with you.Sr. SRE & DevOps EngineerAbout the Role:At CrowdStrike, our engineering organization depends on shared...  ...ownership to operate reliably, scale safely, harden for security...  ...systems including relational databases (PostgreSQL), NoSQL (Cassandra... 
    Full time
    Work experience placement
    Work at office
    Local area

    CrowdStrike

    Sunnyvale, CA
    4 days ago
  • $144k - $230k

     ...join the team and see how you can make a lasting impact on the world.We are seeking a passionate AI Tools Engineer to join the Site Reliability Engineering (SRE) Data Team. Applicants with SRE or equivalent experience are encouraged.What you will be doing:You will build... 
    Senior
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    4 days ago
  • $100k - $200k

     ...OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible...  ...as AWS, Azure, and Google Cloud, including services like databases, middleware, message queues, and load balancers.... 
    Full time

    OPPO

    Palo Alto, CA
    22 hours ago
  • $168k - $270.25k

     ...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually...  ...and security products. Work with your global SRE team to optimize operations, increase efficiency in our use... 
    Senior
    Flexible hours

    Sumo Logic

    San Jose, CA
    3 days ago
  •  ...Job Description Job Description Java SRE Engineer Onsite San Francisco Bay Area Infrastructure Engineer (2 Positions) We are...  ...platforms. This role is focused on infrastructure, reliability, and automation , with Java exposure as a supporting skill.... 

    Eitacies Inc

    Santa Clara, CA
    more than 2 months ago
  • $262k - $364k

     ...and evolve systems by pushing for changes that improve reliability and velocity.Practice sustainable incident response...  ...qualifications:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) combines software and systems engineering to build... 
    Senior

    Google

    San Jose, CA
    3 days ago
  • $100k - $170k

     ...We are seeking a  DevOps Engineer who is eager to have an immediate...  ...on Amazon EKS with focus on reliability and performance Build and...  ...~5+ years of experience in SRE, DevOps, or Platform Engineering...  ...optimization ~ Experience with database operations and scaling (RDS,... 
    Full time
    Work at office
    Immediate start
    Visa sponsorship
    Night shift

    E-Space

    Saratoga, CA
    a month ago
  •  ...A leading technology firm is in search of a Senior Wireless Network Site Reliability Engineer to manage and enhance their wireless network infrastructure. The ideal candidate has over 8 years of experience in wireless network operations and a strong background in wireless... 
    Senior

    TechDigital Group

    Santa Clara, CA
    5 days ago
  • Elevate your engineering prowess to unprecedented levels by joining...  ...among the top echelon in site reliability. As a Senior Lead Site Reliability...  ...systems; and expertise in graph databases (Neo4j, TigerGraph),...  ...projects, particularly in SRE, observability, or AI/ML domains... 
    Senior

    JP Morgan Chase

    Palo Alto, CA
    7 hours ago
  •  ..., and the challenges of building in a high-growth startup, we’d love to talk. This is more than a job—it’s a journey. Site Reliability Engineers (SREs) are responsible for the overall performance and reliability of ASAPP's infrastructure and products. The team owns... 
    Senior
    Remote work

    ASAPP

    Mountain View, CA
    a month ago
  • $113.3k - $205.52k

     ...#OneJamf. What you'll do at Jamf: As a Senior Site Reliability Engineer, you'll help us balance development...  ...meaningful part in shaping how we practice SRE at Jamf. You may be required to work...  ...optimizing SQL queries and database engine tuning. (Preferred) Experience... 
    Senior
    Work at office
    Worldwide
    Flexible hours

    GrabJobs

    San Jose, CA
    2 days ago
  •  ...DESCRIPTION Elevate your engineering prowess to unprecedented levels...  ...among the top echelon in site reliability. As a Senior Lead Site Reliability...  ...systems; and expertise in graph databases (Neo4j, TigerGraph),...  ...projects, particularly in SRE, observability, or AI/ML domains... 
    Senior

    J.P. Morgan

    Palo Alto, CA
    2 days ago
  •  ...Build and operate one or more bounded contexts of the NeoCloud SRE platform — the multi-region substrate that observes, protects, and...  ..., collection-monitor. Alert, Correlation & SLO: alert-engine-framework, alert-correlation, slo-framework, default M-series alert... 
    Senior
    Full time
    Contract work
    Local area

    Bitdeer Technologies Group

    San Jose, CA
    a month ago
  • $230k - $250k

     ...foundation for autonomous networking, giving engineers and AI agents the ability to know the impact...  ...have always been done.Forward is looking for a Site Reliability EngineerAbout the Role This is not a "keep the lights on" SRE role. As our first or early SRE hire you will... 
    Night shift

    Forward Networks

    Santa Clara, CA
    7 hours ago
  • $170k - $200k

     ...are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will...  ...automation, system reliability, and a SRE mindset of continuous improvement.Key...  ...infrastructure (Linux servers, network devices, databases etc,.) Improve observability with... 
    Full time
    Worldwide

    Fortinet

    Sunnyvale, CA
    2 days ago
  • $272k - $431.25k

    NVIDIA is looking for a Cloud Site Reliability Engineering Architect to work in IPP's (Infrastructure, Planning...  ...them?What you'll be doing:Serve as an SRE Architect part of GPU Private Cloud...  ....Experience in working with SQL/NoSQL database systems such as MySQL, Cassandra,... 
    Full time
    Work experience placement
    Worldwide

    Nvidia

    Santa Clara, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior CockroachDB Database Engineer / Site Reliability Engineer (SRE). Be the first to apply!