Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

Neshent Technologies

We are seeking a Senior Database Reliability Engineer (DBRE) to design, operate, and improve reliable, scalable, secure, and highly available database and data platforms. The role combines database engineering, site reliability engineering, Linux systems administration, and infrastructure automation.

The ideal candidate will have strong production experience with PostgreSQL, AWS, Kubernetes, infrastructure as code, and distributed data systems and will work closely with SRE, application, security, and infrastructure teams.

Roles & Responsibilities
  • Design, administer, maintain, and secure PostgreSQL environments in production and cloud environments.
  • Manage PostgreSQL deployments on Kubernetes and Amazon RDS.
  • Perform database administration activities including installation, upgrades, patching, backup/recovery, monitoring, capacity planning, and performance tuning.
  • Troubleshoot production issues across application, database, operating system, storage, and network layers.
  • Design and maintain ETL pipelines and develop scripts and procedures for data migration.
  • Build and automate database and infrastructure operations using Terraform, Ansible, Chef, or Puppet.
  • Develop operational tooling and automation using Python, Bash, Go, Ruby, or Perl.
  • Monitor database health, performance, availability, and capacity and implement improvements.
  • Participate in on-call rotations, incident response, alerting, and post-incident reviews.
  • Improve reliability practices through SLOs, disaster recovery, backup validation, and failover testing.
  • Build database platform tooling and self-service workflows that enable application teams to use data services safely and efficiently.
  • Operate and support distributed data systems such as Kafka/MSK, ClickHouse, Redis, MySQL, Cassandra, or Elasticsearch.
  • Collaborate with SRE, application engineering, security, and infrastructure teams to drive reliable architectural changes.
  • Create technical documentation and design proposals and communicate solutions effectively to engineering stakeholders.
Required Qualifications
  • 5+ years of experience designing, operating, and troubleshooting PostgreSQL in production environments.
  • 5+ years managing production databases or distributed data systems across application, database, OS, storage, and network layers.
  • Hands-on experience with AWS, Amazon RDS, Kubernetes, Terraform, service discovery, and secrets management.
  • 3+ years of Linux systems engineering experience, including performance tuning, memory management, I/O tuning, security, configuration, and networking.
  • Experience automating infrastructure or database operations using Terraform, Ansible, Chef, or Puppet.
  • 2+ years of scripting or programming experience with Python, Bash, Go, Ruby, or Perl.
  • Experience working with at least one non-PostgreSQL data platform such as Kafka/MSK, ClickHouse, Redis, MySQL, Cassandra, or Elasticsearch.
  • Strong understanding of PostgreSQL internals, including replication, concurrency, transactions, indexing, maintenance, backup/recovery, and query performance.
  • Strong troubleshooting, problem-solving, communication, and documentation skills.
Preferred Skills
  • Experience building database platform tooling, self-service workflows, and standardized operational processes.
  • Experience operating Kafka/MSK, ClickHouse, or Redis at scale, including clustering, replication, partitioning/sharding, retention, capacity planning, and workload tuning.
  • Experience defining and improving reliability practices for production data systems.
  • Strong knowledge of SLOs, disaster recovery, backup validation, failover, and resilience testing.
  • Experience with cloud-native database architectures and distributed systems.
  • Ability to create technical design documents and communicate complex database and reliability concepts effectively.
  • Ability to work effectively in a fast-paced, collaborative engineering environment.
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Los Gatos, CA vacancy
  • $168k - $270.25k

    NVIDIA is looking for a Senior Site Reliability Engineer (SRE) to join its GeForce Now (GFN) team. SRE at NVIDIA ensures that our internal and external-facing GPU cloud gaming services have reliability and uptime as promised to the users and at the same time enables developers... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $174k - $252k

     ...systems by pushing for changes that improve reliability and velocity.Practice sustainable...  ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical...  ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you... 
    Senior

    Google

    Sunnyvale, CA
    3 days ago
  • Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work with... 
    Senior
    Flexible hours

    Sumo Logic

    San Jose, CA
    1 day ago
  •  ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and...  ...and networking teams to improve service reliability and deployment workflowsDeploy and...  ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    15 hours ago
  • $168k - $270.25k

     ...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    15 hours ago
  •  ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering...  ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    1 day ago
  • LeanData helps the world’s fastest-growing companies automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly to the SVP of Engineering, this role is... 
    Senior
    Full time
    Work at office
    2 days per week

    LeanData

    Santa Clara, CA
    3 days ago
  • $267k - $356k

     ...day is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-...  ...workloads in the industry, which means reliability and performance aren't just goals—they're...  ...defined storage across new and existing sites using tools such as Ansible, Jenkins etc... 
    Senior
    Work experience placement
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    15 hours ago
  • $101k - $161k

     ...several prestigious awards, such as Best Engineering Team, Best Company for Diversity,...  ...DescriptionWho You'll Work WithWe’re looking for Site Reliability Engineers to join our growing Arista’s...  ...: EngineeringExperience level: Mid-Senior LevelIndustry: Computer Networking
    Senior

    Arista Networks

    Santa Clara, CA
    1 day ago
  • $148k - $235.75k

     ...see how you can make a lasting impact on the world.Join our team of innovative engineers who are building an AI Data Center AIOps platform that turns raw, high-volume telemetry into reliable, job-centric insights and automation for GPU fleets. We’re hiring a DevOps Engineer... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    15 hours ago
  • $90k - $180k

     ...nutritionals and branded generic medicines. Our 115,000 colleagues serve people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.We... 
    Senior
    Remote work

    Abbott

    Sunnyvale, CA
    15 hours ago
  • $160k - $240k

     ...millions of times a day - quickly, reliably, and securely. Any time you...  ...at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our...  ...operations or DevOps at a mid-to-senior level.Strong shell scripting... 
    Senior
    Full time

    Fiserv

    Sunnyvale, CA
    2 days ago
  • $248k - $396.75k

     ...supportive environment, where NVIDIANs are inspired to excel and make a profound global impact.NVIDIA is seeking a Senior Manager of Site Reliability Engineering to lead and reshape how IT operations function at scale. This role goes beyond traditional service management... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    11 hours ago
  • $222k - $300.5k

     ...OverviewAbout the TeamIntuit's Infrastructure and Site Reliability organization owns the operational...  .... The Fintech Platform Systems Engineering team builds and operates the AWS-based...  ...negotiable.The OpportunityWe're hiring a Senior Manager, Site Reliability Engineering to... 
    Senior
    Worldwide
    Shift work

    Intuit

    Mountain View, CA
    11 hours ago
  • $146.7k - $339.3k

    Immigration sponsorship is not available for this positionWhat you can expect As a Senior Lead Site Reliability Engineer, you can anticipate opportunities to work on our hybrid systems across the globe. You will be responsible for installing, configuring, and monitoring... 
    Senior
    Full time
    Work at office
    Remote work
    Worldwide
    Shift work
    Weekend work

    Zoom

    San Jose, CA
    3 days ago
  • $210.6k - $305.1k

     ...Minimum Qualifications:  You have led a distributed team of 5+ engineers, can demonstrate strong technical vision for your team, and ensure...  ..., and basic life insurance. Please see the Cisco careers site to discover more benefits and perks. Employees may be eligible... 
    Senior
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    15 hours ago
  • $145k - $165k

     ...A technology solutions firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key... 
    Senior

    Bolt Graphics, Inc.

    Sunnyvale, CA
    3 days ago
  •  ...Platform powers compute provisioning and infrastructure orchestration across our physical data centers. We are looking for a Senior Site Reliability Engineer to improve the reliability, scalability, and operational maturity of these systems as Lambda’s fleet and customer base... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    3 days ago
  • $168k - $270.25k

     ...deploy and run an AI data center. We take great pride in providing excellent, comprehensive support to our customers! ​Sr Site Reliability Engineer in this role will significantly impact and contribute to the overall success of both external customers running their clusters... 
    Senior
    Full time
    Worldwide

    Nvidia

    Santa Clara, CA
    3 days ago
  • $187.04k - $359.72k

     ...systems by pushing for changes that improve reliability and velocity. Qualifications Minimum...  ...degree in Computer Science, Electrical Engineering, Computer Engineering or related areas....  ...Product Ops, Corporate Functions and more. On-site presence across teams allows the company... 
    Senior
    Temporary work
    Local area
    Overseas
    Shift work

    Tik Tok

    San Jose, CA
    3 days ago
  • $174k - $252k

    Senior Software Engineer, Site Reliability Engineering X Applicants in San Francisco: Qualified applications with arrest or conviction records will be considered for employment in accordance with the San Francisco Fair Chance Ordinance for Employers and the California... 
    Senior
    Full time

    Google Inc.

    Sunnyvale, CA
    1 day ago
  •  ...Palo Alto Networks in Santa Clara seeks a visionary Senior Principal Engineer/Architect to serve as the technical authority for global SRE and Platform Engineering initiatives. You will be the primary architect and driver of our AI-driven Autonomous SRE transformation... 
    Senior

    Jobleads-US

    Santa Clara, CA
    3 hours ago
  • $125k - $160k

     ...Docker and Kubernetes.Collaborate with Product Management and engineering peers from concept through delivery.Maintain high engineering...  ....Continuously improve system performance, scalability, and reliability.Qualifications:7+ years of professional experience with Java.... 
    Senior
    Full time
    Remote work
    Flexible hours

    Centric Software

    Campbell, CA
    3 days ago
  • $166k - $244k

    A leading technology company located in Sunnyvale, California, is seeking a Site Reliability Engineer responsible for building and maintaining large-scale systems. The ideal candidate should possess a degree in Computer Science and have significant experience in programming... 
    Senior

    Google

    Sunnyvale, CA
    3 days ago
  • $150k - $200k

    Espace is seeking a Senior Software Engineer specializing in 5G Physical Layer to enhance IoT solutions from space. You will design and optimize algorithms for 5G systems while collaborating with teams across Europe, India, and the U.S. The role requires a strong background... 
    Senior
    Full time

    Espace

    Saratoga, CA
    3 days ago
  •  ...Software Engineer We are looking for a few exceptional software engineers to work on our cloud based B2B e-commerce, renewals and subscriptions platform. As a member of the engineering team, you will work with product management and other team members to design... 
    Senior
    Flexible hours

    Rainmaker Systems

    Campbell, CA
    6 hours ago
  • E-Space is seeking a Senior Software Engineer to join our Ground Software team, focusing on mission-critical Python-based back-end systems that operate a growing satellite constellation. You will design, implement, and scale real-time data ingestion pipelines, microservices... 
    Senior

    Medium

    Saratoga, CA
    1 day ago
  • $70k - $200k

     ...to go, in any context, for generations to come. Reports To Senior Software Engineering Manager What You Will Be Doing As a Senior Software...  ...driver base. What You Will Bring to ChargePoint Implement reliable APIs and microservices using Java and Spring Boot Contribute... 
    Senior

    ChargePoint

    Campbell, CA
    1 day ago
  • $130k - $200k

     ...Senior Software EngineerReady to make connectivity from space universally accessible, secure and actionable? Then you've come to the...  ...intelligence.What is the role?E-Space is looking for a Senior Software Engineer to join our Ground Software team. You will collaborate with... 
    Senior
    Immediate start

    eSpace

    Saratoga, CA
    1 day ago
  • Imperative Care in Campbell, California, is seeking a Staff Software Engineer in Robotics to design and implement critical software for their innovative robotic platform. This includes working on real-time algorithms and collaborating with cross-functional teams to meet... 
    Senior

    Imperative Care

    Campbell, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!