Senior Site Reliability Engineer
Neshent Technologies
We are seeking a Senior Database Reliability Engineer (DBRE) to design, operate, and improve reliable, scalable, secure, and highly available database and data platforms. The role combines database engineering, site reliability engineering, Linux systems administration, and infrastructure automation.
The ideal candidate will have strong production experience with PostgreSQL, AWS, Kubernetes, infrastructure as code, and distributed data systems and will work closely with SRE, application, security, and infrastructure teams.
Roles & Responsibilities
- Design, administer, maintain, and secure PostgreSQL environments in production and cloud environments.
- Manage PostgreSQL deployments on Kubernetes and Amazon RDS.
- Perform database administration activities including installation, upgrades, patching, backup/recovery, monitoring, capacity planning, and performance tuning.
- Troubleshoot production issues across application, database, operating system, storage, and network layers.
- Design and maintain ETL pipelines and develop scripts and procedures for data migration.
- Build and automate database and infrastructure operations using Terraform, Ansible, Chef, or Puppet.
- Develop operational tooling and automation using Python, Bash, Go, Ruby, or Perl.
- Monitor database health, performance, availability, and capacity and implement improvements.
- Participate in on-call rotations, incident response, alerting, and post-incident reviews.
- Improve reliability practices through SLOs, disaster recovery, backup validation, and failover testing.
- Build database platform tooling and self-service workflows that enable application teams to use data services safely and efficiently.
- Operate and support distributed data systems such as Kafka/MSK, ClickHouse, Redis, MySQL, Cassandra, or Elasticsearch.
- Collaborate with SRE, application engineering, security, and infrastructure teams to drive reliable architectural changes.
- Create technical documentation and design proposals and communicate solutions effectively to engineering stakeholders.
Required Qualifications
- 5+ years of experience designing, operating, and troubleshooting PostgreSQL in production environments.
- 5+ years managing production databases or distributed data systems across application, database, OS, storage, and network layers.
- Hands-on experience with AWS, Amazon RDS, Kubernetes, Terraform, service discovery, and secrets management.
- 3+ years of Linux systems engineering experience, including performance tuning, memory management, I/O tuning, security, configuration, and networking.
- Experience automating infrastructure or database operations using Terraform, Ansible, Chef, or Puppet.
- 2+ years of scripting or programming experience with Python, Bash, Go, Ruby, or Perl.
- Experience working with at least one non-PostgreSQL data platform such as Kafka/MSK, ClickHouse, Redis, MySQL, Cassandra, or Elasticsearch.
- Strong understanding of PostgreSQL internals, including replication, concurrency, transactions, indexing, maintenance, backup/recovery, and query performance.
- Strong troubleshooting, problem-solving, communication, and documentation skills.
Preferred Skills
- Experience building database platform tooling, self-service workflows, and standardized operational processes.
- Experience operating Kafka/MSK, ClickHouse, or Redis at scale, including clustering, replication, partitioning/sharding, retention, capacity planning, and workload tuning.
- Experience defining and improving reliability practices for production data systems.
- Strong knowledge of SLOs, disaster recovery, backup validation, failover, and resilience testing.
- Experience with cloud-native database architectures and distributed systems.
- Ability to create technical design documents and communicate complex database and reliability concepts effectively.
- Ability to work effectively in a fast-paced, collaborative engineering environment.
$168k - $270.25k
NVIDIA is looking for a Senior Site Reliability Engineer (SRE) to join its GeForce Now (GFN) team. SRE at NVIDIA ensures that our internal and external-facing GPU cloud gaming services have reliability and uptime as promised to the users and at the same time enables developers...SeniorFull time$174k - $252k
...systems by pushing for changes that improve reliability and velocity.Practice sustainable... ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical... ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you...Senior- Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually striving for sustained operational excellence of Sumo’s planet-scale observability and security products. Work with...SeniorFlexible hours
- ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building and... ...and networking teams to improve service reliability and deployment workflowsDeploy and... ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production Engineering...SeniorWork at officeLocal areaWork from homeFlexible hours
$168k - $270.25k
...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance...SeniorFull time- ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering... ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or...SeniorWork at officeLocal areaWork from homeFlexible hours
- LeanData helps the world’s fastest-growing companies automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly to the SVP of Engineering, this role is...SeniorFull timeWork at office2 days per week
$267k - $356k
...day is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-... ...workloads in the industry, which means reliability and performance aren't just goals—they're... ...defined storage across new and existing sites using tools such as Ansible, Jenkins etc...SeniorWork experience placementWork at officeLocal areaWork from homeFlexible hours$101k - $161k
...several prestigious awards, such as Best Engineering Team, Best Company for Diversity,... ...DescriptionWho You'll Work WithWe’re looking for Site Reliability Engineers to join our growing Arista’s... ...: EngineeringExperience level: Mid-Senior LevelIndustry: Computer NetworkingSenior$148k - $235.75k
...see how you can make a lasting impact on the world.Join our team of innovative engineers who are building an AI Data Center AIOps platform that turns raw, high-volume telemetry into reliable, job-centric insights and automation for GPU fleets. We’re hiring a DevOps Engineer...SeniorFull time$90k - $180k
...nutritionals and branded generic medicines. Our 115,000 colleagues serve people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.We...SeniorRemote work$160k - $240k
...millions of times a day - quickly, reliably, and securely. Any time you... ...at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our... ...operations or DevOps at a mid-to-senior level.Strong shell scripting...SeniorFull time$248k - $396.75k
...supportive environment, where NVIDIANs are inspired to excel and make a profound global impact.NVIDIA is seeking a Senior Manager of Site Reliability Engineering to lead and reshape how IT operations function at scale. This role goes beyond traditional service management...SeniorFull time$222k - $300.5k
...OverviewAbout the TeamIntuit's Infrastructure and Site Reliability organization owns the operational... .... The Fintech Platform Systems Engineering team builds and operates the AWS-based... ...negotiable.The OpportunityWe're hiring a Senior Manager, Site Reliability Engineering to...SeniorWorldwideShift work$146.7k - $339.3k
Immigration sponsorship is not available for this positionWhat you can expect As a Senior Lead Site Reliability Engineer, you can anticipate opportunities to work on our hybrid systems across the globe. You will be responsible for installing, configuring, and monitoring...SeniorFull timeWork at officeRemote workWorldwideShift workWeekend work$210.6k - $305.1k
...Minimum Qualifications: You have led a distributed team of 5+ engineers, can demonstrate strong technical vision for your team, and ensure... ..., and basic life insurance. Please see the Cisco careers site to discover more benefits and perks. Employees may be eligible...SeniorFull timeTemporary workLocal areaFlexible hours$145k - $165k
...A technology solutions firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key...Senior- ...Platform powers compute provisioning and infrastructure orchestration across our physical data centers. We are looking for a Senior Site Reliability Engineer to improve the reliability, scalability, and operational maturity of these systems as Lambda’s fleet and customer base...SeniorWork at officeLocal areaWork from homeFlexible hours
$168k - $270.25k
...deploy and run an AI data center. We take great pride in providing excellent, comprehensive support to our customers! Sr Site Reliability Engineer in this role will significantly impact and contribute to the overall success of both external customers running their clusters...SeniorFull timeWorldwide$187.04k - $359.72k
...systems by pushing for changes that improve reliability and velocity. Qualifications Minimum... ...degree in Computer Science, Electrical Engineering, Computer Engineering or related areas.... ...Product Ops, Corporate Functions and more. On-site presence across teams allows the company...SeniorTemporary workLocal areaOverseasShift work$174k - $252k
Senior Software Engineer, Site Reliability Engineering X Applicants in San Francisco: Qualified applications with arrest or conviction records will be considered for employment in accordance with the San Francisco Fair Chance Ordinance for Employers and the California...SeniorFull time- ...Palo Alto Networks in Santa Clara seeks a visionary Senior Principal Engineer/Architect to serve as the technical authority for global SRE and Platform Engineering initiatives. You will be the primary architect and driver of our AI-driven Autonomous SRE transformation...Senior
$125k - $160k
...Docker and Kubernetes.Collaborate with Product Management and engineering peers from concept through delivery.Maintain high engineering... ....Continuously improve system performance, scalability, and reliability.Qualifications:7+ years of professional experience with Java....SeniorFull timeRemote workFlexible hours$166k - $244k
A leading technology company located in Sunnyvale, California, is seeking a Site Reliability Engineer responsible for building and maintaining large-scale systems. The ideal candidate should possess a degree in Computer Science and have significant experience in programming...Senior$150k - $200k
Espace is seeking a Senior Software Engineer specializing in 5G Physical Layer to enhance IoT solutions from space. You will design and optimize algorithms for 5G systems while collaborating with teams across Europe, India, and the U.S. The role requires a strong background...SeniorFull time- ...Software Engineer We are looking for a few exceptional software engineers to work on our cloud based B2B e-commerce, renewals and subscriptions platform. As a member of the engineering team, you will work with product management and other team members to design...SeniorFlexible hours
- E-Space is seeking a Senior Software Engineer to join our Ground Software team, focusing on mission-critical Python-based back-end systems that operate a growing satellite constellation. You will design, implement, and scale real-time data ingestion pipelines, microservices...Senior
$70k - $200k
...to go, in any context, for generations to come. Reports To Senior Software Engineering Manager What You Will Be Doing As a Senior Software... ...driver base. What You Will Bring to ChargePoint Implement reliable APIs and microservices using Java and Spring Boot Contribute...Senior$130k - $200k
...Senior Software EngineerReady to make connectivity from space universally accessible, secure and actionable? Then you've come to the... ...intelligence.What is the role?E-Space is looking for a Senior Software Engineer to join our Ground Software team. You will collaborate with...SeniorImmediate start- Imperative Care in Campbell, California, is seeking a Staff Software Engineer in Robotics to design and implement critical software for their innovative robotic platform. This includes working on real-time algorithms and collaborating with cross-functional teams to meet...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- senior lead project manager Los Gatos, CA
- senior network engineer remote Los Gatos, CA
- senior manager accenture Los Gatos, CA
- senior brand designer Los Gatos, CA
- remote senior project manager Los Gatos, CA
- srs Los Gatos, CA
- senior manager pmo Los Gatos, CA
- senior financial analyst remote Los Gatos, CA
- senior manager product development Los Gatos, CA
- senior manager automotive Los Gatos, CA

