Senior Software Engineer, Site Reliability Engineering
$174k - $252kEngage in and improve the whole lifecycle of services—from inception and design, through to deployment, operation and refinement.Support services before they go live through activities such as system design consulting, developing software platforms and frameworks, capacity planning and launch reviews.Maintain services once they are live by measuring and monitoring availability, latency and overall system health.Scale systems sustainably through mechanisms like automation, and evolve systems by pushing for changes that improve reliability and velocity.Practice sustainable incident response and blameless postmortems.Minimum qualifications:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical experience.5 years of experience with software development in one or more programming languages.3 years of experience in designing, analyzing, and troubleshooting large-scale distributed systems.2 years of experience leading projects and providing technical leadership. Preferred qualifications:Master's degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you get when you treat operations as if it’s a software problem. Our mission is to progress, protect, and provide for the software and systems behind all of Google’s public services - Search, Ads, Gmail, Android, YouTube, and AppEngine, to name just a few - with an ever-watchful eye on their availability, latency, performance, and capacity. This is an unusual job, unlike others in the industry. Like traditional operations groups, we keep important, revenue-critical systems up and running despite hurricanes, bandwidth outages, and configuration problems. Unlike traditional operations groups, we also have full access to and authority to fix, extend, and scale the code to keep it working and harden it against all the vagaries of the Internet. We hire people from both systems and software backgrounds. Strong candidates will have experience with both. Just as what we do is unique, where we do it is unique too. At Google, we have the good fortune to have developed many interesting systems ranging from planet-spanning databases to near real-time scalable data warehousing to fault-tolerant datastream joining. In SRE, we flip between the fine-grained detail of disk driver I/O scheduling to the big picture of continental-level service capacity, across a range of systems and a user population measured in billions. We own those products in production. We drive reliability and performance across massive scale by mastering the full depth of the stack. We literally do learn something new every day - usually surprising things - that have the potential to transform the lives of billions of our users around the world.Behind everything our users see online is the architecture built by the Technical Infrastructure team to keep it running. From developing and maintaining our data centers to building the next generation of Google platforms, we make Google's product portfolio possible. We're proud to be our engineers' engineers and love voiding warranties by taking things apart so we can rebuild them. We keep our networks up and running, ensuring our users have the best and fastest experience possible.Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $174000 - $252000 (USD) + 15% bonus target + equity + benefitsLearn more about benefits at Google.Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical experience.5 years of experience with software development in one or more programming languages.3 years of experience in designing, analyzing, and troubleshooting large-scale distributed systems.2 years of experience leading projects and providing technical leadership.
- ...Oracle Cloud Infrastructure (OCI) seeks a Senior Principal Engineer to lead the design and implementation of reliability validation for OCI control plane services, focusing on a high-performance, low-level systems approach. You will mentor engineers, define validation...Senior
$160k - $240k
...Senior Site Reliability Engineer Calling all innovators - find your future at Fiserv. We're Fiserv, a global leader in Fintech and payments, and we move money and information in a way that moves the world. We connect financial institutions, corporations, merchants...Senior$132.6k - $214.5k
...you will collaborate closely with our engineering teams to develop innovative solutions that... ...' performance and health. As a Senior Staff SRE with the Cortex Observability... ...operability of the product and ensure the reliability and availability of our services. Qualifications...SeniorFull timeWork at officeVisa sponsorshipWork visa- ...We are seeking a Senior Database Reliability Engineer (DBRE) to design, operate, and improve reliable, scalable, secure, and highly available database... ...data platforms. The role combines database engineering, site reliability engineering, Linux systems administration, and...Senior
- ...automated detection, drain/cordon/taint, workload rescheduling. Feed the AIOps substrate The remediation-actuator and workflow engine land here — you make the control plane safe for automated action. Your CRDs are the schema the platform's predictors and...SeniorLocal area
$165k - $280k
...the ultimate goal of enabling human life on Mars. SR. SITE RELIABILITY ENGINEER (STARLINK) At SpaceX we’re leveraging our experience in... ...allow users to connect within minutes of unboxing, and the software that brings it all together. We’ve only begun to scratch the...SeniorPermanent employmentTemporary workWorldwideWeekend work$128k - $216k
...of times a day - quickly, reliably, and securely. Any time you... ...difference at Fiserv. Sr. Site Reliability Engineer About Clover Clover... ...What Does A Successful Senior Site Reliability Engineer Do... ...and bridge the gap between software and infrastructure. What...SeniorWorldwide- ...the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines... ...the team for you. Your Impact You will be the most senior technical individual contributor on the team — setting the...Senior
- ...role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining... ...yourself among the top echelon in site reliability. As a Senior Lead Site Reliability Engineer at... ...auditability outcomes. Advanced knowledge of software applications and technical processes...Senior
- LeanData helps the world’s fastest-growing companies automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly to the SVP of Engineering, this role is...SeniorFull timeWork at office2 days per week
$168k - $270.25k
...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance...SeniorFull time$101k - $161k
...artificial intelligence, and software-defined networking to... ...prestigious awards, such as Best Engineering Team, Best Company for Diversity... ...Work WithWe’re looking for Site Reliability Engineers to join our... ...EngineeringExperience level: Mid-Senior LevelIndustry: Computer NetworkingSenior$148k - $235.75k
...on the world.Join our team of innovative engineers who are building an AI Data Center AIOps... ...turns raw, high-volume telemetry into reliable, job-centric insights and automation for... ...that operators depend on. You’ll partner Software Engineering and Systems Engineering team...SeniorFull time$152k - $241.5k
...artificial intelligence.We’re looking for a Senior SRE to join our Compute Farm team and... ...host lifecycle management, fleet reliability/auto-healing, E2E observability or data-... ...Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through...SeniorFull time$267k - $356k
...currently Tuesday.Lambda's Storage Engineering team is the backbone behind... ...the industry, which means reliability and performance aren't just... ...behind Lambda's own software-defined data plane.Build and... ...storage across new and existing sites using tools such as Ansible,...SeniorPart timeWork experience placementWork at officeLocal areaWork from homeFlexible hours$262k - $364k
...within the AViD ecosystem have reliability and uptime appropriate to... ...and performance.Build creative engineering solutions to operations and infrastructure... ....8 years of experience with software development in one or more... ...in a strategic way.Site Reliability Engineering (SRE)...Senior$222k - $300.5k
...TeamIntuit's Infrastructure and Site Reliability organization owns the... ...The Fintech Platform Systems Engineering team builds and operates the... ...The OpportunityWe're hiring a Senior Manager, Site Reliability Engineering... ...You'll partner closely with software engineering, product,...SeniorWorldwideShift work$262k - $364k
Lead a team of Software/Systems Engineers on projects for users and be directly responsible for uptime... ...or Engineering, or a related field.Site Reliability Engineering (SRE) combines software... ...Engineer chose to join SRE.As the Senior Engineering Manager for Collaboration...Senior$168k - $270.25k
...NVIDIA infrastructure. Work with NVIDIA's DGX Cloud team as a Senior Site Reliability Engineer to maintain high-performance DGX Cloud clusters for AI... ...launch through system creation consulting, developing software tools, platforms and frameworks, capacity management, and...SeniorFull timeRemote workWorldwide- ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems... ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes...Senior
- ...Google is seeking a Software Engineering Manager II in Site Reliability Engineering, based in Sunnyvale, California. This onsite role leads a team to ensure reliability and performance of critical systems, partnering with product and engineering teams to deliver scalable...
- ...California, invites an experienced CDN Solutions Engineer to join the Content Delivery Network... ...engineering groups across Apple to ensure reliable delivery at scale. The ideal candidate has 4+ years in CDNs and production software development (Python/Go/Node.js), plus...
- ...ServiceNow in Santa Clara, CA, seeks a Staff Software Engineer – SRE & AIOps to drive infrastructure automation, resilience, and toil... ...remediation for global engineering teams. Embedded within the Site Reliability & Database Engineering organization, you will architect...
$276.1k - $311.4k
...technology. Our advanced AI software and foundation models enable... ...charter, hiring its founding engineers, establishing the operating model... ...strategy that makes reliability a first-class property of the... ...and emerging leaders, and give senior leadership the clarity on reliability...Permanent employmentFull timeWork at officeWork from home$248k - $396.75k
...Santa Clara Full time JR2023973 Site Reliability Engineering (SRE) at NVIDIA is an engineering... ...resilience, and availability. It combines software and systems engineering practices... ...automated anomaly detection. Partner with senior leaders and engineers across Cloud,...Full time- ...services, powered by the Wafer-Scale Engine (WSE). This team will help deliver world-class, ultra-reliable inference infrastructure for... ...for delivering and running software reliably and at scale across... ...-prem environments. Mentor senior SREs, support critical incident...Shift work
$151.6k - $245.3k
...Summary Your Career Palo Alto Networks runs a large hybrid infrastructure and is one of the largest GCP customers. As a Site Reliability Engineer, you will be part of a team supporting the services running on this infrastructure. This includes automation, architecture...Full timeWork at office$110k - $130k
...the World's leading AI-first Quality Engineering Company? Ready to advance your career,... ...at QualityAI! We are looking for a Site Reliability Engineer to join our growing team in Riverwoods... ...Site Reliability Engineer (SRE). ~ Software development "hands on" engineer with...Casual workLocal areaFlexible hours$230k - $250k
...Site Reliability Engineer Forward was founded in 2013 by four Stanford Ph.D.s, building the industry's first network digital twin: a mathematically accurate model of the production network. It's the foundation for autonomous networking, giving engineers and AI agents...Night shift$65 - $85 per hour
...Site Reliability Engineer Sustainable Talent is partnering with a global leader who's been transforming computer graphics, PC gaming, and... ...This group works with various other groups within NVIDIA Software such as Graphics Processors, Mobile Processors, Deep Learning...Full timeContract workWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Software Engineer, Site Reliability Engineering. Be the first to apply!
- software developer positions Sunnyvale, CA
- senior software engineer remote Sunnyvale, CA
- software engineer contract Sunnyvale, CA
- software qa engineer Sunnyvale, CA
- cybersecurity software engineer Sunnyvale, CA
- part time software developer remote Sunnyvale, CA
- junior software developer internship Sunnyvale, CA
- software system engineer Sunnyvale, CA
- software engineer remote Sunnyvale, CA
- work from home software developer Sunnyvale, CA



