Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Software Engineer, Site Reliability Engineering

$174k - $252k

CompliancePoint

Site Reliability Engineering (SRE) Job

Site Reliability Engineering (SRE) is what you get when you treat operations as if it's a software problem. Our mission is to progress, protect, and provide for the software and systems behind all of Google's public services - Search, Ads, Gmail, Android, YouTube, and AppEngine, to name just a few - with an ever-watchful eye on their availability, latency, performance, and capacity. This is an unusual job, unlike others in the industry. Like traditional operations groups, we keep important, revenue-critical systems up and running despite hurricanes, bandwidth outages, and configuration problems. Unlike traditional operations groups, we also have full access to and authority to fix, extend, and scale the code to keep it working and harden it against all the vagaries of the Internet. We hire people from both systems and software backgrounds. Strong candidates will have experience with both. Just as what we do is unique, where we do it is unique too. At Google, we have the good fortune to have developed many interesting systems ranging from planet-spanning databases to near real-time scalable data warehousing to fault-tolerant datastream joining. In SRE, we flip between the fine-grained detail of disk driver I/O scheduling to the big picture of continental-level service capacity, across a range of systems and a user population measured in billions. We own those products in production. We drive reliability and performance across massive scale by mastering the full depth of the stack. We literally do learn something new every day - usually surprising things - that have the potential to transform the lives of billions of our users around the world.

Behind everything our users see online is the architecture built by the Technical Infrastructure team to keep it running. From developing and maintaining our data centers to building the next generation of Google platforms, we make Google's product portfolio possible. We're proud to be our engineers' engineers and love voiding warranties by taking things apart so we can rebuild them. We keep our networks up and running, ensuring our users have the best and fastest experience possible.

Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $174000 - $252000 (USD) + 15% bonus target + equity + benefits.

Minimum Qualifications:
  • Bachelor's degree in Computer Science, Engineering, a related field, or equivalent practical experience.
  • 5 years of experience with software development in one or more programming languages.
  • 3 years of experience in designing, analyzing, and troubleshooting large-scale distributed systems.
  • 2 years of experience leading projects and providing technical leadership.
Preferred Qualifications:
  • Master's degree in Computer Science or Engineering.
Responsibilities:
  • Engage in and improve the whole lifecycle of services—from inception and design, through to deployment, operation and refinement.
  • Support services before they go live through activities such as system design consulting, developing software platforms and frameworks, capacity planning and launch reviews.
  • Maintain services once they are live by measuring and monitoring availability, latency and overall system health.
  • Scale systems sustainably through mechanisms like automation, and evolve systems by pushing for changes that improve reliability and velocity.
  • Practice sustainable incident response and blameless postmortems.
Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Senior Software Engineer, Site Reliability Engineering in Mountain View, CA vacancy
  • $174k - $252k

     ...design consulting, developing software platforms and frameworks,...  ...pushing for changes that improve reliability and velocity.Practice...  ...degree in Computer Science, Engineering, a related field, or equivalent...  ...Computer Science or Engineering.Site Reliability Engineering (SRE)... 
    Senior

    Google

    Sunnyvale, CA
    3 days ago
  • $160k - $240k

     ...millions of times a day - quickly, reliably, and securely. Any time you...  ...at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our...  ...operations or DevOps at a mid-to-senior level.Strong shell scripting... 
    Senior
    Full time

    Fiserv

    Sunnyvale, CA
    2 days ago
  • $222k - $300.5k

     ...TeamIntuit's Infrastructure and Site Reliability organization owns the...  ...The Fintech Platform Systems Engineering team builds and operates the...  ...The OpportunityWe're hiring a Senior Manager, Site Reliability Engineering...  ...You'll partner closely with software engineering, product,... 
    Senior
    Worldwide
    Shift work

    Intuit

    Mountain View, CA
    5 days ago
  • $170k - $219k

     ...Site Reliability Engineer RADAR runs data infrastructure across 1,600+ live retail stores, processing tens of billions of real-world events every day. We're hiring a Site Reliability Engineer to own the reliability of that system end to end — leading incident response... 
    Senior
    Flexible hours
    Shift work
    Night shift

    Radar

    Sunnyvale, CA
    1 day ago
  •  ...the architecture and design of reliable, scalable, cost-effective,...  ...experience; at least eight years of software development experience; four...  ...in computer science or engineering is preferred. Key Skills...  ...Learning, Artificial Intelligence, Site Reliability Engineering, AI... 
    Senior

    Jobleads-US

    Sunnyvale, CA
    2 days ago
  •  ...home day is currently Tuesday.Engineering at Lambda is responsible for...  ...plane services and dataplane software running on SmartNICsDevelop tooling...  ...teams to improve service reliability and deployment...  ...rotationYouHave 5+ years of experience in Site Reliability Engineering, Production... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    4 days ago
  • $168k - $270.25k

     ...phenomenal people like you to help us accelerate the next wave of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing, implementing, and optimizing on-prem High-Performance... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  •  ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering...  ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    1 day ago
  • $148k - $235.75k

     ...on the world.Join our team of innovative engineers who are building an AI Data Center AIOps...  ...turns raw, high-volume telemetry into reliable, job-centric insights and automation for...  ...that operators depend on. You’ll partner Software Engineering and Systems Engineering team... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $267k - $356k

     ...currently Tuesday.Lambda's Storage Engineering team is the backbone behind...  ...the industry, which means reliability and performance aren't just...  ...behind Lambda's own software-defined data plane.Build and...  ...storage across new and existing sites using tools such as Ansible,... 
    Senior
    Work experience placement
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    5 days ago
  • $101k - $161k

     ...artificial intelligence, and software-defined networking to...  ...prestigious awards, such as Best Engineering Team, Best Company for Diversity...  ...Work WithWe’re looking for Site Reliability Engineers to join our...  ...EngineeringExperience level: Mid-Senior LevelIndustry: Computer Networking
    Senior

    Arista Networks

    Santa Clara, CA
    5 days ago
  • $104.9k - $174.7k

     ...responsible for improving the reliability, availability, performance,...  ...through completion.Follow up with engineering, development, security,...  ...Qualifications5+ years of experience in Site Reliability Engineering,...  ..., and support hardware, software, storage, network, cloud, Kubernetes... 
    Senior
    Full time
    Local area

    LexisNexis Risk Solutions Group

    San Jose, CA
    5 days ago
  • $152k - $241.5k

     ...artificial intelligence.We’re looking for a Senior SRE to join our Compute Farm team and...  ...host lifecycle management, fleet reliability/auto-healing, E2E observability or data-...  ...Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through... 
    Senior
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $165k - $280k

     ...with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building...  ...users to connect within minutes of unboxing, and the software that brings it all together. We’ve only begun to scratch the... 
    Senior
    Permanent employment
    Temporary work
    Worldwide
    Weekend work

    SpaceX

    Palo Alto, CA
    3 days ago
  •  ...Job Title : Senior Site Reliability Engineer Location : Santa Clara, CA Contract ENGAGEMENT SUMMARY The Candidate will provide SRE services for AI platforms and supporting infrastructure with emphasis on reliability engineering, incident response... 
    Senior
    Contract work

    VDart

    Santa Clara, CA
    5 days ago
  •  ...keep the world running. Location: 5 on-site days a week in Sunnyvale, CA Headquarters. Our Team's Vision: Our Engineering team is shaping the future of...  ...: We are looking for an experienced Senior Site Reliability Engineer (SRE) with a strong background... 
    Senior
    Work experience placement

    Illumio

    Sunnyvale, CA
    3 days ago
  • $192.4k - $275.8k

     ...the world's most demanding enterprise customers, blending Site Reliability Engineering, Systems Engineering, and Service Engineering disciplines...  ...this is the team for you Your ImpactYou will be the most senior technical individual contributor on the team — setting the... 
    Senior
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    3 days ago
  • $165k - $265k

     ...goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD) At SpaceX we’re...  ...users to connect within minutes, and the software that brings it all together. We’ve only...  ...availabilityMentor and train junior engineersAs a senior engineer you must lead the team to... 
    Senior
    Permanent employment
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Palo Alto, CA
    2 days ago
  • $132.6k - $214.5k

     ...you will collaborate closely with our engineering teams to develop innovative solutions that...  ...’ performance and health. As a Senior Staff SRE with the Cortex Observability...  ...operability of the product and ensure the reliability and availability of our services.... 
    Senior
    Full time
    Work at office
    Visa sponsorship
    Work visa

    Palo Alto Networks

    Santa Clara, CA
    1 day ago
  • $150.4k - $277.6k

     ...Cupertino, California, United States Software and Services The Media Platforms SRE team under the Apple Service Engineering division is one of the most exciting examples...  ...years experience At least 6 years in a Reliability Engineering, DevOps or infrastructure focused... 
    Senior
    Relocation
    Day shift

    Apple

    Cupertino, CA
    2 days ago
  • $180k - $230k

     ...Power Acceleration Job Description We're looking for a Senior SRE to own the reliability, scalability, and observability of our production systems. You'll work closely with platform and data engineering to keep high-throughput, data-intensive services running at... 
    Senior
    Work at office
    Local area
    Immediate start
    Remote work
    3 days per week

    GridCARE

    Redwood City, CA
    5 days ago
  •  ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems...  ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes... 
    Senior

    Saviynt

    Milpitas, CA
    a month ago
  • $255.7k - $300k

    Design, develop, test, and deploy scalable software solutions that maintain and enhance...  ...providing feedback to ensure best practices in reliability, security, and efficiency.Triage and...  ...development initiatives.Mentor other engineers and contribute to the engineering... 
    Full time

    Google

    Sunnyvale, CA
    4 days ago
  • $170k - $200k

    We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining, and troubleshooting cloud service/cluster, infrastructure, and monitoring systems to ensure high availability, performance,... 
    Full time
    Worldwide

    Fortinet

    Sunnyvale, CA
    4 days ago
  • $207.4k - $259.2k

     ...are seeking a highly experienced and passionate Sr. Staff Site Reliability Engineer (SRE) to join our growing team. In this critical role, you...  ...development teams to ensure reliability is built into the software development lifecycle from inception.Troubleshoot complex... 
    Senior
    Permanent employment
    Local area
    Worldwide
    Visa sponsorship

    Archer Aviation

    San Jose, CA
    5 days ago
  •  ...Investigate and resolve performance and reliability issues across application, infrastructure, database, Kubernetes, and Linux layers...  ..., plan capacity, improve observability, and collaborate with engineering teams and business stakeholders. Requirements: Requires hands... 

    engineeringjobs.net, Inc.

    Sunnyvale, CA
    1 day ago
  • $104.4k - $171k

     ...The mission of the Cloud Intelligence Group SRE (Site Reliability Engineering) Team is to ensure the stability of production environments, enterprise-grade cloud data reliability, and service continuity for the Cloud Intelligence Group. Our greatest challenge lies in... 

    Alibaba Cloud

    Sunnyvale, CA
    2 days ago
  • $145k - $165k

     ...: Selflessly collaborate towards our shared purpose. About the role Bolt Graphics is seeking a highly experienced Site Reliability Engineer (SRE) to design, build, and operate highly reliable developer and production systems. This role is mission-critical to maintaining... 
    Work at office
    Immediate start

    Bolt Graphics, Inc.

    Sunnyvale, CA
    3 days ago
  • $100k - $200k

     ...OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about... 
    Full time

    OPPO

    Palo Alto, CA
    3 days ago
  •  ...Site Reliability Engineer Onsite- Bay Area, CA Skills Relevant Skills and Experience What You’ll Do (Day-to-Day) Own and manage our cloud infrastructure (GCP or AWS, on-prem). Build, maintain, and optimize Kubernetes clusters (including GPU-backed clusters... 

    Amiri Recruiting

    Mountain View, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Software Engineer, Site Reliability Engineering. Be the first to apply!