Principal Site Reliability Engineer, Google Cloud
Saviynt
Job Description
Job Description
Saviynt's AI-powered identity platform manages and governs human and non-human access to all of an organization's applications, data, and business processes. Customers trust Saviynt to safeguard their digital assets, drive operational efficiency, and reduce compliance costs. Built for the AI age, Saviynt is today helping organizations safely accelerate their deployment and usage of AI. Saviynt is recognized as the leader in identity security, with solutions that protect and empower the world’s leading brands, Fortune 500 companies and government institutions. For more information, please visit
Why This Role Matters
Saviynt’s platform is mission-critical for our customers. As we scale globally, reliability, availability, and performance are not optional—they are core product features.
As a Principal Engineer, you will define and drive the reliability strategy for our SaaS platform. This is a high-impact, hands-on engineering role with broad influence across infrastructure, platform, and application teams. You will shape how Saviynt designs, operates, and measures reliability at scale.
This role is ideal for engineers who want to work on hard reliability problems, influence architecture across teams, and leave a lasting mark on a growing SaaS platform.
What You Will Do
In this pivotal role, you will be instrumental in designing, building, and maintaining the shared infrastructure services and platforms that our product and application teams will depend on
• You will focus on creating reusable, reliable, and scalable solutions that abstract away complexity, enabling other teams to focus on their core business logic and deliver features faster in a multi-cloud environment
• Design and build core platform components and shared infrastructure services that other development teams will integrate with and leverage to deploy and operate their applications
• Architect, implement, and manage highly available and scalable Kubernetes platforms as a service for internal consumers
• Develop robust, internal-facing tools and automation for infrastructure provisioning and management primarily using Go (Golang)
• Architect and optimize foundational solutions within Cloud environments (AWS, Azure, etc.), focusing on creating reusable patterns and modules for other teams
• Design and implement shared Event-Driven Architecture components and messaging platforms using technologies like Kafka or Google Pub/Sub that product teams can easily utilize
• Develop and maintain robust CI/CD pipelines (e.g., GitLab CI and ArgoCD) as a service, providing standardized and automated deployment workflows for various development teams
• Design and build resilient Distributed Systems components that serve as building blocks for other applications, focusing on reliability, fault tolerance, and performance
• Manage and optimize our shared infrastructure across Multi-Region Cloud Environments, ensuring that platform services are globally available and performant for all consumers
• Establish and enhance centralized Observability and Monitoring platforms and tools that provide self-service insights for consuming teams
• Define and implement clear, well-documented RESTful API designs for the infrastructure services you build, ensuring ease of integration for internal clients
• Implement and manage Service Mesh (e.g., Envoy, Istio) capabilities, providing traffic management, security, and policy enforcement as a shared platform for services
• Design, implement, and optimize highly available Relational Database services or shared data platforms for broad organizational use
• Collaborate closely with product development teams to understand their infrastructure needs and pain points, providing technical guidance and support
• Participate in on-call rotations to support the critical shared infrastructure you build
What Are We Looking For• 1+ years of experience as a Principal SRE with a strong focus on building tools and services for other engineers
• Deep expertise with Kubernetes in production environments, particularly in providing it as a platform(i.e single tenant and multi-tenant deployment architectures)
• Strong programming skills in Go (Golang) and Python, with experience building robust, maintainable backend services and automation
• Extensive hands-on experience with at least one major Cloud Provider (GCP is a must); multi-cloud experience is a strong plus, especially in building abstractions over them.
• Proven experience designing and implementing Event-Driven Architecture and message queuing systems (e.g., Kafka, RMQ, NATS) as shared services
• Solid understanding and practical experience with CI/CD pipeline tools (especially GitLab CI) and experience establishing automated delivery processes for other teams
• Demonstrable experience designing and operating Distributed Systems, with an understanding of patterns for creating reliable, shared components
• Familiarity with Multi-Region Cloud Environments and strategies for building globally distributed and highly available platform
• Proficiency in establishing and utilizing comprehensive Observability and Monitoring platforms (e.g., Prometheus, Grafana, ELK stack, Datadog) for shared infrastructure
• Strong experience with RESTful API design principles and building well-documented, consumable APIs
• Knowledge of Service Mesh concepts and practical experience with solutions like Istio in a platform context
• Hands-on experience with Relational Databases (e.g., MySQL, PostgresSQL), ideally in managing them as a service
• Excellent communication skills and the ability to clearly articulate complex technical concepts to both technical and non-technical audiences
• A strong customer-centric mindset, treating internal development teams as your primary customers
• Advanced Professional GCP Certification is required.
• Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience or equivalent military experience required
If required for this role, you will:
- Complete security & privacy literacy and awareness training during onboarding and annually thereafter
- Review (initially and annually thereafter), understand, and adhere to Information Security/Privacy Policies and Procedures such as (but not limited to):
> Data Classification, Retention & Handling Policy
> Incident Response Policy/Procedures
> Business Continuity/Disaster Recovery Policy/Procedures
> Mobile Device Policy
> Account Management Policy
> Access Control Policy
> Personnel Security Policy
> Privacy Policy
Saviynt is an amazing place to work. We are a high-growth, Platform as a Service company focused on Identity Authority to power and protect the world at work. You will experience tremendous growth and learning opportunities through challenging yet rewarding work which directly impacts our customers, all within a welcoming and positive work environment. If you're resilient and enjoy working in a dynamic environment you belong with us!
Saviynt is an equal opportunity employer and we welcome everyone to our team. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
$170k - $277k
...Job Summary As a Principal Software Engineer within the Engineering team, you will drive the technical... ...-to-end delivery of next-generation cloud security solutions. You will collaborate... .... Prior experience working with Google Cloud Platform (GCP) or Amazon Web Services...PrincipalGoogle- ...Principal Software Engineer Join a forward-thinking team at JPMorgan... ...shape the future of cloud platform engineering.... ...solutions that are secure, reliable, and scalable. This... ...built on top of Google ADK, Anthropic SDKs,... ...health care coverage, on-site health and wellness...PrincipalGoogleWork at officeShift work
$235k - $250k
...Solve complex reliability challenges at scale... ...architecture and engineering culture at a company... ...faster in a multi-cloud environment Design... ...like Kafka or Google Pub/Sub that product... ...BRING ~1+ years of Principal-level of... ...Platform Engineering, or Site Reliability Engineering...PrincipalGoogle- ...platform runs on complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in... ...leadership role. You will own reliability for major platform domains,... ...using technologies like Kafka or Google Pub/Sub that product teams can easily...Google
- ...information, please visit PRINCIPAL SOFTWARE ENGINEER Saviynt is an identity... ...Saviynt’s Enterprise Identity Cloud gives customers... ...platforms like AWS Bedrock, Google AgentSpace , Salesforce AgentForce... ..., tooling, and operational reliability. Collaborate with internal...PrincipalGoogle
- ...Google is seeking a Software Engineering Manager II in Site Reliability Engineering, based in Sunnyvale, California. This onsite role leads a team to ensure reliability and performance of critical systems, partnering with product and engineering teams to deliver scalable...Google
$174k - $253k
...Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical... ...Science or Engineering. ABOUT THE JOB: Site Reliability Engineering (SRE) is what you get when... ...the software and systems behind all of Google’s public services - Search, Ads, Gmail,...Google- ...execution, with experience spanning Deepmind, Google, Amazon, Miro, Elise AI, IBM and... ...SRE to establish, lead, and scale our Site Reliability Engineering function. This role combines strategic... ...role. Strong hands‑on expertise in cloud infrastructure (AWS or Azure preferred...GoogleShift work
$100k - $200k
...Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role,... ...ideal candidate is passionate about cloud infrastructure, automation, and... ...cloud platforms such as AWS, Azure, and Google Cloud, including services like databases...GoogleFull time$272k - $431.25k
NVIDIA is seeking a Sr. Principal Systems Software Engineer for the Apache Spark Acceleration group. GPU accelerated... ...-node GPU deployments will reduce cloud computing costs and lower latency... ...such as AWS EMR, Databricks, Google Dataproc, Oracle Cloud Data Flow, Bytedance...PrincipalGoogleWork experience placement$217.57k - $260k
...otherwise, all roles are on-site five days per week at one of... ...Overview The Staff Site Reliability Engineer, Infrastructure role is... ...least 5 years of experience in cloud platforms-GCP preferred, AWS... ...large-scale tech companies (Google, Meta, Amazon, etc.) or...GoogleFull timeTemporary workWork at officeRemote workFlexible hoursShift work$114k - $253k
...possible. Observability Lead - Cloud SRE & Network Reliability Date: Jul 21, 2026... ...94538 Worker Category: On-site Flex The group you’ll be a... ...a strong Site Reliability Engineering (SRE) and multi-cloud networking... ...Monitor, CloudWatch, and Google Cloud Operations into a...GoogleLocal areaRemote workFlexible hours2 days per week3 days per week1 day per week$349k - $431k
...trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo... ...role, you will report to our Director of Engineering who leads Systems Intelligence and Machine... .... We are seeking a deeply experienced Principal Software Engineer to provide the overarching...PrincipalGoogle$126k - $204.5k
...Job Summary Threat Prevention and Cloud Service Infrastructure team - We are... ...cybersecurity. Your Career As a Principal Software Engineer, you will play a key role in the... ...on cloud platforms such as Google Cloud Platform (GCP), Amazon Web Services...PrincipalGoogleFull timeTemporary workWork at office$60 - $65 per hour
...Principal AI Architect Pay Range: $60hr - $65hr The Principal AI Architect will lead... ...will possess deep expertise in AI/ML engineering, cloud platforms, and modern AI orchestration... ...AI frameworks such as CrewAI, AutoGen, Google ADK, and LangGraph. ~ Experience with...PrincipalGoogle$258k - $387k
...raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and... ...leading investors.About the RoleAs a Principal Software Engineer, you will help define and build the high-performance, highly reliable foundation of the Nuro Driver. This role...PrincipalGooglePart timeImmediate startFlexible hours$307k - $427k
...basic coding elements to interface effectively with AI research engineers. Active engagement with ongoing national security or... ...intelligence will be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional...PrincipalGoogle$165.8k - $307.9k
...throughout its development lifecycle. As a Principal Software Developer in Test, you will be... ...this role, you will represent quality engineering and verification on behalf of your team... ...Groovy, Bash, Jira, Github, Confluence, Google Suite, macOS/Linux, MS SQL, Vector HIL Solutions...PrincipalGoogleWork at officeLocal areaRelocation package$275.8k - $340.5k
...the productivity of ML engineers, and drive the... ...workloads and managing reliable ML inference pipelines... ...and inference across cloud and on-prem compute resources... ...Position Overview: The Principal AI/ML Engineer will... ...: Experience with Google Cloud Platform, Microsoft...PrincipalGoogleLocal areaRemote workWork from homeRelocationRelocation packageFlexible hours$220k - $225k
...platforms like AWS Bedrock, Google AgentSpace, Salesforce AgentForce... ...solutions, using cloud, SAAS and AI design patterns... ...will contribute to the team's engineering standards: API design conventions... ...years of Staff or Associate Principal level professional software engineering...PrincipalGoogleLocal area$272k - $431.25k
...Principal Systems Software Engineer at NVIDIA is an engineering discipline to design, build and maintain large... ...and deployment and open source cloud enabling technologies like Kubernetes... ...facing GPU cloud services run maximum reliability and uptime as promised to the users...Principal$230k - $260k
...Principal ML Engineer Palo Alto, CA About Typeface We help the world's biggest brands move... ...practices for ML system evaluation, safety, reliability, and performance at scale Identify... ...with veterans from Adobe, Microsoft, Google, and top AI companies. Backed by...PrincipalGoogleWork at officeImmediate startFlexible hours3 days per week$248k - $396.75k
...infrastructure both on-prem and cloud. Join us in this... ...a highly skilled Principal AI/ML Engineer to join our dynamic... ...scalable data models for sites, fabrics, roles,... ...by setting standards (reliability, security, testability... ...expertise across Google Cloud, Azure, and Oracle...PrincipalGoogle$120.5k - $243k
...DevOps Engineer This role has been designed as 'Hybrid' with an expectation that you will... ...Packard Enterprise is the global edge-to-cloud company advancing the way people live and... ...with Cloud infrastructure (AWS/Azure/Google Cloud Platform) a definite asset. Excellent...GoogleWork experience placementWork at officeLocal areaImmediate start2 days per week$207k - $301k
...threat detection and hunting, malware intelligence, or Cloud Security Posture Management (CSPM). About The Job... ...create and maintain the safest operating environment for Google's users and developers. Security Engineers work with network equipment and actively monitor our...PrincipalGoogleLocal area$272k - $431.25k
## Principal Platform Software Engineer - RASApplylocations: US, CA, Santa Clara: US, Remotetime type: Full timeposted on: Posted Yesterdayjob requisition id: JR2011909NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer...Principal$248k - $391k
...the world!NVIDIA is hiring a Principal Engineer to lead three interconnected... ...and validates content for reliable consumption by AI agents — with... ..., OneDrive, SharePoint, Google Drive, Glean).Design and implement... ...document stores, wikis, and cloud drives — into the AI...PrincipalGoogleFull timePart time$204k - $259k
...trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo... ...system debugging, system telemetry, and reliability. We work closely with the Hardware, Compute... ...: Work on a small team of Software Engineers to develop system software components...GoogleFull timeWork experience placementRemote work$388k
...and Innovation team develops the Netflix application that runs on the largest SmartTV+ providers worldwide on devices from Amazon, Google, Roku, LG, Samsung, and others. We succeed by unlocking new capabilities and continuous improvements measured by increased velocity...GoogleHourly payFull timeImmediate startWorldwideFlexible hours$145k - $182k
...deliver code to their users reliably, efficiently, securely... ..., feature flags and cloud costs. The Harness... ...Testing Orchestration, Chaos Engineering, Software Engineering... ...Ventures, GV (formerly Google Ventures), Alkeon Capital... ...code Work alongside Site Reliability Engineers...GoogleFull timeLocal areaImmediate startFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Site Reliability Engineer, Google Cloud. Be the first to apply!
- chief engineer Milpitas, CA
- engineering director Milpitas, CA
- data center chief engineer Milpitas, CA
- general engineer Milpitas, CA
- principal engineer Milpitas, CA
- hotel chief engineer Milpitas, CA
- principal developer Milpitas, CA
- cloud developer Milpitas, CA
- senior principal cloud computing engineer Milpitas, CA
- senior cloud network engineer Milpitas, CA





