Principal Site Reliability Engineer, Google Cloud
Saviynt
Job Description
Job Description
Saviynt's AI-powered identity platform manages and governs human and non-human access to all of an organization's applications, data, and business processes. Customers trust Saviynt to safeguard their digital assets, drive operational efficiency, and reduce compliance costs. Built for the AI age, Saviynt is today helping organizations safely accelerate their deployment and usage of AI. Saviynt is recognized as the leader in identity security, with solutions that protect and empower the world’s leading brands, Fortune 500 companies and government institutions. For more information, please visit
Why This Role Matters
Saviynt’s platform is mission-critical for our customers. As we scale globally, reliability, availability, and performance are not optional—they are core product features.
As a Principal Engineer, you will define and drive the reliability strategy for our SaaS platform. This is a high-impact, hands-on engineering role with broad influence across infrastructure, platform, and application teams. You will shape how Saviynt designs, operates, and measures reliability at scale.
This role is ideal for engineers who want to work on hard reliability problems, influence architecture across teams, and leave a lasting mark on a growing SaaS platform.
What You Will Do
In this pivotal role, you will be instrumental in designing, building, and maintaining the shared infrastructure services and platforms that our product and application teams will depend on
• You will focus on creating reusable, reliable, and scalable solutions that abstract away complexity, enabling other teams to focus on their core business logic and deliver features faster in a multi-cloud environment
• Design and build core platform components and shared infrastructure services that other development teams will integrate with and leverage to deploy and operate their applications
• Architect, implement, and manage highly available and scalable Kubernetes platforms as a service for internal consumers
• Develop robust, internal-facing tools and automation for infrastructure provisioning and management primarily using Go (Golang)
• Architect and optimize foundational solutions within Cloud environments (AWS, Azure, etc.), focusing on creating reusable patterns and modules for other teams
• Design and implement shared Event-Driven Architecture components and messaging platforms using technologies like Kafka or Google Pub/Sub that product teams can easily utilize
• Develop and maintain robust CI/CD pipelines (e.g., GitLab CI and ArgoCD) as a service, providing standardized and automated deployment workflows for various development teams
• Design and build resilient Distributed Systems components that serve as building blocks for other applications, focusing on reliability, fault tolerance, and performance
• Manage and optimize our shared infrastructure across Multi-Region Cloud Environments, ensuring that platform services are globally available and performant for all consumers
• Establish and enhance centralized Observability and Monitoring platforms and tools that provide self-service insights for consuming teams
• Define and implement clear, well-documented RESTful API designs for the infrastructure services you build, ensuring ease of integration for internal clients
• Implement and manage Service Mesh (e.g., Envoy, Istio) capabilities, providing traffic management, security, and policy enforcement as a shared platform for services
• Design, implement, and optimize highly available Relational Database services or shared data platforms for broad organizational use
• Collaborate closely with product development teams to understand their infrastructure needs and pain points, providing technical guidance and support
• Participate in on-call rotations to support the critical shared infrastructure you build
What Are We Looking For• 1+ years of experience as a Principal SRE with a strong focus on building tools and services for other engineers
• Deep expertise with Kubernetes in production environments, particularly in providing it as a platform(i.e single tenant and multi-tenant deployment architectures)
• Strong programming skills in Go (Golang) and Python, with experience building robust, maintainable backend services and automation
• Extensive hands-on experience with at least one major Cloud Provider (GCP is a must); multi-cloud experience is a strong plus, especially in building abstractions over them.
• Proven experience designing and implementing Event-Driven Architecture and message queuing systems (e.g., Kafka, RMQ, NATS) as shared services
• Solid understanding and practical experience with CI/CD pipeline tools (especially GitLab CI) and experience establishing automated delivery processes for other teams
• Demonstrable experience designing and operating Distributed Systems, with an understanding of patterns for creating reliable, shared components
• Familiarity with Multi-Region Cloud Environments and strategies for building globally distributed and highly available platform
• Proficiency in establishing and utilizing comprehensive Observability and Monitoring platforms (e.g., Prometheus, Grafana, ELK stack, Datadog) for shared infrastructure
• Strong experience with RESTful API design principles and building well-documented, consumable APIs
• Knowledge of Service Mesh concepts and practical experience with solutions like Istio in a platform context
• Hands-on experience with Relational Databases (e.g., MySQL, PostgresSQL), ideally in managing them as a service
• Excellent communication skills and the ability to clearly articulate complex technical concepts to both technical and non-technical audiences
• A strong customer-centric mindset, treating internal development teams as your primary customers
• Advanced Professional GCP Certification is required.
• Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience or equivalent military experience required
Saviynt is an amazing place to work. We are a high-growth, Platform as a Service company focused on Identity Authority to power and protect the world at work. You will experience tremendous growth and learning opportunities through challenging yet rewarding work which directly impacts our customers, all within a welcoming and positive work environment. If you're resilient and enjoy working in a dynamic environment you belong with us!
Security & Compliance
This role requires adherence to Saviynt’s information security and privacy policies and procedures, including annual security training.
Saviynt is an equal opportunity employer and we welcome everyone to our team. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.
- ...information, please visit PRINCIPAL SOFTWARE ENGINEER Saviynt is an identity... ...’s Enterprise Identity Cloud gives customers unparalleled... ...platforms like AWS Bedrock, Google AgentSpace , Salesforce AgentForce... ..., tooling, and operational reliability. Collaborate with...PrincipalGoogle
$235k - $250k
...Solve complex reliability challenges at scale... ...architecture and engineering culture at a company... ...faster in a multi-cloud environment Design... ...like Kafka or Google Pub/Sub that product... ...BRING ~1+ years of Principal-level of... ...Platform Engineering, or Site Reliability Engineering...PrincipalGoogle$195k - $232k
...platforms like AWS Bedrock, Google AgentSpace, Salesforce AgentForce... ...solutions, using cloud, SAAS and AI design patterns... ...will contribute to the team's engineering standards: API design conventions... ...years of Staff or Associate Principal level professional software engineering...PrincipalGoogleLocal area- ...Google is seeking a Software Engineering Manager II in Site Reliability Engineering, based in Sunnyvale, California. This onsite role leads a team to ensure reliability and performance of critical systems, partnering with product and engineering teams to deliver scalable...Google
- ...software development, cloud infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated... ...for accuracy and reliability. Work with AWS, Azure... ...across tools such as Slack, Google Drive, and Microsoft 36... ...Infrastructure Site Reliability Engineering...GoogleFor contractorsRemote work
$247.5k - $267k
...Principal Software Engineer Palo Alto, CA About Typeface We help the world'... ...grade standards for security, reliability, and compliance.... .... ~ Strong background in cloud-native architecture (AWS, GCP... ...veterans from Adobe, Microsoft, Google, and top AI companies. Backed...PrincipalGoogleWork at officeFlexible hours3 days per week$183.6k - $297k
...Job Summary The Team Engineering - Our engineering team is at... ...environment. Job Summary As a Principal Software Engineer within the... ...delivery of next-generation cloud security solutions. You will... ...experience working with Google Cloud Platform (GCP) or Amazon...PrincipalGoogleFull timeWork at office$307k - $427k
...technical decisions align with Google-level OKRs and long-term... ...between capacity infrastructure, engineering organizations, and strategic... ...spread across all continents. As Principal Engineer, you will own the... ....As the Principal Engineer, Cloud Capacity Planning, you will bridge...PrincipalGoogle$272k - $431.25k
NVIDIA is seeking a Sr. Principal Systems Software Engineer for the Apache Spark Acceleration group. GPU accelerated... ...-node GPU deployments will reduce cloud computing costs and lower latency... ...such as AWS EMR, Databricks, Google Dataproc, Oracle Cloud Data Flow, Bytedance...PrincipalGoogleWork experience placement$100k - $200k
...Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role,... ...ideal candidate is passionate about cloud infrastructure, automation, and... ...cloud platforms such as AWS, Azure, and Google Cloud, including services like databases...GoogleFull time$165.8k - $307.9k
...throughout its development lifecycle. As a Principal Software Developer in Test, you will be... ...this role, you will represent quality engineering and verification on behalf of your team... ...Groovy, Bash, Jira, Github, Confluence, Google Suite, macOS/Linux, MS SQL, Vector HIL...PrincipalGoogleWork at officeLocal areaRelocation package$272k - $431.25k
NVIDIA is looking for a Cloud Site Reliability Engineering Architect to work in IPP's (Infrastructure, Planning and Process) Cloud Infrastructure Team. IPP is a global organization within NVIDIA. This group works with various other groups within NVIDIA such as Graphics...PrincipalFull timeWork experience placementWorldwide$280k - $350k
...top researchers and engineers, building the... ...team as a Staff / Principal Platform Engineer... ...force behind our cloud infrastructure, partnering... ..., and maintain reliable, high-performance,... ...major cloud provider (Google Cloud Platform,... ...be working on-site in our South Bay office...PrincipalGoogleFull timeWork at officeRelocation$272k - $431.25k
...accelerate business results across engineering, IT, supply chain, finance, HR, and sales. We need a Principal or Distinguished Engineer-... ...quality, latency, cost, reliability, and safety.Comprehensive knowledge... ..., Microsoft Copilot Studio, Google Agentspace, or similar...PrincipalGoogleFull time$240k - $250k
...Optimize agent performance, scalability, reliability, and resource utilization. Lead architecture and mentor engineers building endpoint platform capabilities.... ...operations. WHAT YOU BRING ~1+ years of Principal-level of systems software engineering experience...PrincipalLocal area$183.6k - $297k
...outcomes.Job SummaryYour Career Our ATP Cloud team is at the forefront of... ...and resolve incidents at scale. As a Principal Software Engineer, you will own the technical vision for... ...native AI tools (Vertex AI Agent Builder, Google ADK) and define how our team builds and...PrincipalGoogleFull timeWork at office- ...help shape the future of cloud platform engineering. As a Principal Software Engineer, you'll... ...solutions that are secure, reliable, and scalable. This is... ...frameworks built on top of Google ADK, Anthropic SDKs, etc.... ...health care coverage, on-site health and wellness...PrincipalGoogleWork at officeShift work
$258k - $387k
...raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and... ...investors. About the Role As a Principal Software Engineer, you will help define and build the high-performance, highly reliable foundation of the Nuro Driver. This role...PrincipalGoogleImmediate startFlexible hours- ...Job Title - Senior Google Cloud Architect Google Cloud Platform (GCP) & AI Services... ...optimization, operational excellence, and reliability initiatives. Leadership & Stakeholder... ...leadership teams. Mentor architects, engineers, and cloud practitioners. Support...GoogleTemporary work
$240k - $260k
...You define the standards ML engineers and scientists build on, and... ...database: operate Pgvector (Cloud SQL) for POC and Qdrant on GKE... ...~1+ years of experience as a Principal SWE at a SaaS company ~ Demonstrated... ...Solve challenging cloud and reliability problems at scale...Principal$143k - $286k
...Summary... What you'll do... Principal, Software Engineer We are seeking a talented... ...design, develop, and enhance a reliable and easy-to-maintain backend... ...skills.?? Familiarity with public cloud technologies such as Azure or Google Cloud Platform.?? Extensive...PrincipalGoogleFull timeTemporary workPart timeWork at officeFlexible hours- ...Principal GCP Data Architect VLink, founded in 2006, is a leading... ...global provider of software engineering services with next-gen technologies... ...who has led massive cloud migrations, built self-serve... ...including BigLake and Omni), Google Cloud Storage. Processing : Dataflow...PrincipalGoogle
$180k - $260k
...California, United StatesProducts - Engineering /Fulltime /HybridOver 50,000... ...trust our end-to-end, cloud-driven networking solutions.... ...passionate about building highly reliable, scalable and secure... ...service provider platforms (AWS, Google Cloud, and Azure) to enhance...PrincipalGoogleFull timeWork experience placementWork at officeLocal area$176.8k - $221k
...rapidly growing Fintech CompanyAs a Principal Partner Solution Engineer for our Partner Embed Platform... ...of web services, cloud-based platforms (AWS, Azure, or Google Cloud), and middleware technologies... ...benefits, and teams on our career site, LinkedIn Life, or YouTube...PrincipalGoogleTemporary workWork at officeRemote workFlexible hours$207k - $300k
Lead a team of Software/Systems Engineers on projects for users and be directly responsible... ...scalability, latency and efficiency of Google's services.Minimum qualifications:... ...Engineering.Experience with Large Language Model.Site Reliability Engineering (SRE) combines software and...Google$240k - $250k
...microservices in Go (Golang) that power the Cloud Control Plane and distributed Edge... ...modern cloud-native technologies. Mentor engineers, lead architecture discussions, and... ...operations. WHAT YOU BRING ~1+ years of Principal-level backend software engineering...PrincipalLocal area$249k
...build for travelers everywhere. Principal Data & AI Engineer, Reporting and Insights... ...Anthropic, or similar solutions, ensuring reliability, security, and cost efficiency in production... ..., such as Azure OpenAI, AWS Bedrock, Google Vertex AI, or an equivalent platform....PrincipalGoogleWork at office$200k - $250k
...PayPal, Venmo, Cash App Pay, Apple Pay and Google Pay, and even cash at more than 62,000... ...a time. Responsibilities: As Principal Product Manager, Data & Insights, you will... ...of the technical landscape across data engineering, analytics, machine learning, and software...PrincipalGoogleWork from homeFlexible hours$150k - $217k
...qualifications Advanced degree in a management, technical, or engineering field. 5 years of experience in a management consulting firm with... ...group working on a range of critical projects and issues. Google's leadership team hand-picks thorny business challenges, and members...PrincipalGoogle- ...looking for a Senior Software Engineer to join our Platform team and... ...commands must execute reliably in real time, and ML models need... ...production infrastructure on a public cloud platform (we run on GCP, but... ...tooling (we use Grafana + Google Cloud Monitoring)...GoogleFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Site Reliability Engineer, Google Cloud. Be the first to apply!
- principal engineer Milpitas, CA
- senior civil engineer project manager Milpitas, CA
- chief engineer Milpitas, CA
- director software engineering Milpitas, CA
- engineering director Milpitas, CA
- data center chief engineer Milpitas, CA
- hotel chief engineer Milpitas, CA
- general engineer Milpitas, CA
- principal developer Milpitas, CA
- senior cloud solutions architect Milpitas, CA



