Software Engineer, Site Reliability (SRE)
$230k - $390kSierra
About usAt Sierra, we’re creating a platform to help businesses build better, more human customer experiences with AI. We are primarily an in-person company based in San Francisco, with growing offices in Atlanta, New York, London, Paris, Madrid, Munich, Singapore, Tokyo, and Sydney.We are guided by a set of values that are at the core of our actions and define our culture: Trust, Customer Obsession, Craftsmanship, Intensity, and Family. These values are the foundation of our work, and we are committed to upholding them in everything we do.Our co-founders are Bret Taylor and Clay Bavor. Bret currently serves as Board Chair of OpenAI. Previously, he was co-CEO of Salesforce (which had acquired the company he founded, Quip) and CTO of Facebook. Bret was also one of Google's earliest product managers and co-creator of Google Maps. Before founding Sierra, Clay spent 18 years at Google, where he most recently led Google Labs. Earlier, he started and led Google’s AR/VR effort, Project Starline, and Google Lens. Before that, Clay led the product and design teams for Google Workspace. What you'll doAs a Software Engineer on our Site Reliability team at Sierra, you will be responsible for defining and building the foundation of reliability, observability, and scalability across Sierra’s AI-driven infrastructure. You’ll partner closely with our core engineering and product teams to ensure our systems are highly available, efficient, and built for growth.Own Sierra’s observability stack—monitoring, alerting, logging, and tracing—to give engineers clear visibility into system health and performance.Partner with product and platform engineers to design systems that are reliable and scalable from day one—not as an afterthought.Design and implement scalable, reliable, and secure cloud infrastructure (AWS) using Terraform and modern DevOps tooling.Improve the reliability and scalability of our LLM deployments, ensuring robust, performant, and cost-effective operation.Lead improvements to deployment pipelines, CI/CD tooling, and incident management processes to reduce downtime and response time.Define the foundation of SRE practices at Sierra, influencing culture, tooling, and best practices across the engineering org.What you'll bring5+ years of hands-on experience in Site Reliability or Infrastructure engineering roles for complex SaaS or cloud-based systems.Experience designing for availability, scalability, and reliability at both infrastructure and application layers.Deep experience with Terraform, AWS services, container orchestration, and cloud networking (including IAM and VPC architecture).Strong background in observability systems (e.g., Prometheus, Grafana, Datadog, or similar).Experience working with enterprise customers and familiarity with their compliance and networking needs along with integration patterns.Comfortable working in fast-moving environments and collaborating across product, ML, and core engineering teams.Degree in Computer Science or a related field, or equivalent professional experience.Even better...Experience with LLM infrastructure — optimizing inference performance, managing fine-tuned models, or large-scale model deployment.Past experience in an early-stage startup environment, especially defining SRE culture and tooling from scratch.Familiarity with incident management automation or self-healing infrastructure patterns.Our valuesTrust: We build trust with our customers with our accountability, empathy, quality, and responsiveness. We build trust in AI by making it more accessible, safe, and useful. We build trust with each other by showing up for each other professionally and personally, creating an environment that enables all of us to do our best work.Customer Obsession: We deeply understand our customers’ business goals and relentlessly focus on driving outcomes, not just technical milestones. Everyone at the company knows and spends time with our customers. When our customer is having an issue, we drop everything and fix it.Craftsmanship: We get the details right, from the words on the page to the system architecture. We have good taste. When we notice something isn’t right, we take the time to fix it. We are proud of the products we produce. We continuously self-reflect to continuously self-improve.Intensity: We know we don’t have the luxury of patience. We play to win. We care about our product being the best, and when it isn’t, we fix it. When we fail, we talk about it openly and without blame so we succeed the next time.Family: We know that balance and intensity are compatible, and we model it in our actions and processes. We are the best technology company for parents. We support and respect each other and celebrate each other’s personal and professional achievements.What we offerWe want our benefits to reflect our values and offer the following to full-time employees:Flexible (unlimited) paid time offMedical, dental, and vision benefits for you and your familyLife insurance and disability benefitsRetirement plan dependent on country of employmentParental leaveFertility and family building benefits through CarrotLunch, as well as delicious snacks and coffee to keep you energized Discretionary benefit stipend giving people the ability to spend where it matters mostFree alphorn lessonsThese benefits are further detailed in Sierra's policies, may vary by region, and are subject to change at any time, consistent with the terms of any applicable compensation or benefits plans. Eligible full-time employees can participate in Sierra's equity plans subject to the terms of the applicable plans and policies.Be you, with usWe're working to bring the transformative power of AI to every organization in the world. To do so, it is important to us that the diversity of our employees represents the diversity of our customers. We believe that our work and culture are better when we encourage, support, and respect different skills and experiences represented within our team. We encourage you to apply even if your experience doesn't precisely match the job description. We strive to evaluate all applicants consistently without regard to race, color, religion, gender, national origin, age, disability, veteran status, pregnancy, gender expression or identity, sexual orientation, citizenship, or any other legally protected class.Compensation Range: $230K - $390KLocationSan Francisco, CAEmployment TypeFull timeLocation TypeOn-siteDepartmentEngineeringPlatform EngineeringCompensation$230K – $390K • Offers Equity
- ...design teams for Google Workspace. What you'll do As a Software Engineer on our Site Reliability team at Sierra, you will be responsible for defining and... ...reduce downtime and response time. Define the foundation of SRE practices at Sierra, influencing culture, tooling, and...WebsiteFull timeFlexible hours
$170k - $250k
...Site Reliability Engineer (SRE) Location: San Francisco, CA / Palo Alto, CA Company Stage of Funding: Growth-Stage AI Infrastructure Company ($80M Raised) Office Type: Onsite (4 Days Per Week) Salary: $170,000-$250,000 + Competitive Equity Company Description...WebsiteWork at officeVisa sponsorshipFlexible hours- ...management-and change lives along the way. The Role As a Site Reliability Engineer (SRE) at Air Apps, you will be responsible for ensuring the... ...of our systems. You will work at the intersection of software development and operations, implementing automation,...WebsiteTemporary workWorldwide
- ...Site Reliability Engineer (SRE) FLUIX is building the AI operating system that plans, designs, and optimizes AI infrastructure. We are based... ...are not afraid to get our hands dirty with physical and software systems. We are eager to visit and work with clients and...WebsiteWork at officeWeekend work
$170k - $230k
...Site Reliability Engineer (SRE) Palo Alto / San Francisco Bay Area About Mithril Mithril is an AI infrastructure platform built to make GPU compute more accessible and affordable for the world's leading enterprises, AI startups, and the AI research community,...WebsiteWork at officeLocal area1 day per week$325k
...s mission is to create reliable, interpretable, and steerable... ...committed researchers, engineers, policy experts, and... ...-- critical for both site reliability and... ...for reliability-minded software engineers and SREs.... ...also: Have been an SRE, Production Engineer, or...WebsiteFull timeWork at officeVisa sponsorshipFlexible hours$350k
...the Tinker community. About the Role We're looking for a Site Reliability Engineer to drive the reliability of Tinker end-to-end. You'll work... ..., or site reliability engineering. Proficiency writing software to solve reliability problems, including building tooling...WebsiteFull timeVisa sponsorshipWork visaRelocation package$150k - $176k
...is a Y Combinator 2024 Breakthrough Company . As a Software Engineer II on the Site Reliability Engineering team within the Platform Engineering group... ...Contribute to architectural discussions within the SRE team and with cross-functional teams Influence cross...WebsiteFull timeWork at officeLocal areaRemote workRelocationFlexible hours3 days per week$180.5k - $236.91k
Hi, we're Oscar. We're hiring a Senior Software Engineer, Cloud Infrastructure / SRE to join our Engineering team.Oscar is the first health insurance... ...'s business and technical domains such as DevOps, site reliability, and cloud best practicesLead the planning, execution...WebsiteFull timeWork at officeRemote work$163.71k - $306k
...WHY WE'RE LOOKING FOR YOU: Good software has to run where customers need it. For... ...infrastructure, behind their own controls, with the reliability and operational clarity they would... ...TAMs to trust. Partner with product engineers on infrastructure requirements for new...Website$300k
...experimentation, full-scale model training, or inference. As a Platform Engineer/Senior Site Reliability Engineer, you’ll own the reliability, performance, and... .... Skills / Must Have: ~7+ years of experience in SRE, DevOps, or Infrastructure Engineering roles supporting...WebsitePermanent employment- About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the...Website
- ...OpportunityTo achieve our ambitious goals, we’re looking for an SRE to join our infrastructure team. This role will be responsible for building software to ensure the reliability of our back-end systems, working with engineers who develop them, and planning for our future growth....WebsiteWorldwideHome officeFlexible hours
$152.5k - $205k
...you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design... .... This role is for an experienced SRE or infrastructure engineer who enjoys... ...Infrastructure Engineering, or a closely related software engineering role supporting production...WebsiteFlexible hours$113.4k - $162k
...everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring,... ...-Team Engagement: Work closely with software engineers, DevOps, and product teams to... ...the design and implementation of new SRE best practices.You'll be a great fit if...WebsiteTemporary work$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure... ...components that ensure cluster reliability and security (e.g., CoreDNS, cert-... ...youHave 6+ years of experience in software development and operating...WebsiteWork at officeLocal areaRemote workWorldwideFlexible hours$140k - $205k
Senior Technology Site Reliability EngineerCooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development... ...Site Reliability Engineer (“SRE”) is responsible for ensuring... ...applications. The SRE blends software engineering with systems engineering...WebsiteFull timeTemporary workWork at officeFlexible hoursWeekend work$117k - $209.33k
...help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and... ...Autodesk GovCloud products.As part of a new SRE team supporting Autodesk GovCloud, you... ...-facing services.You will combine software engineering and production operations...WebsiteFull timeFor contractors$165k - $225.6k
...across functions to drive scale, reliability, and innovation through technology.The Senior Site Reliability Engineer OpportunityReporting to the Manager... ...& Advocacy: Partner with software engineering teams to champion DevOps and SRE best practices, deliver excellent...WebsitePermanent employmentLocal areaWorldwideFlexible hours- ...proprietary infrastructure and software, we empower over 250,00... ...next.About the teamThe Engineering team at Airwallex is a... ...to build scalable, reliable, and secure products... ...grow without borders.Our SRE team is breaking new... ...What you’ll doAs a Senior Site Reliability Engineer,...WebsiteTemporary workLocal areaWorldwide
$195k - $257.5k
...is a stakeholder.What you’ll be responsible for:As a Staff Site Reliability Engineer on Circle’s Platform team, you’ll design, build, and operate... ..., leveraging Kubernetes, Infrastructure as Code, and modern SRE practices to deliver resilient, secure, and highly available...WebsiteFlexible hours$165k - $241.4k
...portfolios Your ImpactThe FedRAMP SRE team is focused on our... ...We’re looking for talented engineers with a software or operations background,... ...development teams to ensure the reliability, performance and security... ...see the Cisco careers site to discover more benefits and...WebsiteFull timeTemporary workWork at officeLocal areaFlexible hours1 day per week$148.5k - $223.9k
...Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco.... ...protected. The ExperienceAs an SRE, you will be a technical leader of... ...that prevent them, applying software engineering principles to operations...WebsiteFull timeWorldwideWeekend work$160k - $200k
Job Purpose:BTIG seeks a DevOps/Site Reliability Engineer to join our technology team. This role is central... ...thrives at the intersection of software operations, infrastructure engineering... ...3-5+ years of experience in a DevOps, SRE, platform engineering, or production support...WebsiteFull time$194k - $267k
...too, let's talk.The TeamThe Site Reliability team is dedicated to architecting... ...that support Okta’s SRE ecosystem. In this development... ...maximize platform reliability and engineering velocity.The ideal candidate... ...debugging software using gdb, strace, ltrace, tcpdump...WebsiteLocal areaWorldwideFlexible hours$220k - $235k
...a strategic, high-output Staff/Senior Staff SRE to define the future of our cloud platform and champion engineering excellence across Ironclad. In this role, you... ...technical leadership and strategic direction for the Site Reliability Engineering team and our broader Cloud...WebsiteFull timeContract workWork at office$204k - $306k
...mission. If you are too, let's talk.Manager, Site Reliability EngineeringSan Francisco,... ...Francisco Office. The IDaaS Site Reliability Engineering GroupOkta authenticates, authorizes and... ...What you’ll be doing Managing a team of SRE’s supporting various workloads and teams...WebsitePermanent employmentWork at officeLocal areaWorldwideFlexible hours2 days per week- ...Apple Service Engineering (ASE) seeks a senior SRE software engineer to own the architectural direction of Kubernetes internals powering Apple services... ...will define controllers and namespace management, raise reliability, and contribute to upstream Kubernetes. The role...Website
$194k - $267k
...let's talk.We are seeking a highly technical ObservabilitySite Reliability Engineer with a specialty in Google Cloud, to own and expand our... ...comprehensive, scalable Observability Platform that enables our SRE teams and business partners. You will treat infrastructure as...WebsitePermanent employmentLocal areaWorldwideFlexible hours$227.2k - $324.5k
About the Role:Site Reliability Engineering (SRE) at Tubi is not a traditional operations team. We are a software engineering organization that applies a developer's mindset and toolkit to the challenges of building and running large-scale, distributed systems. Our mission...WebsiteFull timeContract workTemporary workLocal areaFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Software Engineer, Site Reliability (SRE). Be the first to apply!
- ngo software engineer San Francisco, CA
- software data engineer San Francisco, CA
- graduate software engineer San Francisco, CA
- software development engineer (robotics engineer) San Francisco, CA
- software system engineer San Francisco, CA
- software developer no experience San Francisco, CA
- graduate software developer San Francisco, CA
- software engineer - early career San Francisco, CA
- entry level software engineer remote San Francisco, CA
- software engineer intern San Francisco, CA


