Senior Site Reliability Engineer
$117k - $209.33kAutodesk
Job Requisition ID #26WD99273Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products.As part of a new SRE team supporting Autodesk GovCloud, you will have a unique opportunity to help shape how Autodesk deploys, runs, and improves production services in restricted cloud environments. This is a foundational role where you will help establish the operating model, reliability practices, automation, and engineering standards needed to support critical customer-facing services.You will combine software engineering and production operations to deploy, run, monitor, improve, and automate Autodesk services in GovCloud. You will partner closely with product engineering, security, compliance, platform, and infrastructure teams to ensure services are reliable, scalable, secure, and ready for production.The ideal candidate has deep experience operating production systems at scale, an automation-first mindset, and the ability to improve reliability through engineering practices such as SLOs/SLIs, production readiness, incident management, observability, resilience testing, and toil reduction. Success in this role requires strong technical judgment, a customer-focused mindset, and a passion for using software engineering to solve operational problems at scale.In accordance with GovCloud Cloud Service Provider Security Requirements, this role must be performed by U.S. Citizens. Employment is contingent upon meeting all applicable government security and eligibility requirements, including necessary background investigations and government issued security clearances.ResponsibilitiesServe as a primary owner for the reliability, availability, performance, operability, and capacity of one or more production servicesDeploy, operate, maintain, and continuously improve production services running in Autodesk GovCloud environmentsPartner with engineering teams to ensure services are designed with reliability, scalability, security, and operability in mindDefine and operate reliability practices such as SLOs/SLIs, error budgets, production readiness reviews, service reviews, and operational health reviewsBuild automation to improve deployment safety, operational efficiency, incident response, and service recoveryDesign, develop, and maintain software, automation, and tooling that improve the reliability, scalability, and efficiency of production systemsImplement and improve monitoring, alerting, logging, tracing, and observability capabilities across supported servicesLead and participate in incident response, troubleshooting, and post-incident reviews focused on learning and continuous improvementDevelop and maintain operational documentation, runbooks, and recovery proceduresScale and enhance resilience testing and Gameday practices to validate system behavior, recovery capabilities, and operational readinessContinuously identify and eliminate operational toil through software engineering, automation, and process improvementEnsure supported services remain compliant with Autodesk security, privacy, and regulatory requirements, including FedRAMP and related controls where applicableParticipate in a 24x7 on-call rotation for production servicesFunction effectively in a fast-paced environment while helping establish and mature operational excellence practices for Autodesk GovCloudMinimum QualificationsB.S. or higher in Computer Science, Engineering, or a related technical discipline, or equivalent practical experience7+ years of experience in Site Reliability Engineering, Software Engineering, Platform Engineering, Cloud Infrastructure, or Production OperationsExperience operating and supporting customer-facing production services in large-scale cloud environmentsStrong understanding of reliability engineering principles, including SLOs/SLIs, observability, incident management, capacity planning, production readiness, and automationExperience with AWS, Azure, or other public cloud platformsExperience developing automation using languages such as Python, Go, Java, PowerShell, Bash, or similarExperience with Infrastructure as Code, CI/CD pipelines, deployment automation, and modern cloud operations practicesUnderstanding of security, compliance, and operational risk management in production environmentsStrong written and verbal communication skillsPreferred Qualifications10+ years of experience operating highly available, customer-facing production systemsExperience with AWS GovCloud, FedRAMP, IL4/IL5, or other regulated cloud environmentsExperience supporting services with stringent availability, reliability, and security requirementsExperience with containers, Kubernetes, cloud-native architectures, APIs, load balancing, networking, DNS, and distributed systemsExperience with observability platforms such as Splunk, Dynatrace, Datadog, CloudWatch, or similar technologiesExperience operating databases, storage platforms, messaging systems, caching technologiesExperience designing and implementing operational automation at scaleExperience leading or participating in Gamedays, disaster recovery exercises, resilience testing, or operational readiness reviewsStrong incident management experience, including technical leadership during major incidents and stakeholder communicationStrong collaboration skills and ability to work effectively across engineering, security, compliance, and operations teamsPassion for building reliable, secure, and scalable systems that customers can trustLearn MoreAbout AutodeskWelcome to Autodesk! Amazing things are created every day with our software – from the greenest buildings and cleanest cars to the smartest factories and biggest hit movies. We help innovators turn their ideas into reality, transforming not only how things are made, but what can be made.We take great pride in our culture here at Autodesk – it’s at the core of everything we do. Our culture guides the way we work and treat each other, informs how we connect with customers and partners, and defines how we show up in the world.When you’re an Autodesker, you can do meaningful work that helps build a better world designed and made for all. Ready to shape the world and your future? Join us!BenefitsFrom health and financial benefits to time away and everyday wellness, we give Autodeskers the best, so they can do their best work. Learn more about our benefits in the U.S. by visiting Salary transparencySalary is one part of Autodesk’s competitive compensation package. For U.S.-based roles, we expect a starting base salary between $117,000 and $209,330. Offers are based on the candidate’s experience and geographic location, and may exceed this range. In addition to base salaries, our compensation package may include annual cash bonuses, commissions for sales roles, stock grants, and a comprehensive benefits package.Equal Employment OpportunityAt Autodesk, we're building a diverse workplace and an inclusive culture to give more people the chance to imagine, design, and make a better world. Autodesk is proud to be an equal opportunity employer and considers all qualified applicants for employment without regard to race, color, religion, age, sex, sexual orientation, gender, gender identity, national origin, disability, veteran status or any other legally protected characteristic. We also consider for employment all qualified applicants regardless of criminal histories, consistent with applicable law.BelongingWe take pride in cultivating a culture of belonging where everyone can thrive. Learn more here: In-Person Onboarding and Identity VerificationThis role may require in-person onboarding and/or in-person ID verification.Are you an existing contractor or consultant with Autodesk? Please search for open jobs and apply internally (not on this external site).SummaryLocation: San Francisco, CA, USAType: Full time
- About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure...Senior
$190.8k - $267.1k
...helping Reddit grow its business. The reliability of our Ads systems directly impacts advertiser... ...team partners closely with Ads Engineering to improve reliability, scalability, operational... ...advertiser trust. We’re looking for a Senior Site Reliability Engineer to build, operate,...SeniorFor contractorsWork experience placement- ...Apple Service Engineering (ASE) seeks a senior SRE software engineer to own the architectural direction of Kubernetes internals powering Apple services... ...will define controllers and namespace management, raise reliability, and contribute to upstream Kubernetes. The role...Senior
$127k - $249k
The TeamPlatform Engineering sits within SRE and builds the core infrastructure powering MongoDB... ...a pivotal role in engineering the reliable, globally connected, multi-cloud... ...Role OverviewWe are seeking a talented Senior Site Reliability Engineer (SRE) with a strong...SeniorLocal areaRemote workWorldwideFlexible hours$152.5k - $205k
...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible forThe Site Reliability Engineer builds and maintains shared platform capabilities, common libraries, and infrastructure that help Circle teams ship secure...SeniorFlexible hours$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours- ...A tech startup in San Francisco is looking for Site Reliability Engineers to enhance system reliability and performance. Ideal candidates have over 5 years of relevant experience and strong expertise in cloud infrastructure, including AWS and Kubernetes. The role involves...Senior
$148.5k - $223.9k
...right place! Agentforce is the future of AI, and you are the future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts in the Infrastructure and R&D organizations...SeniorFull timeWorldwideWeekend work$165k - $227k
...opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk.The Engineering OpportunityWe are looking for an experienced Senior Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly reliable...SeniorLocal areaWorldwideFlexible hours- ...’s build what’s next.About the teamThe Engineering team at Airwallex is a diverse group of... ...ownership, working together to build scalable, reliable, and secure products that empower... ...our Global services.What you’ll doAs a Senior Site Reliability Engineer, you’ll work...SeniorTemporary workLocal areaWorldwide
- ...getting here.)About the RoleWe're building infrastructure that has to perform under real-world scale, reliability, and security demands — and we're looking for an engineer who wants to own the foundation it runs on. This isn't a traditional "keep the lights on" role.You'...Senior
$200.7k - $250.9k
...washed away in a flood in 1942, the Royal Engineers rebuilt it. Then it washed away again in... ...opportunities for improvements in reliability/observability/performance/preparedness and... ...candidate for the role: Has past Site Reliability Engineering or DevOps experience...Senior$200k - $240k
...systems across all product teams. You will collaborate closely with engineering leadership, product managers, and cross-functional teams to... ...and Helm ~ Understand the importance of performant and reliable systems ~ Education - Ideally looking for a B.A. / B.S. degree...SeniorWork at officeImmediate start3 days per week$167.7k - $245.2k
...assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Your Impact As a Senior Site Reliability Engineer (SRE), you will lead the design and management of large-scale, highly available distributed systems, collaborating...SeniorFull timeTemporary workWork experience placementWork at officeLocal areaFlexible hours$167.7k - $245.2k
...within Cisco’s Networking, Security, Collaboration, and Observability portfolios.Your ImpactWe are seeking a skilled Senior Site Reliability Engineer (SRE) in Production Engineering with a strong background in SaaS and operations. You will design and manage large-scale...SeniorFull timeTemporary workWork at officeLocal areaFlexible hours1 day per week$232k - $319k
...to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and... ...enabled with self-serviceAccelerate the velocity of SRE and product engineering by developing robust platforms, powerful tooling, and...SeniorPermanent employmentLocal areaWorldwideFlexible hours$250k
...across Europe, while now significantly expanding its footprint in the United States. The company is looking for a Senior / Staff Site Reliability Engineer to support and scale large-scale HPC and cloud environments powering GPU-intensive workloads. The role involves...SeniorFull timeRemote work$175k - $250k
...00.00/yr - $250,000.00/yr Job Title: Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On-Site only. Must live within commuting distance... ...scalability, performance, and reliability across environments. What You’ll Do...SeniorFull timeRemote workRelocationRelocation package$15k
...benefits packages, technology talks by our experts, a beautiful modern office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to meet our growing needs, and you will leverage...SeniorWork at officeLocal areaRemote work$139.76k - $287.75k
...to grow their business.We are seeking a Senior Site ReliabilityEngineer to help operate,... ...will be instrumental in advancing the reliability, scalability, automation, observability... ...The ideal candidate is a highly hands-on engineer with strong production experience and a...SeniorWork at officeLocal areaRelocationRelocation package$106k - $130k
..., for any employer, at the date of hire. This position is ineligible for employment Visa sponsorship.Role Summary The Senior Site Reliability Engineer applies software engineering and systems engineering practices to improve the reliability, resilience, scalability, and...SeniorHourly payFull timeImmediate startVisa sponsorshipWork visaFlexible hours$120k - $175k
...level of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's possible...SeniorFull timeRemote workWork visaFlexible hours$262k - $364k
Lead a team of software/systems engineers on projects for users and be directly responsible for uptime.Own end-to-end availability... ...Engineering.Experience with machine learning infrastructure.Site Reliability Engineering (SRE) combines software and systems engineering...Senior$232.34k - $290.42k
...same: to make access to data as simple and reliable as electricity. With Fivetran, customer... ..., canonical and ready to query, with no engineering or maintenance required. We’re proud... ...integrate our teams, systems, and career sites. About the Role Fivetran and dbt...SeniorFull timeWork at officeRemote work$300k
...thousands of H100s, H200s, and B200s, ready for experimentation, full-scale model training, or inference. As a Platform Engineer/Senior Site Reliability Engineer, you’ll own the reliability, performance, and automation of this GPU-powered infrastructure, ensuring...SeniorPermanent employment- ...Infrastructure team builds the platforms and tooling that help engineering teams develop, deploy, and operate production systems safely... ...safe shipping the default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll define and scale...Permanent employmentWork experience placementWork at officeLocal area
$113.4k - $162k
...break down barriers to communication and free the flow of conversation for people everywhere.TextNow is looking for motivated Site Reliability Engineer to own infrastructure, monitoring, logging, ci/cd, reliability and everything in between!This role is about impact at...Temporary work- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Enterprise Technology, Infrastructure Platforms team, you will solve complex and broad business...
- ...The Team Platform Engineering is the department within SRE that is responsible for a range... ...role in developing and maintaining the reliable and globally connected multi-cloud network... ...Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong...SeniorFull timeWork at officeRemote workWorldwide
$194k - $267k
...all in on this mission. If you are too, let's talk.Position Overview:We are seeking a highly technical StaffObservabilitySite Reliability Engineer with a specialty in Splunk to own and evolve our Splunk ecosystem. In this role, you will move beyond simple monitoring to...Permanent employmentWork at officeLocal areaWorldwideFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer sre San Francisco, CA
- site reliability engineer San Francisco, CA
- site reliability engineer remote San Francisco, CA
- senior operations technician San Francisco, CA
- senior operations associate San Francisco, CA
- senior cloud service delivery manager San Francisco, CA
- senior it service manager San Francisco, CA
- senior project engineer San Francisco, CA
- senior chief engineer San Francisco, CA
- sr operations manager San Francisco, CA


