Senior Site Reliability Engineer
$117k - $209.33kAutodesk
Job Requisition ID #26WD99273Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable, secure, and scalable cloud services for Autodesk GovCloud products.As part of a new SRE team supporting Autodesk GovCloud, you will have a unique opportunity to help shape how Autodesk deploys, runs, and improves production services in restricted cloud environments. This is a foundational role where you will help establish the operating model, reliability practices, automation, and engineering standards needed to support critical customer-facing services.You will combine software engineering and production operations to deploy, run, monitor, improve, and automate Autodesk services in GovCloud. You will partner closely with product engineering, security, compliance, platform, and infrastructure teams to ensure services are reliable, scalable, secure, and ready for production.The ideal candidate has deep experience operating production systems at scale, an automation-first mindset, and the ability to improve reliability through engineering practices such as SLOs/SLIs, production readiness, incident management, observability, resilience testing, and toil reduction. Success in this role requires strong technical judgment, a customer-focused mindset, and a passion for using software engineering to solve operational problems at scale.In accordance with GovCloud Cloud Service Provider Security Requirements, this role must be performed by U.S. Citizens. Employment is contingent upon meeting all applicable government security and eligibility requirements, including necessary background investigations and government issued security clearances.ResponsibilitiesServe as a primary owner for the reliability, availability, performance, operability, and capacity of one or more production servicesDeploy, operate, maintain, and continuously improve production services running in Autodesk GovCloud environmentsPartner with engineering teams to ensure services are designed with reliability, scalability, security, and operability in mindDefine and operate reliability practices such as SLOs/SLIs, error budgets, production readiness reviews, service reviews, and operational health reviewsBuild automation to improve deployment safety, operational efficiency, incident response, and service recoveryDesign, develop, and maintain software, automation, and tooling that improve the reliability, scalability, and efficiency of production systemsImplement and improve monitoring, alerting, logging, tracing, and observability capabilities across supported servicesLead and participate in incident response, troubleshooting, and post-incident reviews focused on learning and continuous improvementDevelop and maintain operational documentation, runbooks, and recovery proceduresScale and enhance resilience testing and Gameday practices to validate system behavior, recovery capabilities, and operational readinessContinuously identify and eliminate operational toil through software engineering, automation, and process improvementEnsure supported services remain compliant with Autodesk security, privacy, and regulatory requirements, including FedRAMP and related controls where applicableParticipate in a 24x7 on-call rotation for production servicesFunction effectively in a fast-paced environment while helping establish and mature operational excellence practices for Autodesk GovCloudMinimum QualificationsB.S. or higher in Computer Science, Engineering, or a related technical discipline, or equivalent practical experience7+ years of experience in Site Reliability Engineering, Software Engineering, Platform Engineering, Cloud Infrastructure, or Production OperationsExperience operating and supporting customer-facing production services in large-scale cloud environmentsStrong understanding of reliability engineering principles, including SLOs/SLIs, observability, incident management, capacity planning, production readiness, and automationExperience with AWS, Azure, or other public cloud platformsExperience developing automation using languages such as Python, Go, Java, PowerShell, Bash, or similarExperience with Infrastructure as Code, CI/CD pipelines, deployment automation, and modern cloud operations practicesUnderstanding of security, compliance, and operational risk management in production environmentsStrong written and verbal communication skillsPreferred Qualifications10+ years of experience operating highly available, customer-facing production systemsExperience with AWS GovCloud, FedRAMP, IL4/IL5, or other regulated cloud environmentsExperience supporting services with stringent availability, reliability, and security requirementsExperience with containers, Kubernetes, cloud-native architectures, APIs, load balancing, networking, DNS, and distributed systemsExperience with observability platforms such as Splunk, Dynatrace, Datadog, CloudWatch, or similar technologiesExperience operating databases, storage platforms, messaging systems, caching technologiesExperience designing and implementing operational automation at scaleExperience leading or participating in Gamedays, disaster recovery exercises, resilience testing, or operational readiness reviewsStrong incident management experience, including technical leadership during major incidents and stakeholder communicationStrong collaboration skills and ability to work effectively across engineering, security, compliance, and operations teamsPassion for building reliable, secure, and scalable systems that customers can trustLearn MoreAbout AutodeskWelcome to Autodesk! Amazing things are created every day with our software – from the greenest buildings and cleanest cars to the smartest factories and biggest hit movies. We help innovators turn their ideas into reality, transforming not only how things are made, but what can be made.We take great pride in our culture here at Autodesk – it’s at the core of everything we do. Our culture guides the way we work and treat each other, informs how we connect with customers and partners, and defines how we show up in the world.When you’re an Autodesker, you can do meaningful work that helps build a better world designed and made for all. Ready to shape the world and your future? Join us!BenefitsFrom health and financial benefits to time away and everyday wellness, we give Autodeskers the best, so they can do their best work. Learn more about our benefits in the U.S. by visiting Salary transparencySalary is one part of Autodesk’s competitive compensation package. For U.S.-based roles, we expect a starting base salary between $117,000 and $209,330. Offers are based on the candidate’s experience and geographic location, and may exceed this range. In addition to base salaries, our compensation package may include annual cash bonuses, commissions for sales roles, stock grants, and a comprehensive benefits package.Equal Employment OpportunityAt Autodesk, we're building a diverse workplace and an inclusive culture to give more people the chance to imagine, design, and make a better world. Autodesk is proud to be an equal opportunity employer and considers all qualified applicants for employment without regard to race, color, religion, age, sex, sexual orientation, gender, gender identity, national origin, disability, veteran status or any other legally protected characteristic. We also consider for employment all qualified applicants regardless of criminal histories, consistent with applicable law.BelongingWe take pride in cultivating a culture of belonging where everyone can thrive. Learn more here: In-Person Onboarding and Identity VerificationThis role may require in-person onboarding and/or in-person ID verification.Are you an existing contractor or consultant with Autodesk? Please search for open jobs and apply internally (not on this external site).SummaryLocation: San Francisco, CA, USAType: Full time
- About the RoleWe’re looking for an experienced Site Reliability Engineer (SRE) to help us scale our platform with reliability, observability, and operational excellence at the core. You’ll partner with engineers and data scientists to build, automate, and maintain the infrastructure...Senior
$152.5k - $205k
...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible for:As a Senior Site Reliability Engineer on Circle’s platform team, you’ll design, build, and operate the secure, scalable platform infrastructure behind...SeniorFlexible hours$152.5k - $205k
...flexible work environment where new ideas are encouraged and everyone is a stakeholder.What you’ll be responsible forThe Site Reliability Engineer builds and maintains shared platform capabilities, common libraries, and infrastructure that help Circle teams ship secure...SeniorFlexible hours$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours- DescriptionWe are looking for a Senior or Staff level Site Reliability Engineer to strengthen the reliability, scalability, and operational maturity of our platform in San Francisco, California. This role will focus on improving service health, refining observability, and...Senior
$165k - $225.6k
...From core infrastructure to enterprise platforms, we partner across functions to drive scale, reliability, and innovation through technology.The Senior Site Reliability Engineer OpportunityReporting to the Manager, Site Reliability Engineering, this role will help build,...SeniorPermanent employmentLocal areaWorldwideFlexible hours- ...’s build what’s next.About the teamThe Engineering team at Airwallex is a diverse group of... ...ownership, working together to build scalable, reliable, and secure products that empower... ...our Global services.What you’ll doAs a Senior Site Reliability Engineer, you’ll work...SeniorTemporary workLocal areaWorldwide
$148.5k - $223.9k
...right place! Agentforce is the future of AI, and you are the future of Salesforce.Salesforce is seeking a senior engineering candidate to join the Site Reliability organization in San Francisco. Working closely with counterparts in the Infrastructure and R&D organizations...SeniorFull timeWorldwideWeekend work$167.7k - $245.2k
...very effective.We’re looking for talented engineers with a software or operations background... ...development teams to ensure the reliability, performance and security of our infrastructure... ...insurance. Please see the Cisco careers site to discover more benefits and perks....SeniorFull timeTemporary workWork at officeLocal areaFlexible hours1 day per week- ...getting here.)About the RoleWe're building infrastructure that has to perform under real-world scale, reliability, and security demands — and we're looking for an engineer who wants to own the foundation it runs on. This isn't a traditional "keep the lights on" role.You'...Senior
$127k - $249k
We are looking for an experienced Senior or Staff Engineer for our SRE, InfraSec team, to guide the security of our cloud-based infrastructure. As a Staff SRE, you will be very hands-on technically while also mentoring a small team of SREs.The InfraSec team collaborates...SeniorLocal areaRemote workWorldwideFlexible hours$175k - $250k
...000.00/yr - $250,000.00/yr Job Title: Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On-Site only. Must live within commuting distance... ...scalability, performance, and reliability across environments. What You’ll Do Design...SeniorFull timeRemote workRelocationRelocation package$165k - $241.4k
...within Cisco’s Networking, Security, Collaboration, and Observability portfolios.Your ImpactWe are seeking a skilled Senior Site Reliability Engineer (SRE) in Production Engineering with a strong background in SaaS and operations. You will design and manage large-scale...SeniorFull timeTemporary workWork at officeLocal areaFlexible hours1 day per week- ...career: come shape the future and be part of a truly unique global culture at OutSystems! Hybrid Onsite in Menlo Park, CA Site Reliability Engineering (SRE) is a discipline that incorporates aspects of software engineering and applies them to infrastructure and...SeniorImmediate startRemote workWorldwide
$210.8k - $272.8k
Thumbtack is hiring for a Site Reliability Engineer to enhance the reliability and scalability of our services in San Francisco, CA. You'll design and support resilient systems, ensuring a smooth user experience. The ideal candidate will have extensive experience with AWS...Senior$181k - $263k
...and supporting deployments of global products, and providing first line operational support. We are looking for a Senior Staff Site Reliability Engineer who will set the technical direction for reliability engineering across LiveRamp's global infrastructure. This is a...SeniorFull timeWork from homeWorldwideFlexible hoursNight shift$232k - $319k
...to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and... ...enabled with self-serviceAccelerate the velocity of SRE and product engineering by developing robust platforms, powerful tooling, and...SeniorPermanent employmentLocal areaWorldwideFlexible hours- ...Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will...Senior
$250k
...across Europe, while now significantly expanding its footprint in the United States. The company is looking for a Senior / Staff Site Reliability Engineer to support and scale large-scale HPC and cloud environments powering GPU-intensive workloads. The role involves...SeniorFull timeRemote work$210.8k - $272.8k
About Thumbtack Thumbtack helps millions of people confidently care for their homes. About the Site Reliability Engineering Team The Site Reliability Engineering team focuses on creating and maintaining a reliable, secure, and scalable platform vital for a seamless user...SeniorLocal area- CloudDevs works with fast-moving, venture-backed startups across the US. We’re building a pool of world-class Site Reliability Engineers for current roles and for upcoming opportunities. You will either be placed directly into one of our partner startups or added to our...SeniorLocal area
$165k - $227k
...opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Engineering Opportunity We are looking for an experienced Senior Site Reliability Engineer to join Okta's Emerging Products Group (EPG). Our mission is to build highly reliable...SeniorWorldwideFlexible hours$15k
...benefits packages, technology talks by our experts, a beautiful modern office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to meet our growing needs, and you will leverage...SeniorWork at officeLocal areaRemote work$139.76k - $287.75k
...to grow their business.We are seeking a Senior Site ReliabilityEngineer to help operate,... ...will be instrumental in advancing the reliability, scalability, automation, observability... ...The ideal candidate is a highly hands-on engineer with strong production experience and a...SeniorWork at officeLocal areaRelocationRelocation package$174.92k - $209.91k
...same: to make access to data as simple and reliable as electricity. With Fivetran, customer... ..., canonical and ready to query, with no engineering or maintenance required. We’re proud... ...integrate our teams, systems, and career sites.About the RoleFivetran is building data...SeniorFull timeWork at officeRemote work$174.92k - $209.91k
...High-Performance Engineer For Site Reliability Engineering Team Fivetran is building data pipelines to power the modern data stack for thousands of companies. Fivetran is looking for a high-performance, experienced engineer to be a part of a team of Site Reliability...SeniorFull timeWork at officeRemote work- ...better than we found it. The Apple Service Engineering (ASE) team builds and provides systems... ...The ASE Compute team is looking for a senior SRE software engineer to own the... ...management infrastructure, strengthen the reliability of our Kubernetes services, and engage...Senior
- Apple Service Engineering (ASE) seeks a senior SRE software engineer to own the architectural direction of Kubernetes internals powering Apple services... ...will define controllers and namespace management, raise reliability, and contribute to upstream Kubernetes. The role...Senior
- ...Infrastructure team builds the platforms and tooling that help engineering teams develop, deploy, and operate production systems safely... ...safe shipping the default for every product team.As a Staff Site Reliability Engineer on Release Engineering, you'll define and scale...Permanent employmentWork experience placementWork at officeLocal area
$300k
...thousands of H100s, H200s, and B200s, ready for experimentation, full-scale model training, or inference. As a Platform Engineer/Senior Site Reliability Engineer, you’ll own the reliability, performance, and automation of this GPU-powered infrastructure, ensuring...SeniorPermanent employment
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer remote San Francisco, CA
- site reliability engineer San Francisco, CA
- site reliability engineer sre San Francisco, CA
- srs distribution San Francisco, CA
- senior operations coordinator San Francisco, CA
- senior associate architect San Francisco, CA
- senior dynamics crm developer San Francisco, CA
- senior application security San Francisco, CA
- senior account director San Francisco, CA
- sr hr business partner San Francisco, CA


