Senior Site Reliability Engineer
$91.4k - $187kOracle
Job Description
OCI Incident Response is the first line of defense in maintaining the high availability of Oracle’s cloud. We minimize customer-impacting events by making them shorter, less frequent, and less impactful through large-scale incident management. We are at the forefront of reducing event duration by leveraging our operational experience, knowledge of best practices, and ability to develop tools that automate incident management.
Description
We are looking for a Senior Site Reliability Engineer to join our OCI team. This role is part of a globally distributed team responsible for detecting, triaging, and mitigating OCI service-impacting events as quickly as possible. You will be part of one of these regional teams and will be responsible for minimizing the downtime of OCI services. You will achieve this by delivering excellent major incident management and operating systems with high scalability, performance, and security that help prevent incidents from occurring.
Oracle’s Cloud is state-of-the-art and constantly evolving. When issues arise, your team will respond within minutes to ensure customer impact is minimized. This role will expose you to the inner workings of OCI’s systems and organization. You will interact with and influence leaders across Oracle and drive broad, cross-organization programs aimed at iteratively improving OCI-wide service availability. We are an agile team with significant impact. If you want to be part of a fast-moving team breaking new ground, we would love to speak with you!
We are looking for candidates who are flexible to work AMER shift hours (9:30 AM to 5:30 PM PST) on a rotating roster, including occasional weekends and public holidays.
Career Level - IC3
Responsibilities
Responsibilities:
Solve complex problems related to infrastructure cloud services and automate common tasks to ensure continuous availability with minimal human intervention.
Command and coordinate SMEs and service leaders to restore services as quickly as possible during major incidents, while keeping accurate and timely data on the progress of such incidents.
Utilize a deep understanding of cloud computing design patterns and their dependencies to mitigate complex major incidents.
Embed a methodical approach to troubleshoot large, complex, interconnected systems used in incident detection and orchestration.
Document pertinent information related to incidents that aids process improvement, identifies deviations, and enables the creation of an incident knowledge base.
Monitor and evaluate high-level service and infrastructure dashboards, taking action to address identified anomalies.
Identify opportunities and take ownership of automation and/or continuous improvement of incident management process steps and best practices.
Define and document the technical architecture of large-scale distributed systems.
Understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of production services.
Be responsible for the design and delivery of the mission-critical stack, with a focus on security, resiliency, scalability, and performance.
Partner with development teams to define operational requirements for product roadmaps.
Articulate the technical characteristics of services and technology areas, and guide development teams to engineer and add premier capabilities to the Oracle Cloud service portfolio.
Act as the ultimate escalation point for complex or critical issues that have not yet been documented as Standard Operating Procedures (SOPs).
Minimum Qualifications:
Bachelor’s degree or higher in Computer Science or relevant work experience..
3+ years’ experience in Site Reliability Engineering, DevOps, or System Engineering.
Must have public cloud operations experience (e.g., AWS, Azure, GCP, OCI).
Extensive experience with Major Incident Management in a cloud-based environment.
Demonstrate clear understanding of automation and orchestration principles.
Experience having worked in at least one modern object-oriented programming language.
Experience with professional software engineering standard methodologies such as Agile project management, coding standards, code reviews, source control management, build processes, testing, and operations.
Familiarity with infrastructure automation tools such as Chef, Ansible, Jenkins, Terraform
Excellent expertise with several of following technologies: Infrastructure-as-a-Service, CI/CD systems, Docker, RESTful APIs, log analysis tools, debugging tools
Disclaimer:
Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.
Range and benefit information provided in this posting are specific to the stated locations only
US: Hiring Range in USD from: $91,400 to $187,000 per annum. May be eligible for bonus and equity.
Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.
Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.
Oracle US offers a comprehensive benefits package which includes the following:
Medical, dental, and vision insurance, including expert medical opinion
Short term disability and long term disability
Life insurance and AD&D
Supplemental life insurance (Employee/Spouse/Child)
Health care and dependent care Flexible Spending Accounts
Pre-tax commuter and parking benefits
401(k) Savings and Investment Plan with company match
Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
11 paid holidays
Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
Paid parental leave
Adoption assistance
Employee Stock Purchase Plan
Financial planning and group legal
Voluntary benefits including auto, homeowner and pet insurance
The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.
Career Level - IC3
About Us
Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.
True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.
We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing View email address on click.appcast.io or by calling View phone number on click.appcast.io in the United States.
Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....SeniorWork at officeLocal areaRemote workWorldwideFlexible hours$86.9k - $198k
Site Reliability Engineer, SeniorThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if...SeniorFull timeContract workPart timeWork at officeLocal areaRemote work$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling...SeniorWork experience placementWork at office$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...SeniorTemporary workImmediate startFlexible hoursShift work- ...Position - Senior Site Reliability Engineer Location - 100% Remote Experience - 10+ Years Full Time Hiring Job Description - Must Have Technical/Functional Skills: • 10+ years of experience in SRE, DevOps, or infrastructure engineering...SeniorFull timeRemote work
$110k - $145k
...global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role...SeniorContract workTemporary workWork at officeWork from homeFlexible hours$130k - $160k
...Description NBCU is looking for creative engineers willing to learn from the current process... .../existing systems; propose & deploy more reliable scalable solutions Responsible for... ...monitoring deliverables to improve site reliability Evaluate new software releases...SeniorFull timeWork at officeLocal areaRotating shift$160k - $200k
...Job Description Job Description Description TL;DR Kharon is seeking a full-time Senior Site Reliability Engineer based in Denver. This role requires in-office attendance at least 3 days a week. RESPONSIBILITIES: Spearhead the full lifecycle management of our...SeniorFull timeWork at officeImmediate startFlexible hours3 days per week$104.43k - $156.65k
...Comcast. (In most cases, Comcast prefers to have employees on-site collaborating unless the team has been designated as virtual... ..., Fox, Disney, NBC, Paramount+, and many others.Our Site Reliability Engineering (SRE) team is at the heart of our mission to deliver seamless...Permanent employmentFull timeWork at officeRemote workWorldwideFlexible hours$105.6k - $145.2k
Architect the Future as our Site Reliability Engineer!Are you ready to take your skills to the next level as a self-motivated and enthusiastic Site Reliability Engineer with hands-on experience supporting multiple connected Cloud-based products? Trimble is a global technology...Ongoing contractFull timeWork at officeLocal areaWorldwide$112.5k - $187.5k
...TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering... ...Site Reliability Engineer at TransUnion, you will serve as a senior technical leader and force multiplier on the SRE team....Full timeTemporary workWork experience placementWork at officeFlexible hours2 days per week$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office$87.5k - $143.75k
...Observability Engineer DISH is transforming the future of connectivity. We're doing it by building the country's first virtualized... ...Observability Tools Personal responsibility for the quality, reliability, and usability of the NOC Observability tools, including...Flexible hoursNight shift$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$87.4k - $123.4k
...or any other status protected by applicable state or local law. ***For remote and hybrid positions you will be required to provide reliable high-speed internet with a wired connection as well as a place in your home to work with limited disruption. You must have...16 hoursContract workTemporary workWork experience placementCasual workWork at officeLocal areaRemote workWork from homeWork visaFlexible hours- ...integration into Creo and Onshape products. Deliver high-quality, innovative solutions by applying first principles to address the needs of engineers. Collaborate with other developers, quality assurance and software engineers. Master's degree or higher in Computational Mechanics...SeniorFull time
$160k - $190k
...Site Reliability Engineer (Classified Deployments) Location: Southern California or Washington, D.C. Clearance: Active Secret required; TS/SCI strongly preferred Work Mode: Hybrid/On-site with government customers Citizenship: U.S. Citizen Compensation:...$110k - $145k
...content reflecting our world. NBCU's Distribution engineering is responsible for the automation and reliability of NBCU's Live sources. Reasonable for the... ...Distribution Engineering is looking to add a talented Site Reliability Engineer to be part of our Video Streaming...Work experience placementWork at officeLocal area$98.58k - $138.02k
...Site Reliability Engineer II Restaurant365 is a SaaS company disrupting the restaurant industry! Our cloud-based platform provides a unique, centralized solution for accounting and back-office operations for restaurants. Restaurant365's culture is focused on empowering...Work at office$94.85k - $135.5k
...AI-powered business communications. This is where you and your skills come in. We're currently looking for: An experienced Site Reliability Engineer (SRE) to join the RingCentral Collaboration team. As a SRE, you will be responsible for maintaining and improving uptime...Full timeLocal areaFlexible hours$148.5k - $260.1k
...CAC/PIV.Distributed Systems Software Engineer - GovCloud (Senior/Lead) Our Public Cloud engineering teams... ...count on our platform to be highly reliable, lightning fast, supremely secure, and... ...services. You have experience balancing live-site management, feature delivery, and...SeniorFull timeLocal area$114k - $165.3k
.... We are unable to sponsor or take over sponsorship of an employment visa at this time, including CPT/OPT.*** The Lead Site Reliability Engineer will combine deep technical expertise with team leadership to drive reliability across Empower's financial services platform...16 hoursContract workTemporary workWork experience placementCasual workWork at officeLocal areaRemote workWork from homeWork visaFlexible hours$120k - $160k
...office at least 3 days a week for collaboration and connection. Why this Role Matters The Manager, Site Reliability Engineering plays a critical role in ensuring Litera's platforms remain reliable, scalable, secure, and high performing for our customers...Work experience placementWork at officeWorldwide3 days per week$175k - $220k
...is global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. The Director, Site Reliability Engineering (SRE) will lead reliability, performance, and observability initiatives for a portfolio of Vertafore products. This role...Contract workTemporary workWork at officeWork from homeFlexible hours- ...Senior Software Engineer/Developer - Android Description seeking a Sr. Engineer/Developer Android to join their team in Denver... ...or SDKs Writing unit and integration tests to increase reliability and quality of solutions Completing documentation and procedures...SeniorFull timeWork at office2 days per week1 day per week
$120.54k - $140k
...Job Title: Senior Software Engineer Employer: Procare Software, LLC Job Location: Denver, CO Salary: $120,536 - $140,000 Job Duties: Create and support enterprise software solutions both web and mobile applications by maintaining and supporting existing...SeniorFull timeRemote work$165k - $216.56k
...redefine the future of how work gets done.We are looking for a Senior Solution Engineer who is accustomed to solving customer’s most complex... ...States, please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.comCompensation...Senior- Legion Technical Solutions is seeking a Senior Principal Systems Engineer to support a premier national security space program in Aurora, CO. You will lead system architecture, requirements elaboration, and large-scale integration across multidisciplinary teams to deliver...Senior
$98.5k - $206.8k
Job Title: Senior Platform EngineerJob Category: EngineeringTime Type: Full timeMinimum Clearance Required to Start: TS/SCIEmployee... ...accepted). 5+ years of platform, infrastructure, or software engineering experience. Hands-on experience with Kubernetes and container...SeniorContract workWork experience placementFlexible hours$143.5k
...Wireless, OnTech and GenMobile.Job Duties and ResponsibilitiesSenior Engineer-Software Quality sought by DISH Network, LLC in Englewood,... ...review test codes for functionality, code coverage, quality, reliability and determine whether it sufficiently covers all conditions...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Denver, CO
- site reliability engineer sre Denver, CO
- srs distribution Denver, CO
- senior associate architect Denver, CO
- senior dynamics crm developer Denver, CO
- senior application security Denver, CO
- senior account director Denver, CO
- sr hr business partner Denver, CO
- senior supervisor Denver, CO
- senior plumbing designer Denver, CO



