Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

$91.4k - $187k

Oracle

Job Description

OCI Incident Response is the first line of defense in maintaining the high availability of Oracle’s cloud. We minimize customer-impacting events by making them shorter, less frequent, and less impactful through large-scale incident management. We are at the forefront of reducing event duration by leveraging our operational experience, knowledge of best practices, and ability to develop tools that automate incident management.

Description

We are looking for a Senior Site Reliability Engineer to join our OCI team. This role is part of a globally distributed team responsible for detecting, triaging, and mitigating OCI service-impacting events as quickly as possible. You will be part of one of these regional teams and will be responsible for minimizing the downtime of OCI services. You will achieve this by delivering excellent major incident management and operating systems with high scalability, performance, and security that help prevent incidents from occurring.

Oracle’s Cloud is state-of-the-art and constantly evolving. When issues arise, your team will respond within minutes to ensure customer impact is minimized. This role will expose you to the inner workings of OCI’s systems and organization. You will interact with and influence leaders across Oracle and drive broad, cross-organization programs aimed at iteratively improving OCI-wide service availability. We are an agile team with significant impact. If you want to be part of a fast-moving team breaking new ground, we would love to speak with you!

We are looking for candidates who are flexible to work AMER shift hours (9:30 AM to 5:30 PM PST) on a rotating roster, including occasional weekends and public holidays.

Career Level - IC3

Responsibilities

Responsibilities:

  • Solve complex problems related to infrastructure cloud services and automate common tasks to ensure continuous availability with minimal human intervention.

  • Command and coordinate SMEs and service leaders to restore services as quickly as possible during major incidents, while keeping accurate and timely data on the progress of such incidents.

  • Utilize a deep understanding of cloud computing design patterns and their dependencies to mitigate complex major incidents.

  • Embed a methodical approach to troubleshoot large, complex, interconnected systems used in incident detection and orchestration.

  • Document pertinent information related to incidents that aids process improvement, identifies deviations, and enables the creation of an incident knowledge base.

  • Monitor and evaluate high-level service and infrastructure dashboards, taking action to address identified anomalies.

  • Identify opportunities and take ownership of automation and/or continuous improvement of incident management process steps and best practices.

  • Define and document the technical architecture of large-scale distributed systems.

  • Understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of production services.

  • Be responsible for the design and delivery of the mission-critical stack, with a focus on security, resiliency, scalability, and performance.

  • Partner with development teams to define operational requirements for product roadmaps.

  • Articulate the technical characteristics of services and technology areas, and guide development teams to engineer and add premier capabilities to the Oracle Cloud service portfolio.

  • Act as the ultimate escalation point for complex or critical issues that have not yet been documented as Standard Operating Procedures (SOPs).

Minimum Qualifications:

Bachelor’s degree or higher in Computer Science or relevant work experience..

  • 3+ years’ experience in Site Reliability Engineering, DevOps, or System Engineering.

  • Must have public cloud operations experience (e.g., AWS, Azure, GCP, OCI).

  • Extensive experience with Major Incident Management in a cloud-based environment.

  • Demonstrate clear understanding of automation and orchestration principles.

  • Experience having worked in at least one modern object-oriented programming language.

  • Experience with professional software engineering standard methodologies such as Agile project management, coding standards, code reviews, source control management, build processes, testing, and operations.

  • Familiarity with infrastructure automation tools such as Chef, Ansible, Jenkins, Terraform

  • Excellent expertise with several of following technologies: Infrastructure-as-a-Service, CI/CD systems, Docker, RESTful APIs, log analysis tools, debugging tools

Disclaimer:

Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.

Range and benefit information provided in this posting are specific to the stated locations only

US: Hiring Range in USD from: $91,400 to $187,000 per annum. May be eligible for bonus and equity.

Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.

Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.

Oracle US offers a comprehensive benefits package which includes the following:

  1. Medical, dental, and vision insurance, including expert medical opinion

  2. Short term disability and long term disability

  3. Life insurance and AD&D

  4. Supplemental life insurance (Employee/Spouse/Child)

  5. Health care and dependent care Flexible Spending Accounts

  6. Pre-tax commuter and parking benefits

  7. 401(k) Savings and Investment Plan with company match

  8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.

  9. 11 paid holidays

  10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.

  11. Paid parental leave

  12. Adoption assistance

  13. Employee Stock Purchase Plan

  14. Financial planning and group legal

  15. Voluntary benefits including auto, homeowner and pet insurance

The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.

Career Level - IC3

About Us

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.

We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing View email address on click.appcast.io or by calling View phone number on click.appcast.io in the United States.

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Denver, CO vacancy
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper).... 
    Senior
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Denver, CO
    21 hours ago
  • $86.9k - $198k

    Site Reliability Engineer, SeniorThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether you come from a background in network engineering, systems administration, or software development, if... 
    Senior
    Full time
    Contract work
    Part time
    Work at office
    Local area
    Remote work

    Booz Allen Hamilton

    Aurora, CO
    21 hours ago
  • $121.4k - $218.6k

     ...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner...  ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling... 
    Senior
    Work experience placement
    Work at office

    Akamai

    Denver, CO
    1 day ago
  • $81.1k - $187k

     ...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection... 
    Senior
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle

    Denver, CO
    4 days ago
  •  ...Position - Senior Site Reliability Engineer Location - 100% Remote Experience - 10+ Years Full Time Hiring Job Description - Must Have Technical/Functional Skills: • 10+ years of experience in SRE, DevOps, or infrastructure engineering... 
    Senior
    Full time
    Remote work

    vaaridatech

    Denver, CO
    29 days ago
  • $110k - $145k

     ...global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. We are seeking a Senior Site Reliability Engineer to own the reliability, scalability, performance, and operational integrity of critical production services. This role... 
    Senior
    Contract work
    Temporary work
    Work at office
    Work from home
    Flexible hours

    Vertafore

    Denver, CO
    2 days ago
  • $130k - $160k

     ...Description NBCU is looking for creative engineers willing to learn from the current process...  .../existing systems; propose & deploy more reliable scalable solutions Responsible for...  ...monitoring deliverables to improve site reliability Evaluate new software releases... 
    Senior
    Full time
    Work at office
    Local area
    Rotating shift

    NBCUniversal

    Centennial, CO
    15 days ago
  • $160k - $200k

     ...Job Description Job Description Description TL;DR Kharon is seeking a full-time Senior Site Reliability Engineer based in Denver. This role requires in-office attendance at least 3 days a week.  RESPONSIBILITIES: Spearhead the full lifecycle management of our... 
    Senior
    Full time
    Work at office
    Immediate start
    Flexible hours
    3 days per week

    Kharon

    Denver, CO
    14 days ago
  • $104.43k - $156.65k

     ...Comcast. (In most cases, Comcast prefers to have employees on-site collaborating unless the team has been designated as virtual...  ..., Fox, Disney, NBC, Paramount+, and many others.Our Site Reliability Engineering (SRE) team is at the heart of our mission to deliver seamless... 
    Permanent employment
    Full time
    Work at office
    Remote work
    Worldwide
    Flexible hours

    Comcast

    Centennial, CO
    5 days ago
  • $105.6k - $145.2k

    Architect the Future as our Site Reliability Engineer!Are you ready to take your skills to the next level as a self-motivated and enthusiastic Site Reliability Engineer with hands-on experience supporting multiple connected Cloud-based products? Trimble is a global technology... 
    Ongoing contract
    Full time
    Work at office
    Local area
    Worldwide

    Trimble Navigation

    Westminster, CO
    4 days ago
  • $112.5k - $187.5k

     ...TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering...  ...Site Reliability Engineer at TransUnion, you will serve as a senior technical leader and force multiplier on the SRE team.... 
    Full time
    Temporary work
    Work experience placement
    Work at office
    Flexible hours
    2 days per week

    TransUnion

    Greenwood Village, CO
    1 day ago
  • $75.7k - $136.3k

     ...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and... 
    Work experience placement
    Work at office

    Akamai

    Denver, CO
    1 day ago
  • $87.5k - $143.75k

     ...Observability Engineer DISH is transforming the future of connectivity. We're doing it by building the country's first virtualized...  ...Observability Tools Personal responsibility for the quality, reliability, and usability of the NOC Observability tools, including... 
    Flexible hours
    Night shift

    Phenom People

    Littleton, CO
    2 days ago
  • $95k - $171k

     .... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts... 
    Permanent employment
    Work experience placement
    Work at office
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Akamai

    Denver, CO
    1 day ago
  • $87.4k - $123.4k

     ...or any other status protected by applicable state or local law. ***For remote and hybrid positions you will be required to provide reliable high-speed internet with a wired connection as well as a place in your home to work with limited disruption. You must have... 
    16 hours
    Contract work
    Temporary work
    Work experience placement
    Casual work
    Work at office
    Local area
    Remote work
    Work from home
    Work visa
    Flexible hours

    Empower Retirement

    Greenwood Village, CO
    21 hours ago
  •  ...integration into Creo and Onshape products. Deliver high-quality, innovative solutions by applying first principles to address the needs of engineers. Collaborate with other developers, quality assurance and software engineers. Master's degree or higher in Computational Mechanics... 
    Senior
    Full time

    Ptc

    Denver, CO
    21 hours ago
  • $160k - $190k

     ...Site Reliability Engineer (Classified Deployments) Location: Southern California or Washington, D.C. Clearance: Active Secret required; TS/SCI strongly preferred Work Mode: Hybrid/On-site with government customers Citizenship: U.S. Citizen Compensation:... 

    Zachary Piper Solutions

    Arvada, CO
    4 days ago
  • $110k - $145k

     ...content reflecting our world. NBCU's Distribution engineering is responsible for the automation and reliability of NBCU's Live sources. Reasonable for the...  ...Distribution Engineering is looking to add a talented Site Reliability Engineer to be part of our Video Streaming... 
    Work experience placement
    Work at office
    Local area

    NBCUniversal

    Greenwood Village, CO
    8 hours ago
  • $98.58k - $138.02k

     ...Site Reliability Engineer II Restaurant365 is a SaaS company disrupting the restaurant industry! Our cloud-based platform provides a unique, centralized solution for accounting and back-office operations for restaurants. Restaurant365's culture is focused on empowering... 
    Work at office

    Restaurant365

    Denver, CO
    1 day ago
  • $94.85k - $135.5k

     ...AI-powered business communications. This is where you and your skills come in. We're currently looking for: An experienced Site Reliability Engineer (SRE) to join the RingCentral Collaboration team. As a SRE, you will be responsible for maintaining and improving uptime... 
    Full time
    Local area
    Flexible hours

    RingCentral

    Denver, CO
    4 days ago
  • $148.5k - $260.1k

     ...CAC/PIV.Distributed Systems Software Engineer - GovCloud (Senior/Lead) Our Public Cloud engineering teams...  ...count on our platform to be highly reliable, lightning fast, supremely secure, and...  ...services. You have experience balancing live-site management, feature delivery, and... 
    Senior
    Full time
    Local area

    Salesforce

    Denver, CO
    2 days ago
  • $114k - $165.3k

     .... We are unable to sponsor or take over sponsorship of an employment visa at this time, including CPT/OPT.*** The Lead Site Reliability Engineer will combine deep technical expertise with team leadership to drive reliability across Empower's financial services platform... 
    16 hours
    Contract work
    Temporary work
    Work experience placement
    Casual work
    Work at office
    Local area
    Remote work
    Work from home
    Work visa
    Flexible hours

    Empower Retirement

    Greenwood Village, CO
    4 days ago
  • $120k - $160k

     ...office at least 3 days a week for collaboration and connection. Why this Role Matters The Manager, Site Reliability Engineering plays a critical role in ensuring Litera's platforms remain reliable, scalable, secure, and high performing for our customers... 
    Work experience placement
    Work at office
    Worldwide
    3 days per week

    Litera Corp.

    Denver, CO
    1 day ago
  • $175k - $220k

     ...is global, with headquarters in Denver, Colorado, and offices across the U.S., Canada, and India. The Director, Site Reliability Engineering (SRE) will lead reliability, performance, and observability initiatives for a portfolio of Vertafore products. This role... 
    Contract work
    Temporary work
    Work at office
    Work from home
    Flexible hours

    Vertafore

    Denver, CO
    2 days ago
  •  ...Senior Software Engineer/Developer - Android Description seeking a Sr. Engineer/Developer Android to join their team in Denver...  ...or SDKs Writing unit and integration tests to increase reliability and quality of solutions Completing documentation and procedures... 
    Senior
    Full time
    Work at office
    2 days per week
    1 day per week

    Esrhealthcare

    Denver, CO
    21 hours ago
  • $120.54k - $140k

     ...Job Title: Senior Software Engineer Employer: Procare Software, LLC Job Location: Denver, CO Salary: $120,536 - $140,000 Job Duties: Create and support enterprise software solutions both web and mobile applications by maintaining and supporting existing... 
    Senior
    Full time
    Remote work

    Procare Solutions

    Denver, CO
    21 hours ago
  • $165k - $216.56k

     ...redefine the future of how work gets done.We are looking for a Senior Solution Engineer who is accustomed to solving customer’s most complex...  ...States, please visit the job posting on the Snowflake Careers Site for salary and benefits information: careers.snowflake.comCompensation... 
    Senior

    Snowflake

    Denver, CO
    5 days ago
  • Legion Technical Solutions is seeking a Senior Principal Systems Engineer to support a premier national security space program in Aurora, CO. You will lead system architecture, requirements elaboration, and large-scale integration across multidisciplinary teams to deliver... 
    Senior

    Legion Technical Solutions LLC

    Aurora, CO
    4 days ago
  • $98.5k - $206.8k

    Job Title: Senior Platform EngineerJob Category: EngineeringTime Type: Full timeMinimum Clearance Required to Start: TS/SCIEmployee...  ...accepted). 5+ years of platform, infrastructure, or software engineering experience. Hands-on experience with Kubernetes and container... 
    Senior
    Contract work
    Work experience placement
    Flexible hours

    CACI International

    Denver, CO
    2 days ago
  • $143.5k

     ...Wireless, OnTech and GenMobile.Job Duties and ResponsibilitiesSenior Engineer-Software Quality sought by DISH Network, LLC in Englewood,...  ...review test codes for functionality, code coverage, quality, reliability and determine whether it sufficiently covers all conditions... 
    Senior

    EchoStar

    Englewood, CO
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!