Senior Site Reliability Engineer
$81.1k - $187kOracle
Job Description
We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection and resolution of issues.
The engineer will work closely with development, infrastructure, security, and operations teams to monitor service health, troubleshoot production issues, participate in incident response, improve observability, and implement reliability best practices. This role also includes analyzing recurring failures, building automation, supporting deployments, and contributing to capacity planning, disaster recovery, and operational readiness.
Also works on number of different region/realm rollouts, deployments. Forecasts demands and responds to capacity needs. Collaborates with software development teams to develop reliable and scalable infrastructures. Performs data collection to maintain and optimize operations and reliability. Leverages knowledge to perform incident response and/or maintenance tasks. Provides health and performance reporting. Identifies opportunities for automation. Communicates about services and identifies and explains the potential impact of changes. Provides support for technology and document incidents. Experiments with new tools and assesses potential impact and develops knowledge of site reliability trends.
Responsibilities
Key Responsibilities
Capacity Ingestion and Management:
-Takes proactive steps to design and architect infrastructure and/or service according to terms for reliability and functionality.
-Forecasts demands for infrastructure and responds to capacity needs, ensuring systems have sufficient resources to handle current and future workloads.
-Collaborates with the software development team to develop infrastructures and features that are reliable and scalable according to deployment requirements.
-Independently identifies opportunities for and drives prototyping (e.g., testing new applications or infrastructures, assisting in onboarding).
Incident and Service Lifecycle Management:
-Performs data collection, triage, technical analysis, and redirection to maintain and optimize operations and infrastructure reliability.
-Independently monitors services, maintains up-to-date knowledge of their performance, and documents their condition.
-Leverages comprehensive knowledge to perform incident response, root cause analyses, and/or maintenance on assigned services (e.g., software installs, version upgrades, security updates, backup and recovery).
-Provides health and performance reporting and takes appropriate actions based on trends in data.
-May independently perform provisioning to support infrastructure, applications, and services.
-May perform standard and non-standard decommissioning (e.g., shutting down servers, removing data from databases) to remove objects that are no longer needed.
Automation:
-Identifies opportunities for automation and assesses potential benefits.
-Develops automation tools or scripts to provide solutions, gather metrics, monitor, analyze, mitigate, or remediate issues/defects within infrastructures.
-Independently conducts testing to ensure automation performs the task correctly and produces expected results.
Technical Communication and Guidance:
-Communicates the scale, capacity, security, performance attributes, and requirements of services and technology within and sometimes beyond immediate team.
-Identifies and explains the potential impact of infrastructure, feature, and tool changes, considering their impact on team operations.
Troubleshooting and Resolution:
-Provides operational support for technology, escalating incidents and other standard and non-standard issues arising within Oracle services.
-Participates in on-call shifts to address issues.
-Resolves technical issues spanning various services, investigating and debugging products in order to reach SLOs (service level objectives).
-Documents incidents and performs root cause analyses according to standard reporting methods.
-Independently performs post-mortem procedures to prevent incident reoccurrence.
Innovation and Improvement:
-Experiments with new tools and technologies to assess their potential impact on and improve infrastructure performance and reliability, ensuring adherence to security standards.
-Independently identifies and executes improvements for performance bottlenecks and deployments to ensure efficient resource usage, speed, and scalability.
-Develops knowledge of site reliability trends and shares new information with team members, management, and beyond to help others build, test, deploy and run services.
-Performs standard and non-standard analyses and provides clear data on production to contribute to business development decisions (e.g., design changes).
Core Responsibilities
Planning & Execution:
Independently manages work, monitoring timelines and deliverables to ensure projects or initiatives stay on track and meet requirements. Proactively prioritizes work and adapts to resource or timeline shifts, suggesting adjustments to maintain project efficiency.
Collaboration & Partnership:
Collaborates across teams to align on expectations and achieve shared objectives. Builds and maintains a comprehensive understanding of business, stakeholder, and/or customer needs to build and support effective partnerships. Actively listens to diverse perspectives and asks questions to ensure understanding of others.
Problem Solving:
Independently identifies and addresses standard and non-standard issues in accordance with standard practices, escalating more complex issues as appropriate. Analyzes data and/or information from multiple sources to troubleshoot standard and non-standard errors. Contributes to knowledge sharing and best practices.
Continuous Learning:
Embraces continuous learning by actively seeking to build knowledge and new skills and/or tools and staying current with industry trends and best practices. Seeks out and leverages feedback and training to improve skills. Contributes to a culture of continuous learning and knowledge sharing with team members.
Continuous Improvement:
Develops ideas and recommends updates to increase the efficiency and effectiveness of processes, protocols, and workflows within a team. Seeks input from team members on alternative approaches and methods for improving work.
IAC: Terraform, Chef, Ansible
Languages: Python, Java, Bash
Orchestration: Kubernetes, Helm
CI/CD: Jenkins
Observability: Grafana, Prometheus
Disclaimer:
Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.
Range and benefit information provided in this posting are specific to the stated locations only
US: Hiring Range in USD from: $81,100 to $187,000 per annum. May be eligible for bonus and equity.
Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.
Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.
Oracle US offers a comprehensive benefits package which includes the following:
Medical, dental, and vision insurance, including expert medical opinion
Short term disability and long term disability
Life insurance and AD&D
Supplemental life insurance (Employee/Spouse/Child)
Health care and dependent care Flexible Spending Accounts
Pre-tax commuter and parking benefits
401(k) Savings and Investment Plan with company match
Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
11 paid holidays
Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
Paid parental leave
Adoption assistance
Employee Stock Purchase Plan
Financial planning and group legal
Voluntary benefits including auto, homeowner and pet insurance
The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.
Career Level - IC3
About Us
Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.
True innovation starts when everyone is empowered to contribute. That's why we're committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.
We're committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing View email address on click.appcast.io or by calling View phone number on click.appcast.io in the United States.
Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans' status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
$146.4k - $263.6k
...Do you want to shape reliability practices for a new AI inference platform? Are you a senior technical leader who drives solutions... ...architecture decisions with product engineering teams, and shape SRE... ...at scale. As a Senior II Site Reliability Engineer, you will...SeniorWork experience placementWork at office$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...SuggestedPermanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours- ...Senior Systems Programmer – StorageRemote - United StatesJR012900 At Ensono, our Purpose is to be a relentless ally, disrupting... ...five traits are the key to achieving our purpose: Honesty, Reliability, Curiosity, Collaboration, and Passion. Role Summary :...SeniorTemporary workWork experience placementWork at officeRemote workFlexible hours
- ...Wayfair in Bengaluru seeks a Senior Staff Engineer to act as local SME, planning and building a platform that connects a wide range of internal and external frameworks while shaping the technical direction. The role requires 10+ years in software design, leading high...SeniorLocal area
$186.07k - $218.9k
...will be responsible for owning the design, development, and reliability of core platform services that underpin account and identity... ...Coinbase Workspace (unified organization management) Championing engineering standards, code and design review culture, and technical...SeniorLocal area- ...AI infrastructure, working with server, cloud, and platform engineering teams. Operationalize machine learning workflows and support... ...implement system enhancements to improve performance, scalability, reliability, and cost efficiency. Collaborate across divisions to support...SeniorWork at officeRemote work
$180k - $220k
...experiences to realize our bold vision for healthcare. Senior Software Engineer The Role As a Senior Software Engineer, you will lead... ...that advance Datavant’s platform scalability and reliability. You’ll drive technical design, coach peers, and ensure system...Senior- A financial services company is seeking a Sr. Distinguished Machine Learning Engineer to define and drive technical strategies for their personalization platform. The role involves collaboration across teams to enhance recommendation systems and build robust machine learning...SeniorRemote job
$170k - $210k
...countries. To learn more, visit franklincovey.com. Title: Senior Software Engineer Payroll Title: Sr. Software Engineer Division &... ...maintaining our high standards for quality, security, and reliability. You’ll also be a hands-on coach and mentor, helping...SeniorFull timeRemote work$115k - $192.9k
...In this position... As a Full Stack Software Engineer, you will collaborate with a broad array of talented software engineers to build amazing software that will transform the Ford Reservation Services into the standard for providing an excellent digital experience for...SeniorImmediate startFlexible hours$92.5k - $209.5k
...multi-tenant cloud environment. OCI gives engineers the opportunity to work on systems that... ..., security posture, and operational reliability. The team works on distributed build... ...We Looking For? We are looking for a senior software engineer with strong distributed...SeniorTemporary workFixed term contractFlexible hours$197.4k - $232k
...Location Type: Remote Department Engineering Compensation: $197.4K - $232K -... ...Streaming Platform. About the Role: Senior Software Engineers II at Confluent take... ...and technical decisions that balance reliability, scalability, performance, and...SeniorFull timeRemote work$79.2k - $209.5k
...feature and subsystem design, including tradeoffs around security, reliability, scalability, maintainability, performance, and change... ...standards, and reduce technical debt. Improve the team through engineering practices, operational practices, development process,...SeniorTemporary workFlexible hours$84.63k - $112.84k
...ownership, deliver meaningful impact, and help shape the future of AI-ready connectivity, join us today. The Role At Lumen, a Senior Software Developer joins a team focused on building and strengthening the infrastructure. In an environment where performance,...SeniorTemporary workRemote workWork from home$89.2k - $209.5k
...Role Summary Oracle Health Platform Engineering builds core platform capabilities that... ...documentation, and operations. We are seeking a Senior Software Developer (IC3) to design,... ...and strengthen platform security and reliability. Responsibilities Key...SeniorTemporary workFlexible hours$253.9k - $298.7k
...foster collaboration, connection, and alignment. Attendance is expected and fully supported. We're looking to hire a Senior Staff Software Engineer to join one of our teams within the Platform Product Group. The Platform Product Group's mission is to build a trusted,...SeniorLocal area$105k - $141.75k
...Team is looking for a highly skilled Mainframe Modernization Senior Consultant to provide technical support and/or leadership in the... ...application code across platforms. Familiarity with agile engineering practices like Test Automation, Test-Driven Development (TDD),...SeniorRemote workWorldwide$186.07k - $218.9k
...Attendance is expected and fully supported. We are looking for a Senior Software Engineer to join the Payment Rails team within Coinbase's Platform... ...that ensure every transaction is fast, secure, and reliable, directly enabling millions of customers to move value seamlessly...SeniorLocal area$186.07k - $218.9k
...while delivering exceptional customer experiences. As a Software Engineer on our team, you will play a key role in this transformation,... ...adoption. Debug complex technical issues to enhance system reliability, scalability, and ease of operation. Review and ensure the...SeniorLocal area- ...Parexel is seeking a Senior Real-World Evidence (RWE) Analyst Programmer Join a high-impact team as a remote Sr. Real-World Evidence... ...validate Real-World Data (RWD) to ensure consistency and reliability Implement programming based on RWE protocols using a...SeniorRemote workFlexible hours
$116k - $144k
...Amentum is a global leader in advanced engineering and innovative technology solutions, trusted by the United States and its allies to... ...across all 7 continents. We are seeking a highly skilled Senior Secure DevOps Engineer to join our team and play a critical role...SeniorHourly payContract workLocal areaRemote work- ...delivers real business value with AI. What You'll Do As a Senior AI Engineer, you will be a key contributor to our internal engineering... ..., and applications with an emphasis on scalability, reliability, and performance. Collaborate with cross-functional engineers...SeniorPermanent employmentFlexible hours
$126.07k - $196.98k
...industrial infrastructure---sustainable solutions and more modern living depend on Chemours chemistry. Chemours is seeking a Senior AI Engineer to join our growing AI & Data Science team. This is a remote role and will report directly to the Head of Data Science & AI....SeniorWork at officeLocal areaRemote work$165k - $241.4k
...work with many different people within Engineering and throughout Cisco to help build the infrastructure... ...environment. Focus on increasing reliability, performance, scalability, testability,... ...insurance. Please see the Cisco careers site to discover more benefits and perks....SeniorPermanent employmentFull timeTemporary workLocal areaFlexible hours- ...Senior Life Sciences Knowledge Engineer Company: Norstella Location: Remote, United States Date Posted: Jul 24, 2026 Employment Type: Full Time Job ID: R-2094 Description Senior Life Sciences Knowledge Engineer About us: Why Norstella? Norstella...SeniorFull timeTemporary workWork at officeLocal areaRemote workFlexible hoursShift work
- A leading data streaming platform company is seeking a Senior Software Engineer II to take ownership of critical backend systems that underpin their data streaming platform. This role involves the design and delivery of large-scale services, making architectural decisions...SeniorRemote job
$123.4k - $176.3k
...organization's software systems in a cross-functional team environment through adherence to established design control processes and good engineering practices. This job family programs and configures end user applications, systems, databases and websites to achieve the...SeniorTemporary workLocal areaImmediate startFlexible hours$94.2k
...driven risk involving PHI while advising engineering and security leadership on emerging AI threats... ...Ability to operate effectively as a senior individual contributor in a large, matrixed... ...or Remote Position Physical work site required Occasionally Disclaimer:...SeniorFor contractorsWork at officeLocal areaRemote work$135.2k - $306.4k
...Job Description Senior / Principal Software Engineer or Architect Database Engine, Search & Document Systems We are looking for senior systems... ...surfaces, performance, correctness, and operational reliability. This area is especially relevant if you have experience...SeniorFull timeTemporary workFlexible hours$132.23k - $176.31k
...transformation depends on trust-trust in our networks, our platforms, and our ability to protect what matters most. The Senior Lead Cloud Engineer position plays a critical role in delivering on that promise. Provide expert technical direction in the analysis of...SeniorTemporary workRemote workWork from home
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- senior technical product manager Santa Fe, NM
- senior medical science liaison Santa Fe, NM
- senior accountant remote Santa Fe, NM
- senior robotics software engineer Santa Fe, NM
- senior dynamics crm developer Santa Fe, NM
- senior compensation manager Santa Fe, NM
- senior storage engineer Santa Fe, NM
- senior vice president of operations Santa Fe, NM
- senior inventory accountant Santa Fe, NM
- senior cloud security engineer Santa Fe, NM


