Principal Site Reliability Engineer
$84.9k - $209.5kOracle
Job Description
Building off our Cloud momentum, Oracle has formed a new organization - Health Data Intelligence Platform. This team will focus on product development and product strategy for Oracle Health, while building out a complete platform supporting modernized, automated healthcare. This is a net new line of business, constructed with an entrepreneurial spirit that promotes an upbeat and creative environment. We are unencumbered and will need your contribution to make it a special engineering center with the focus on excellence.
Health Data Intelligence Platform has a rare opportunity to play a critical role in how Oracle Health products impact and redefine the healthcare industry by redefining how healthcare and technology intersect.
You will have the opportunity to:
· Reach billions of people with our products & services
· Create technology in which truly impacts the world
· Ability to have immediate impact on developing technology
· Unlimited growth potential with inspiring work
· Work with the best minds in the industry
· Enjoy working in an open, diverse, and productive environment
Responsibilities
What You'll Do
Service Ownership –You will be part of the SRE team, whose mission is the shared full stack ownership of a collection of services and/or technology areas, with our Development partners.
Ownership Scope – As an SRE, you will understand the end-to-end configuration, technical dependencies, and overall behavioral characteristics of the production services you own. In partnership with your Development partners, you will have the responsibility to ensure that services are designed and delivered to be critical with focus on security, resiliency, scale, and performance. SREs are the ultimate authority and are accountable for the end-to-end performance and operability of the services they own.
Service Design – As the Oracle Cloud evolves; you will partner with development teams in defining and implementing improvements in service architecture, both current and future. As an SRE, you will be a guide at articulating technical characteristics of your services and the dependencies between services, and guide Development teams to engineer and add premier capabilities to the Oracle Cloud service portfolio. As an SRE, you will support federal project submission process and security compliance for new platforms and system resources.
Operations Engineering – You will understand and be able to communicate the scale, capacity, security, performance attributes and requirements of the services you own. You are a domain guide, able to understand and communicate every characteristic of your service stack, such as:
o degradation and behavior under load of the services and their dependencies
o end-to-end tuning needs, optimizing resource utilization, as load patterns fluctuate
o Instrumentation and metrics that clearly describe the service behaviors
o scaling requirements and patterns
o resiliency and recoverability, ensuring that backup / restore and disaster recovery capabilities are implemented, tested and maintained
o Security operations and vulnerability remediation, verifying vulnerabilities are patched or remediated while conforming to corporate and federal security standards and processes
Automation – You will have a clear understanding of automation and orchestration principles, and will be eager to automate, wherever and whenever the possibility arises, while simultaneously eliminating technical debt. Automation must be part of your DNA.
· Prevention - Once you have authoritatively resolved an issue, you will immediately work on how to more quickly resolve the problem next time, with the goal to eventually prevent the problem happening ever again
· Technical Experts - As service owner, you are the ultimate partner concern point for complex or critical issues that have not yet been documented as SOPs for Level1 staff. You will usually get called in during major incidents as an SME, when the source of a problem is unclear. You will have the deep understanding of service topology and their dependencies required to solve issues and define mitigations.
· Broad Interests - SREs are a rare mix of sysadmins and development Engineers, and as such have the ability to understand and explain the affect of product architecture decisions on the ability to run as distributed systems. They are driven by professional curiosity and a desire to a develop deep understanding of the their services and the technologies they depend upon.
· Represent SRE - Proactive, self-motivated, customer-focused, organized, and a good communicator. SRE can be expected to represent Cloud products and engineering in critical forums.
Disclaimer:
Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.
Range and benefit information provided in this posting are specific to the stated locations only
US: Hiring Range in USD from: $84,900 to $209,500 per annum. May be eligible for bonus and equity.
Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.
Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.
Oracle US offers a comprehensive benefits package which includes the following:
Medical, dental, and vision insurance, including expert medical opinion
Short term disability and long term disability
Life insurance and AD&D
Supplemental life insurance (Employee/Spouse/Child)
Health care and dependent care Flexible Spending Accounts
Pre-tax commuter and parking benefits
401(k) Savings and Investment Plan with company match
Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
11 paid holidays
Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
Paid parental leave
Adoption assistance
Employee Stock Purchase Plan
Financial planning and group legal
Voluntary benefits including auto, homeowner and pet insurance
The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.
Career Level - IC4
About Us
Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.
True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.
We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing View email address on click.appcast.io or by calling View phone number on click.appcast.io in the United States.
Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
- ...Site Reliability Engineering (SRE) Team Lead The Site Reliability Engineering (SRE) team is foundational to the growth and scale of our platform. You and your team will help advance several initiatives tied to automation, SRE culture, and cloud architecture. You will...PrincipalShift work
- A leading software company in Massachusetts is seeking a Principal Software Engineer to ensure the long-term reliability and operational excellence of its platform. This role involves leading complex reliability initiatives, developing strategies for large-scale systems...Principal
$180k - $205k
...patient outcomes in cardiovascular care. They are hiring a Principal-level Platform Engineer to join a small, highly impactful infrastructure team... ..., CI/CD, observability, developer tooling, and platform reliability across a regulated engineering environment. This is a...PrincipalWork at officeRemote work$140k - $210.9k
...Senior Site Reliability Engineer Federal Reserve Financial Services (FRFS) delivers a suite of payments services to financial institutions via FedLine® Solutions, FedNowSM, Fedwire®, National Settlement Service (NSS), FedCash®, FedACH® (Automated Clearing House), and...SuggestedFull timeTemporary workPart timeWork at officeShift work$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling...SuggestedWork experience placementWork at office$104.9k - $174.7k
...scale, 24x7, distributed and fault-tolerant systems within agreed reliability objectives, whilst enabling the fast flow of feature and... ...strong automation skills. About team; This diverse team of Engineers in assisting multiple product teams as we continue to innovate...Local areaImmediate startWorldwide$146.4k - $263.6k
...passionate about cutting edge technology? Do you enjoy working with a diverse multi-national team of engineering talents? Join our highly skilled Site Reliability team Our team designs, develops, and manages applications and infrastructure that support Akamai'...Work experience placementWork at office$75.7k - $136.3k
...solve complex challenges? Do you have a passion for automation and building systems that scale? Join our highly skilled Site Reliability Engineering team! Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and...Work experience placementWork at office$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours- ...Job Title: Site Reliability Engineer Location: Remote with Quarterly visits to Chennai, Tamil Nadu, India Duration: Full-Time bout BigRio: BigRio is a remote-based, technology consulting firm headquartered in Boston, MA. We deliver software solutions...Full timeRemote work
$128k - $160k
...at a time. To those who see AI as a driver of progress, come build the future together. The Crown Is Yours As a Senior Site Reliability Engineer, you'll build and scale the critical infrastructure behind every product. In this role, you'll take on complex challenges...Full timeImmediate start$146.4k - $263.6k
...uses large datasets to analyze and measure the performance and reliability of our platform. We are networking data scientists: we... ...organizing project‑specific work. Driving partnership with Engineering, Operations and Product teams to help guide adoption of new features...Work at office- ...commuting distance of one of our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team, you will operate the... ...ideal candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based...Full time
$150k - $185k
...and frameworks that work best. You enjoy building for other engineers equally, if not more, than building for a customer. You know... ...and evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Help architect our systems...Temporary workWork at officeFlexible hours3 days per week$127k - $249k
The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper). As...Work at officeLocal areaRemote workWorldwideFlexible hours- ...Software Development Engineer We're creating a platform that will change the way organizations measure their software development efforts... ...teams can work and the tools they use Location: on-site in Boston We believe that it takes a diverse team to build the...Principal
- hackajob is collaborating with Verisk to connect them with exceptional professionals for this role. Description Come be part of something new and exciting at a well-established analytical company. Help build scalable solutions for an industry leading catastrophe...Principal
$51.9 per hour
...Company: Allegheny Health Network Job Title: Site Reliability Engineering – Clinical & Facility Services General Overview This role ensures the reliability, availability, and performance of critical healthcare IT systems in the Environment of Care (EOC), supporting patients...Local area$160k - $225k
...tens of thousands of users across hundreds of organizations globally. About the Role Manifold is looking for a Staff Site Reliability Engineer (SRE) to work at the intersection of AI, data infrastructure, and life sciences. In this high-impact role, you will help...$121.5k - $306.4k
...infrastructure and service and provides input on best practices for reliability and functionality. Establishes direction to ensure accurate... ...with new technology, executing improvements, building site reliability knowledge, and providing clear data. #LI-ES2 Responsibilities...Temporary workFlexible hours$127k - $249k
...Central time zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support, maintain and grow the Atlas... ...crucial workloads. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong infrastructure background. This...Local areaRemote workFlexible hours- LogRocket is seeking a senior backend/full-stack engineer to design and scale systems that process millions of events. You’ll help build highly available services, improve real-time search, and mentor junior engineers in a fast-growing SaaS environment. You should thrive...Principal
- Vestmark, located in Boston, is seeking a Principal Engineer to architect AI-driven solutions for wealth management technology. You will take the lead in delivering innovative, agentic systems that enhance business processes. With a focus on collaboration across teams,...Principal
$95k - $171k
A leading cloud computing company seeks a Site Reliability Engineer II to join their Inference Cloud Team. The role involves building dashboards, writing automation in Python or Go, and collaborating with engineering teams to ensure AI infrastructure reliability. Candidates...Flexible hours- A leading fintech company is seeking an experienced technical support specialist in Boston, MA. You will troubleshoot production system issues and provide automation support for financial applications. The role requires at least 5 years of experience in Java and database...Remote jobWork at office
- ...work and education) Education Desired: Bachelor of Computer Engineering Travel Percentage: 0% We are FIS. Our technology powers the... ...python, etc. A mindset/desire to improve application systems reliability and automate manual support tasks, to facilitate continuous...Full timeWork at officeRemote workWork from homeFlexible hours
$75.2k - $95.3k
About the Team & Role We are looking for a highly motivated and high‑potential entry‑level Site Reliability Engineer (SRE) to join our team and help drive meaningful business impact while launching your career in reliability engineering. This is a really exciting time...Work experience placementFlexible hours- Anduril Industries is seeking a principal-level security engineer to lead platform security across our autonomous and robotic platforms. You will own cryptographic infrastructure, hardware root-of-trust, and secure communication in contested environments, shaping security...Principal
- Site Reliability Engineer at the organization. Key technologies: Kubernetes, Prometheus, Grafana. Key Responsibilities Define and track SLOs, SLIs and error budgets Design and implement observability stacks (metrics, logging, tracing) Automate toil and improve system...
$75.2k - $95.3k
WEX, Inc. is seeking a highly motivated entry-level Site Reliability Engineer (SRE) in Boston, MA. This role is integral to our SRE transformation, contributing to the performance and reliability of our systems. As an SRE, you'll monitor system reliability, manage incidents...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!
- principal cloud engineer Boston, MA
- senior principal engineer Boston, MA
- assistant chief engineer Boston, MA
- principal infrastructure engineer Boston, MA
- general engineer Boston, MA
- director of electrical engineering Boston, MA
- principal engineer Boston, MA
- director of product engineering Boston, MA
- director data engineering Boston, MA
- data center chief engineer Boston, MA


