Principal Site Reliability Engineer
$84.9k - $209.5kOracle
Solve complex problems related to infrastructure cloud services and build automation to prevent problem recurrence. Design, write, and deploy software to improve the availability, scalability, and efficiency of Oracle products and services. Design and develop designs, architectures, standards, and methods for large-scale distributed systems. Facilitate service capacity planning and demand forecasting, software performance analysis, and system tuningYou will provide cloud operations for Oracle National Security Realms. You’ll be part of a dynamic team with a broad knowledge of how Oracle’s cloud platform works. You’ll partner with customer support, service owners, and engineering teams around the globe to ensure high-quality service for customers.Note - this role is not a Monday to Friday core hours role – it will involve working a 24/7 shift rotation with on-call duties, including nights, weekends and public holidays.Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing or by calling in the United States.Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.Disclaimer:Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.Range and benefit information provided in this posting are specific to the stated locations onlyUS: Hiring Range in USD from: $84,900 to $209,500 per annum. May be eligible for bonus and equity.Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.Oracle US offers a comprehensive benefits package which includes the following:1. Medical, dental, and vision insurance, including expert medical opinion2. Short term disability and long term disability3. Life insurance and AD&D4. Supplemental life insurance (Employee/Spouse/Child)5. Health care and dependent care Flexible Spending Accounts6. Pre-tax commuter and parking benefits7. 401(k) Savings and Investment Plan with company match8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.9. 11 paid holidays10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.11. Paid parental leave12. Adoption assistance13. Employee Stock Purchase Plan14. Financial planning and group legal15. Voluntary benefits including auto, homeowner and pet insuranceThe role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.Career Level - IC4Escalation points for junior site reliability engineers during complex or high-impact incidents.Manage and execute complex manual Change Management tickets, by working closely with the service teams to ensure safety and minimal disruption to services.Support the on-boarding of new services and tools, ensuring they are operationally ready and properly integrated.Provide mentorship and training to SREs, helping build team capability and confidence.Create and maintain clear, useful documentation for operational processes and system support.Identify areas of manual work and drive automation to reduce toil and improve efficiency.Automate tasks to enable continuous delivery and ensure continuous availability with minimal human overheadRecognize unsafe or inefficient practices and work with teams to design safer, more effective solutions.Complete change requests to enable new functionality and maintain realm complianceEnsure timely resolution of incidents, service requests, and change requestsCollaborate with global service and engineering teamsDefine and drive change management, continuous integration, and deployment best practicesHelp create and maintain real-world production architectures, scalability, and system designUse a methodical approach to troubleshoot, large, complex, interconnected systemsWe also use…Linux and Unix operating systemsDocker, Kubernetes, and TerraformScripting languages such as Bash, shells, Perl, or PythonCitizenship/location requirements - i.e. US Citizenship, U.S. Citizenship and possess and maintain TS/SCI w/Poly security clearance, reside in Austin, TX or Reston, VATechnology related bachelor’s degree and/or equivalent work experienceA desire to learn and keep up with modern technologiesProficient with writing services/task automation in any modern development language (e.g. Python, Bash, Ruby, Perl, JavaScript, or Java)Familiarity with core protocols and OSI model (DNS, DHCP, TCP/IP)Deep knowledge of Linux or Unix OS internals and host-based networkingFamiliarity with configuration management solutions such as Chef, Puppet, etcExperience with devising, managing, and extending monitoring solutions for large scale environments.Knowledge of cloud computing conceptsExperience working in a mission-critical environment (Operations, Technical Support, NOC etc)Proficient with communication skills (writing, organization, learning exchange)Experience executing tasks under change management proceduresExperience resolving auto-cut and manual alarms following runbooksA focus on customer satisfactionSpecific experience working with deployment of AI infrastructure to include clustered GPUs, LLM deployment and maintenance, and understanding of model integration for customer solutionsCertifications in VMware or other hypervisors will stand outCertifications in Cisco (e.g. CCNA, CCNP) will stand outCertifications in CISSP or other security related will stand outCertifications in Oracle Databases or RAC will stand outPosting Date:
- ...You will provide cloud operations for Oracle National Security Realms. Responsibilities Escalation points for junior site reliability engineers during complex or high-impact incidents. Manage and execute complex manual Change Management tickets, by working...PrincipalTemporary workWork experience placementFlexible hoursNight shift
$136.2k - $214.01k
...outcomes Visionary in future focused problem-solving Exceptional in execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the various services and applications that come together to...SuggestedFull timeFlexible hours$150k - $180k
...what’s possible in remote sensing, you belong here at Umbra. About the Job We are seeking an experienced Senior Site Reliability Engineer to help design, build, operate, and scale the mission- and business-critical infrastructure that powers Umbra's systems....SuggestedPermanent employmentWork at officeLocal areaRemote workWorldwideFlexible hours- ...Site Reliability Engineer Location: Occasional onsite visits to Reston VA (Zip code 20190). Duration-1 year plus Interview process: The final interview is a mandatory, face-to-face interview in Reston VA. Zip code: 20190 Strong...SuggestedLong term contractTemporary workH1bImmediate startRelocation
- ...Site Reliability Engineer (SRE) Reston, VA Site Reliability Engineer (SRE) Position: Site Reliability Engineer (SRE) Work Authorization: All Work Authorizations Location: Reston, VA Contract: 24 months Description: Site Reliability Engineer (SRE) roles...SuggestedContract work
- ...architecting infrastructure and service for reliability and functionality. Provides day-to-day... ...technology, execute improvements, build site reliability knowledge, and provide clear... ...: 8 years of experience in software engineering, infrastructure management, or related field...Immediate start
$112.5k - $187.5k
...We Collect Your Privacy Choices Team Overview At TransUnion, this role will report to a DevOps Director. The Site Reliability Engineering team drives reliability strategy, elevates engineering standards, and owns some of the most complex and consequential work...Full timeWork experience placementWork at officeFlexible hours2 days per week$113k - $345k
...in an evolving world. Our mission-first software and data engineering platform modernizes data operations, utilizing advanced workflows... ...and refining them into robust, rigorously tested tools with a reliable release cycle. Driven by a customer-centric mindset and...PrincipalContract workImmediate start$175k - $185k
...NVR, Inc. is seeking an experienced Principal Software Engineer to work on site in Reston, VA. NVR's technology teams build and support the applications... .... Lead the design of scalable, secure, reliable, and maintainable software solutions that support long...Principal$111.8k - $221.8k
...more. Join us to drive positive, lasting change that moves missions and the government forward! AFS is seeking a Site Reliability Engineer to join our team and support our client on a full-time basis in Reston, VA or Annapolis Junction, MD. This role requires...Full timeLive inWork at officeLocal area- ## Principal Software EngineerApply: Herndon, VA: Full time: Posted Today: R24667Vantor... ...Senior Developer collaborates with ontology engineers, software engineers, database engineers... ..., and ensure application scalability, reliability, and security.* Produce and maintain...PrincipalFull time
$176k - $282k
...Systems Engineer, Principal Job Locations US-VA-Reston | US-DC-Washington Requisition ID 2026-164241 Position Category Engineering Clearance Top Secret/SCI w/Poly Responsibilities Peraton is seeking a Systems Engineer...PrincipalContract workWork experience placementShift work$118k - $177k
...Everforth ECS is seeking a Senior Site Reliability Engineer to work remotely . Everforth ECS is seeking talented professionals to join our successful and growing team in building the next-generation Continuous Diagnostics and Mitigation (CDM) Cyber data solution...Remote work$120k - $150k
...operates along the way. In This Role, You Will: Own reliability across a complex, global product portfolio. You'll be... ...Daily use of Claude Code, Cursor, or AI coding assistants as an engineering accelerant - not occasional dabbling Experience...Work at officeWorldwide- ...tech SME and PM Period of performance: Up to 2 years in duration MUST HAVES: Minimum of 8 years of experience as a Site Reliability Engineer with a strong understanding of SRE principles for highly scalable and reliable systems Possess a bachelor's degree...Local areaRelocation package3 days per week
$103.5k - $150k
...exceptional people to create extraordinary experiences together. Bring your whole self. The Role and Team The Site Reliability Engineering organization at Medallia brings together the infrastructure and applications that power a highly reliable global SaaS platform...Temporary workWork experience placementLocal area3 days per week$114.6k - $234.6k
...continues its rapid expansion, we are seeking a skilled Software Engineer to join our newly established Cloud Performance Organization.... ...SLOs, KPIs, telemetry, dashboards, and alerts to ensure reliability and performance . Design performance, load, fault-injection...PrincipalFull timeTemporary workFlexible hours$100k - $160k
...all bring to the team; and empower our employees to create innovative and trusted results. We are looking for a dynamic Site Reliability Engineer (SRE) with a Top Secret clearance to join our team! The Site Reliability Engineer (SRE) will manage, monitor, and optimize...Temporary work$62k - $141k
...Site Reliability Engineer Chantilly, VA Top Secret/SCI Polygraph Career Level not specified $62,000 - $141,000 Job Description The Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether...Full timeContract workPart timeWork at officeLocal areaRemote work- ...Site Reliability Engineer Mc Lean, VA Long Term Client's Enterprise Data Machine Learning (EDML) employs innovative minds like yourself to design and develop software-systems that can meet the demand of our ever-growing customer base. Like a...Immediate start
$87.1k - $157.45k
...throughout the entire USG arsenal. Our team of hackers, engineers, makers, and shakers brings deep experience across... ...to come in and help us build systems that stay reliable when things get complicated.We need a Site Reliability Engineer who has experience building, deploying...Full timeWork from homeFlexible hours- ...Detail Description: The AWS Site Reliability Engineer (SRE) is responsible for the operational health, availability, and performance of the AWS and Databricks environments built by the Platform Engineering team. You prepare and take ownership of "day two" operations...
- ..., to act first. Exiger is FedRAMP authorized and a 2x Leader in Gartner Magic Quadrant for Supplier Risk Management. Site Reliability Engineer Location: U.S. (Hybrid) This role requires U.S. citizenship and eligibility for a U.S. security clearance. Role...Work at officeWork from homeFlexible hours
- ...Cloud-Based Machine Learning System Architect Lead a team of engineers to design and implement cloud-based machine-learning system architectures. Develop scalable, generic, highly flexible cloud-based solutions to empower the machine-learning team to deliver algorithms...PrincipalRemote workFlexible hours
$158.5k - $230k
...extraordinary experiences together. Bring your whole self. The Role and Team We are growing our GovCloud team and looking for a Staff Site Reliability Engineer to help scale how we operate Medallia's US public-sector cloud platform. You will support federal agencies and other...Permanent employmentTemporary workWork experience placementWork at officeLocal areaRemote work3 days per week$145k - $160k
...We are seeking a specialized Observability & Infrastructure Engineer to lead the monitoring, telemetry, and platform-as-code initiatives critical to our multi-region disaster recovery roadmap. You will architect and implement robust observability pipelines, ensure deep...Temporary workRemote workFlexible hours- ...Required U.S. Citizenship / No clearance needed / 100% remote within the US Staff Site Reliability Engineer / Cloud SME Location: 100% remote in the continental US Type: Long-term contract (3+ years) Our client, a premier national healthcare provider, is currently...Long term contractRemote work
- ...MANTECH seeks a motivated, career-driven Principal AWS Cloud Engineer to serve as the primary AWS technical authority on an enterprise cloud team in Chantilly, VA The Principal AWS Cloud Engineer anchors the design, deployment, and security management of our...PrincipalWork at office
$191k - $253k
...ABOUT THE TEAM Anduril Intelligence Systems (AIS) is a lead provider of highly specialized engineering products for Intelligence Community (IC) customers. We work within the IC to understand their requirements and shape concepts of operation. We design, develop, and...Full timeWork experience placementImmediate start- ...TITLE: Sr. Software Engineer LOCATION: Herndon, VA CLEARANCE REQUIRED: Active TS/SCI with CI Polygraph EMPLOYMENT TYPE: Full-Time, Onsite POSITION SUMMARY Modern Government Solutions (MGS) is seeking a Sr. Software Engineer to design, develop, and...Full timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Principal Site Reliability Engineer. Be the first to apply!
- principal developer Reston, VA
- data center chief engineer Reston, VA
- engineering director Reston, VA
- hotel chief engineer Reston, VA
- principal engineer Reston, VA
- chief engineer Reston, VA
- director software engineering Reston, VA
- general engineer Reston, VA
- site reliability engineer Reston, VA
- principal Reston, VA



