Manager, Site Reliability Engineering
$102k - $234.6kOracle
Job Description Supports team members in designing and architecting infrastructure and service for reliability and functionality. Provides day-to-day direction to help forecast demands and ensure systems have adequate resources. Assists with collaboration between team members and the software development team to create reliable, scalable infrastructures. Advises on data collection, optimizing operations and infrastructure reliability. Aids in incident response activities to ensure service reliability. Reviews health and performance reports. Trains team members to identify automation. Trains team members to communicate and understand the impact of changes. Serves as an escalation point for incidents and reviews documentation. Enables team members to experiment with new technology, execute improvements, build site reliability knowledge, and provide clear data. Responsibilities Key Responsibilities Capacity Ingestion andManagement: - Supportsteam members designing and architecting infrastructure and/or service accordingto terms for reliability and functionality. - Supervisesimmediate team members and provides day-to-day direction to help forecastdemands for infrastructure and respond to capacity needs, ensuring systems havesufficient resources to handle current and future workloads. - Assiststeam members in collaborating with the software development team to developinfrastructures and features that are reliable and scalable according todeployment requirements. - Developsteam members' ability to identify opportunities for and drive prototyping(e.g., testing new applications or infrastructures, assisting in onboarding). Incident and ServiceLifecycle Management: - Advisesteam members on performing data collection, triage, technical analysis, andredirection and recommends methods to maintain and optimize operations andinfrastructure reliability. - Providessupport to team members monitoring services, ensuring they maintain up-to-dateknowledge of performance and document their condition. - Leveragesworking knowledge to aid team members in performing incident response, rootcause analyses, and/or maintenance on assigned services (e.g., softwareinstalls, version upgrades, security updates, backup and recovery). - Reviewshealth and performance reporting and recommends appropriate actions based ontrends in data. - Helpsteam members follow procedures to perform provisioning to supportinfrastructure, applications, and services. - Educatesteam members on performing decommissioning (e.g., shutting down servers,removing data from databases) to remove objects that are no longer needed. Automation: - Trainsteam to identify opportunities for automation and assesses potential benefits. - Reviewsautomation tools or scripts developed by team members and provides feedback. - Coachesteam members to conduct testing to ensure automation performs the taskcorrectly and produces expected results. Technical Communication andGuidance: - Trainsand enables team members to communicate the scale, capacity, security,performance attributes, and requirements of services and technology within andsometimes beyond immediate team. - Reviewsand helps team members understand the potential impact of infrastructure,feature, and tool changes, considering their impact on team operations. Troubleshooting andResolution: - Servesas the team's escalation point for incidents and other moderately complexissues arising within Oracle services. - Recommendsmethods to resolve technical issues spanning various services, coaching teammembers to investigate and debug products in order to reach SLOs (service levelobjectives). - Reviewsdocumentation for accuracy and trains team members to perform root causeanalyses according to standard reporting methods. - Coachesteam members to independently perform post-mortem procedures to preventincident reoccurrence. Innovation and Improvement: - Enablesteam members to experiment with new tools and technologies and helps assesstheir potential impact on and improve infrastructure performance andreliability, ensuring adherence to security standards. - Supervisesthe execution of improvements for performance bottlenecks and deployments toensure efficient resource usage, speed, and scalability. - Developsteam members' knowledge of site reliability trends and shares new informationto help them build, test, deploy and run services. - Trainsteam members to perform analyses and provide clear data on production tocontribute to business development decisions (e.g., design changes). Core Responsibilities Planning & Execution: - Createsand owns the execution plan for the team's work and multiple projects orinitiatives, monitoring timelines and budgets when applicable to ensureprojects are completed on time and in adherence with requirements. Delegateswork across the team and helps them prioritize their work. Identifies resourceneeds and adapts plans based on changing priorities and business needs. Collaboration &Partnership: - Strengthenscollaborative partnerships across teams to align expectations and sharedobjectives. Guides team members to build relationships with business leaders,stakeholders, and/or customers to ensure effective collaboration. Practicesactive listening and asks insightful questions to promote an inclusive culture. Problem Solving: - Leadsteam to identify and address moderately complex operational and/or technicalissues in accordance with standard practices, providing guidance asappropriate. Directs team to analyze data and/or information from multiplesources to troubleshoot moderately complex errors. Continuous Learning: - Activelyseeks learning opportunities for self and team to enhance knowledge and skillsin key areas and remain current with industry advancements. Uses feedback andtraining to elevate personal and team skills, modeling a commitment tolearning. Identifies skill gaps within the team and supports team members inlearning opportunities by providing resources and fostering an environment thatencourages knowledge-sharing. Continuous Improvement: - Identifiesand recommends improvements as well as encourages team to share ideas toincrease the efficiency and effectiveness of processes, protocols, andworkflows within a team. Reviews and provides feedback on team members' ideasand seeks input from others on alternative approaches and methods for improvingwork. Performance and Development: - Coachesand provides direction to team members in alignment with performance managementprocesses, guidelines, and expectations. Facilitates goal setting discussionswith team members, ensuring individual goals are aligned with broader teamgoals. Identifies development opportunities to help team members improve performance.Contributes to the talent acquisition pipeline by leading candidate interviews,assessing promotion eligibility, and managing talent resources. Qualifications Disclaimer: Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements. Range and benefit information provided in this posting are specific to the stated locations only US: Hiring Range in USD from: $102,000 to $234,600 per annum. May be eligible for bonus, equity, and compensation deferral. Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.
Candidates are typically placed into the range based on the preceding factors as well as internal peer equity. Oracle US offers a comprehensive benefits package which includes the following:
1. Medical, dental, and vision insurance, including expert medical opinion
2. Short term disability and long term disability
3. Life insurance and AD&D
4. Supplemental life insurance (Employee/Spouse/Child)
5. Health care and dependent care Flexible Spending Accounts
6. Pre-tax commuter and parking benefits
7. 401(k) Savings and Investment Plan with company match
8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
9. 11 paid holidays
10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
11. Paid parental leave
12. Adoption assistance
13. Employee Stock Purchase Plan
14. Financial planning and group legal
15. Voluntary benefits including auto, homeowner and pet insurance The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.
Career Level - M2 About Us Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives. True innovation starts when everyone is empowered to contribute. That's why we're committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs. We're committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing View email address on click.appcast.io or by calling View phone number on click.appcast.io in the United States. Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans' status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
Candidates are typically placed into the range based on the preceding factors as well as internal peer equity. Oracle US offers a comprehensive benefits package which includes the following:
1. Medical, dental, and vision insurance, including expert medical opinion
2. Short term disability and long term disability
3. Life insurance and AD&D
4. Supplemental life insurance (Employee/Spouse/Child)
5. Health care and dependent care Flexible Spending Accounts
6. Pre-tax commuter and parking benefits
7. 401(k) Savings and Investment Plan with company match
8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
9. 11 paid holidays
10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
11. Paid parental leave
12. Adoption assistance
13. Employee Stock Purchase Plan
14. Financial planning and group legal
15. Voluntary benefits including auto, homeowner and pet insurance The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.
Career Level - M2 About Us Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives. True innovation starts when everyone is empowered to contribute. That's why we're committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs. We're committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing View email address on click.appcast.io or by calling View phone number on click.appcast.io in the United States. Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans' status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Manager, Site Reliability Engineering in Reston, VA vacancy
$136.2k - $214.01k
...Exceptional in execution and impact The Role As a Senior Site Reliability Engineer at Proofpoint you will develop a deep understanding of the... ..., Ansible, Cloudformation, etc. • Experience automating management and operational tasks using Node.Js, Python, or scripting...SuggestedFull timeFlexible hours$81.1k - $187k
...partner with customer support, service owners, and engineering teams around the globe to ensure high-quality... ...Level - IC3Escalation points for junior site reliability engineers during complex or high-impact incidents.Manage and execute complex manual Change Management tickets...SuggestedTemporary workMonday to FridayFlexible hoursShift workNight shift- ...Site Reliability Engineer (SRE) Reston, VA Site Reliability Engineer (SRE) Position: Site Reliability Engineer (SRE) Work Authorization... ..., Postman) ~ Hands-On experience with a source code management system like GIT or SVN including pull, push, branch,...SuggestedContract work
$87.1k - $157.45k
...throughout the entire USG arsenal. Our team of hackers, engineers, makers, and shakers brings deep experience across... ...to come in and help us build systems that stay reliable when things get complicated. We need a Site Reliability Engineer who has experience building, deploying...SuggestedLocal areaImmediate startWork from homeFlexible hours- ...Its centralized 1EXIGER.AI platform allows organizations to manage their entire operating network, from parts to suppliers,... ...Gartner Magic Quadrant for Supplier Risk Management. Site Reliability Engineer Location: U.S. (Hybrid) This role requires U.S. citizenship...SuggestedWork at officeWork from homeFlexible hours
$62k - $141k
...Site Reliability Engineer Chantilly, VA Top Secret/SCI Polygraph Career Level not specified $62,000 - $141,000 Job Description The Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether...Full timeContract workPart timeWork at officeLocal areaRemote work- ...Site Reliability Engineer Description: We are looking for a dynamic Site Reliability Engineer (SRE) with a Top Secret clearance to join our team! The Site Reliability Engineer (SRE) will manage, monitor, and optimize clusters on Kubernetes. Together, we're accelerating...
$120k - $150k
...delivers a comprehensive event marketing and management platform for marketers and event... ...In This Role, You Will: Own reliability across a complex, global product portfolio... ...Cursor, or AI coding assistants as an engineering accelerant - not occasional dabbling...Work at officeWorldwide$128.5k - $190k
...Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS platform, Medallia Experience Cloud... ...Bring your whole self. The Role and Team The Site Reliability Engineering organization at Medallia brings together the infrastructure...Temporary workWork experience placementLocal area$90k - $130k
...challenges. Credence has an immediate opening for a Site Reliability SME who has hands-on experience working as a Cloud Operations Engineer with experience in IT operations to join our expanding Cloud Managed Services Provider team. This position is based in...Temporary workWork experience placementImmediate startWorldwide$158.5k - $230k
...Medallia is the pioneer and market leader in Experience Management. Our award-winning SaaS platform, Medallia Experience Cloud... ...We are growing our GovCloud team and looking for a Staff Site Reliability Engineer to help scale how we operate Medallia's US public-sector cloud...Permanent employmentTemporary workWork experience placementWork at officeLocal areaRemote work3 days per week$148.5k - $223.9k
...duplicating efforts. Job Category Software Engineering Job Details About Salesforce... .... The Experience Join our Site Reliability Engineering (SRE) team, where you'll work... ...demonstrated experience with incident management and a solid understanding of IT Infrastructure...Full time$114.4k - $125.4k
...Job Category Software Engineering Job Details About Salesforce... ...about ensuring the reliability and performance of mission-critical... ...Salesforce is seeking a talented Site Reliability Engineer to join... ...stability through incident management, automated alerting,...Full timeLocal areaShift workNight shift- ...Job Description Job Description Systems Integrator / Business Process Manager Department: Government Customer- Herndon Location: Herndon, VA TENICA is looking for a Systems Integrator / Business Process Manager. Candidate must have a TOP SECRET/SCI...Work at office
- ...Required Position Summary GC2IT is seeking a Senior Software Engineer - Analytics Services Lead to own the design, delivery,... ...reporting capabilities. · Use telemetry and APM data to improve reliability, identify bottlenecks, support root-cause analysis, and lead...Full timeContract workWork at officeRemote work
- ...is seeking a Senior Software Engineer - Administrative Services Lead... ...interfaces, user/role management, and peripheral/device management... ...personnel, support/field teams, site/regional managers, and role-... ...middleware providers. · Ensure reliable installation, upgrade,...Full timeWork at officeRemote work
$146k - $194k
...Anduril as a lead provider of specialized engineering and products for Intelligence... ...security requirements.ABOUT THE JOBAs a Site Reliability Engineer, your primary mission is to ensure... ...and supporting edge. Patch/update management and remote management tooling.DevOps improvements...Full timeWork experience placementImmediate startRemote work$119.8k - $234.7k
...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole... ...EngineeringDiscipline: Site Reliability EngineeringCompany:... ...for a Senior Site Reliability Engineer (SRE) to join the Azure Silver... ...performance, efficiency, change management, and incident response—to help...Ongoing contractLocal area3 days per week$150k - $180k
...About the Job We are seeking an experienced Senior Site Reliability Engineer to help design, build, operate, and scale the mission- and... ...Collaborate closely with cross-functional teams, product managers, and stakeholders to align on technical strategy and provide...Permanent employmentWork at officeLocal areaRemote workWorldwideFlexible hours$109.18k - $163.77k
...company, we have offices in nine countries and can insert advertisements around the world.Job SummaryThe Site Reliability Engineering team is responsible for managing the critical infrastructure that powers FreeWheel's Streaming Hub platform. Streaming Hub is a large-...Full time$91.4k - $187k
Work with Site Reliability Engineering (SRE) team on the shared full stack ownership of a collection of services and/or technology areas. Understand... ...motivated individuals that have: * Prior SRE experience managing production cloud services.* Prior experience in releasing...Temporary workFlexible hoursShift workWeekend work- ...Job Description Job Description Site Reliability Engineer II DC Metro area, offices in Reston · Hybrid · 24/7 FedRAMP Operations · Rotational Shift · Initial Contract till March 27. KEY REQUIREMENT This role requires US citizenship and residence on US soil...Hourly payContract workFor contractorsShift workNight shiftWeekend work
$128.6k - $184.9k
...to keep their digital systems secure and reliable. Come help organizations be their best, while... ...team! We are looking for an experienced engineer to support the next generation of our... ...insurance. Please see the Cisco careers site to discover more benefits and perks. Employees...Permanent employmentFull timeTemporary workLocal areaRemote workFlexible hoursShift workNight shiftWeekend work$130k - $200k
Summary Position Title: Site Reliability Engineer Position ID: TA247 Location(s): On-site; Aurora, CO; Herndon, VA Application Deadline... ...critical systems by troubleshooting complex software issues, managing production incidents, and ensuring reliable operations...Full timeTemporary workLocal area- ...Type: W2 Compensation: $83.33 Overview We are seeking a Senior ServiceNow Developer with strong experience in Customer Service Management (CSM) to support the development and enhancement of enterprise ServiceNow applications. This role focuses on building and...Contract work
$121.5k - $264.1k
Capacity Ingestion and Management:- Supports team members designing... ...on practices and terms for reliability and functionality.- Supervises... ...maintaining knowledge of site reliability trends and sharing... ...years of experience in software engineering, infrastructure management,...Temporary workImmediate startFlexible hours- ...Project Manager/ Release Train EngineerLocation: Remote (Reston, VA)Duration: 6-12+ Months ContractTeam Overview:Providing release train engineering solutions across the organization.Managing AWS product engineering projects as well as infrastructure and cyber security...Work experience placementRemote work
$125k - $145k
...Job Title: Release Engineer Location: Remote Clearance: Public Trust OR Secret Type: Full-time, W2 About VivSoft: VivSoft... ...Monitoring rule configuration, and ICAM federation testing) Manage the multi-environment Flosum configuration (development,...Full timeRemote workFlexible hours$143.7k - $194.4k
...for an experienced Software Development Engineer to join our growing team that will be building... ...to processes or automation - Manage and grow innovative, production-quality... ...design or architecture (design patterns, reliability and scaling) of new and existing systems...InternshipFlexible hours- ...Service Products built on the ServiceNow platform. Will help lead ServiceNow delivery and work with key business units and process managers to build solutions and processes, supporting maintenance, continual service improvement and new capabilities on the ServiceNow...Long term contractTemporary workImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Manager, Site Reliability Engineering. Be the first to apply!
Related searches
- site reliability engineer Reston, VA
- IT site lead Reston, VA
- site safety Reston, VA
- site leader Reston, VA
- on-site clinical research associate (traveling/remote) Reston, VA
- junior website developer Reston, VA
- historic site Reston, VA
- construction site safety Reston, VA
- official site Reston, VA
- site reliability engineer sre



