Staff Site Reliability Engineer
$131k - $164kDiligent
Staff Site Reliability Engineer
New York, New York, United States
Position Overview
We are seeking a highly skilled Staff Site Reliability Engineer with deep technical expertise across VMware, Linux, and automation frameworks, to join our global Infrastructure & Operations team. This role is a hands-on senior engineering position responsible for designing, maintaining, and optimizing our private cloud environments, which underpin mission-critical SaaS products.
The ideal candidate will have extensive experience operating in enterprise datacenter environments, a strong foundation in Microsoft Active Directory and Windows Server, and a proven ability to build—not just run—automation workflows that improve reliability, scalability, and efficiency.
You will work closely with other engineering teams (Network, Security, SRE, and DevOps) to ensure the stability and performance of our global platform and drive continuous improvement through automation and infrastructure modernization.
Key Responsibilities
- Architect, deploy, and maintain VMware-based private cloud infrastructure across multiple global datacenters.
- Automate infrastructure operations using PowerCLI, Ansible, Python, or other automation tools to streamline provisioning, configuration, and compliance tasks.
- Administer and optimize Linux (RHEL/CentOS/Ubuntu) and Windows Server operating systems supporting enterprise workloads.
- Integrate and maintain Active Directory for authentication, policy, and service account management across hybrid environments.
- Collaborate with network and security teams to manage and troubleshoot firewall rules, VPNs, load balancers, and routing dependencies.
- Support and maintain F5 BIG-IP and AVI (NSX Advanced Load Balancer) for application delivery and traffic management.
- Ensure system availability, performance, and security to meet SLAs and compliance requirements (CIS, NIST, ISO).
- Participate in on-call rotations and change control processes for infrastructure incidents and maintenance.
- Document architecture, procedures, and automations for cross-team knowledge sharing and operational continuity.
- Mentor junior engineers and contribute to long-term technical strategy for infrastructure automation and modernization.
- 10+ years of experience in systems engineering or infrastructure roles, with at least 5 years at a senior or staff level.
- Expert proficiency in VMware vSphere (6.x/7.x/8.x) – including ESXi, vCenter, DRS, HA, vMotion, and distributed switches.
- Advanced Linux administration skills (RHEL/CentOS/Ubuntu), including performance tuning, system hardening, and troubleshooting.
- Strong understanding of Windows Server and Active Directory, including Group Policy, DNS, and authentication integrations.
- Demonstrated experience building automation frameworks using PowerShell, PowerCLI, Ansible, Python, or similar tools.
- Hands-on experience in enterprise datacenter environments, including storage (SAN/NAS), networking, and monitoring systems.
- Solid understanding of TCP/IP networking, email infrastructure, DNS, VPNs, and firewall concepts.
- Experience working with F5 BIG-IP, AVI/NSX Advanced Load Balancer, or similar ADC platforms.
- Familiarity with configuration management, version control (Git), and CI/CD pipelines.
- Strong problem-solving and analytical skills with a focus on reliability and scalability.
- Knowledge of Pure Storage, Cisco UCS, or similar datacenter technologies.
- Experience with Terraform, Jenkins, or Azure DevOps for infrastructure automation.
- Exposure to security hardening and compliance frameworks (CIS, NIST, ISO 27001).
- Experience in SaaS or highly available enterprise environments.
- Creativity is ingrained in our culture. We are innovative collaborators by nature. We thrive in exploring how things can be differently both in our internal processes and to help our clients
- We care about our people. Diligent offers a flexible work environment, global days of service, comprehensive health benefits, meeting free days, generous time off policy and wellness programs to name a few
- We have teams all over the world. We may be headquartered in New York City, but we have office hubs in Washington D.C., Vancouver, London, Galway, Budapest, Munich, Bengaluru, Singapore, and Sydney.
- Diversity is important to us. Growing, maintaining and promoting a diverse team is a top priority for us. We foster and encourage diversity through our Employee Resource Groups and provide access to resources and education to support the education of our team, facilitate dialogue, and foster understanding.
Required Experience/Skills
Nice to Have
U.S pay range
$131,000 - $164,000 USD
About Us
Diligent is the AI leader in governance, risk and compliance (GRC) SaaS solutions, helping more than 1 million users and 700,000 board members to clarify risk and elevate governance. The Diligent One Platform gives practitioners, the C-Suite and the board a consolidated view of their entire GRC practice so they can more effectively manage risk, build greater resilience and make better decisions, faster.
At Diligent, we're building the future with people who think boldly and move fast. Whether you're designing systems that leverage large language models or part of a team reimaging workflows with AI, you'll help us unlock entirely new ways of working and thinking. Curiosity is in our DNA, we look for individuals willing to ask the big questions and experiment fearlessly - those who embrace change not as a challenge, but as an opportunity. The future belongs to those who keep learning, and we are building it together. At Diligent, you're not just building the future - you're an agent of positive change, joining a global community on a mission to make an impact.
What Diligent Offers You
Diligent created the modern governance movement. Our world-changing idea is to empower leaders with the technology, insights and connections they need to drive greater impact and accountability – to lead with purpose. Our employees are passionate, smart, and creative people who not only want to help build the software company of the future, but who want to make the world a more sustainable, equitable and better place.
Headquartered in New York, Diligent has offices in Washington D.C., London, Galway, Budapest, Vancouver, Bengaluru, Munich, Singapore and Sydney. To foster strong collaboration and connection, this role will follow a hybrid work model. If you are within a commuting distance to one of our Diligent office locations, you will be expected to work onsite at least 50% of the time. We believe that in-person engagement helps drive innovation, teamwork, and a strong sense of community.
We are a drug free workplace. Diligent is proud to be an equal opportunity employer. We do not discriminate based on race, color, religious creed, sex, national origin, ancestry, citizenship status, pregnancy, childbirth, physical disability, mental disability, age, military status, protected veteran status, marital status, registered domestic partner or civil union status, gender (including sex stereotyping and gender identity or expression), medical condition (including, but not limited to, cancer related or HIV/AIDS related), genetic information, or sexual orientation in accordance with applicable federal, state and local laws. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Diligent's EEO Policy and Know Your Rights. We are committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or an accommodation due to a disability, you may contact us at View email address on click.appcast.io.
To all recruitment agencies: Diligent does not accept unsolicited agency resumes. Please do not forward resumes to our jobs alias, Diligent employees or any other organization location. Diligent is not responsible for any fees related to unsolicited resumes.
- ...ears, and hands on the ground at a government customer site, ensuring the reliability and performance of Twenty's mission-critical platform running... ...of deep technical ownership and customer-facing engineering: you'll define how we measure reliability, lead incident...SuggestedFull timeWork at officeRemote workFlexible hours
- ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III - DevOps Engineer at JPMorgan Chase within the Commercial and Investment Bank, you will solve complex and broad business...Suggested
$104.9k - $174.7k
...scale, 24x7, distributed and fault-tolerant systems within agreed reliability objectives, whilst enabling the fast flow of feature and... ...strong automation skills. About team; This diverse team of Engineers in assisting multiple product teams as we continue to innovate...SuggestedLocal areaImmediate startWorldwide$166k - $220k
...Site Reliability Engineer (SRE) Anduril Industries is a defense technology company with a mission to transform U.S. and allied military capabilities with advanced technology. By bringing the expertise, technology, and business model of the 21st century's most innovative...SuggestedFull timeWork experience placementImmediate start$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...SuggestedTemporary workImmediate startFlexible hoursShift work- ...Qualifications: ~10+ years of overall experience in IT including, with hands-on Development and Systems engineering background ~3-5 years of experience in a Site Reliability Engineering role ~ Experience with Enterprise Cloud transformation efforts ~ Experience with...Temporary workImmediate start
$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust...Work experience placementWork at office$165k - $241.4k
...can only be performed by a U.S. citizen on U.S. soil. From a reliability standpoint, this role involves evaluating the scalability,... ...coverage, and basic life insurance. Please see the Cisco careers site to discover more benefits and perks. Employees may be eligible...Permanent employmentFull timeTemporary workLocal areaFlexible hours$165k - $230k
...actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars. SR. SITE RELIABILITY ENGINEER (STARSHIELD) Starshield leverages SpaceX’s Starlink technology and launch capability to support national security efforts...Permanent employmentTemporary workImmediate startWeekend work$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$106.3k - $221.1k
...more. Join us to drive positive, lasting change that moves missions and the government forward! Job Description The Site Reliability Engineer will ensure the reliability, performance, and scalability of the Client System. The engineer will define and track Key...Live inWork at officeLocal area$125k - $135k
...Site Reliability Engineer Job number: 880 This is a remote position. Ad Hoc is a technology company that empowers organizations to deliver scalable, impactful digital services. Using modern, agile methods, our team creates products that meet people's needs...Remote workFlexible hours$153k - $185k
...Senior Site Reliability Engineer El Segundo, California, United States About Varda Low Earth orbit is open for business. Varda is accelerating... ...suited to accomplishing this goal, with leadership and staff comprised of veterans from SpaceX, Blue Origin, major...Permanent employmentFull timeImmediate startRelocation packageFlexible hoursWeekend work$131k - $227.13k
...Description: The 1LMX MES COE is seeking an engineer who will own infrastructure‑as‑code, cloud platform, and reliability for the Apriso environment on AWS. This role blends full‑stack development, DevOps, and Site Reliability Engineering (SRE) practices to deliver...Full timeTemporary workWork experience placementWork at officeRemote workRelocationFlexible hoursShift work3 days per week$112k - $179k
...system, network, software, and security solutions. About The Role Peraton is seeking a self-driven and resourceful Site Reliability Engineer to join our dynamic of Network and UC engineers in Washington, DC. This position combines software engineering and systems...Contract workWorldwideShift work- ...Site Reliability Engineer (SRE) Randstad is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our client in the Washington D.C. area, focusing on optimizing the availability, performance, and scalability of critical production services. The ideal...
- ...Site Reliability Engineer (SRE) Dexian is seeking a savvy Site Reliability Engineer (SRE) who will play a key role in building a sustainable... ...resolve technical issues and escalations with other technical staff as the need arises. Work to automate the detection and...Work experience placement
$114.6k - $190.2k
...with MANTECH! ***This is for a future opportunity*** MANTECHseeks motivated, career, and customer-oriented Site Reliability Engineer (SRE) for a new initiative. This effort supports the rapid design, deployment, operation, and sustainment of enterprise-...Hourly payContract workTemporary workWork experience placementWork at officeLocal areaRemote work$51.9 per hour
...This job is responsible for the reliability, availability, and... ...efficiency. This role blends software engineering, clinical engineering, and... ...cross-functionally with AHN site leaders and teams to navigate... ..., performance management and staff productivity.Plan, organize,...For contractorsLocal area- ...Associate Software Engineer The purpose of the Associate Software Engineer position is to provide technical assistance directly to... ...finthrive. Award-winning Culture of Customer-centricity and Reliability At FinThrive we're proud of our agile and committed...Contract workWork experience placementCasual workLocal area
$84.9k - $209.5k
...unencumbered and will need your contribution to make it a special engineering center with the focus on excellence. Health Data... ...critical issues that have not yet been documented as SOPs for Level1 staff. You will usually get called in during major incidents as an SME...Temporary workImmediate startFlexible hours$178k - $213k
...Splunk Ventures, and Vista Credit Partners of Vista Equity Partners 2022 Cybersecurity Excellence Award for MDR Manager, Site Reliability Engineering Reports to: VP, Product Engineering Location: While proximity to Tampa is preferred to support hybrid schedule in Tampa...Permanent employmentWork experience placementWork at officeRemote workWork from homeHome officeFlexible hours- ...Lead Site Reliability Engineer Bridge Defense is redefining how modern defense technology is delivered. Based in Washington, D.C., we are built for the dynamic mission environment facing the Department of Defense, the Intelligence Community, and federal law enforcement...Contract workRemote workRelocation
- ...Principal Site Reliability Engineer The Principal Site Reliability Engineer will be a critical technical leader responsible for driving the operational excellence, resilience, and security of our core systems for a key Randstad client in the Washington D.C. area. This...
$220k - $250k
...Staff Site Reliability Engineer Yugabyte is the company behind YugabyteDB, the AI-ready, multi-modal, distributed PostgreSQL database for cloud-native apps. Trusted by industry leaders including Shopify, Paramount+, GM, Kroger, Fiserv, and NPCI, YugabyteDB has been...H1bLocal areaWorldwideVisa sponsorship- Overview MANTECH seeks motivated, career, and customer-oriented Site Reliability Engineer (SRE) for a new initiative. This effort supports the rapid design, deployment, operation, and sustainment of enterprise-scale AI, data, and mission platform capabilities across cloud...Work at office
- Role Overview We are seeking a high-caliber Site Reliability Engineer (SRE) to join our Forward Engineering team. You will be the guardian of our production ecosystems, ensuring that our complex, data-driven AI platforms remain resilient, scalable, and highly performant...Local area
$100.2k - $203.4k
As a Site Reliability Engineer, you will play a pivotal role in advancing operational AI adoption within a cutting‑edge Hub‑and‑Spoke architecture. Your primary focus will be on ensuring the reliability, scalability, and continuous monitoring of enterprise AI systems that...Live inLocal area$149.4k - $202k
Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is a strategic force driving digital transformation. We treat operations as a software engineering challenge, focusing on...Remote work$116.9k - $234.1k
Site Reliability Engineer The Site Reliability Engineer will ensure the reliability, performance, and scalability of the Client System. The engineer will define and track Key Performance Indicators and Service Level Objectives, identify and resolve performance bottlenecks...Local area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Site Reliability Engineer. Be the first to apply!
- assistant engineer Washington DC
- senior staff systems engineer Washington DC
- engineering aide Washington DC
- senior staff engineer Washington DC
- staff engineer Washington DC
- technology administrator Washington DC
- software engineer staff Washington DC
- site reliability engineer sre Washington DC
- site reliability engineer Washington DC
- site reliability engineer remote Washington DC


