Senior Site Reliability Engineer
2T Consulting
We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.
The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.
Required Technical Skills
- Strong understanding of Site Reliability Engineering principles and operational excellence.
- Experience with infrastructure reliability, service availability, resiliency, and performance optimization.
- Storage Space Direct and failover clustering technical expertise. (Storage Spaces Direct enables you to build highly available, software-defined storage by pooling local disks (SSDs, NVMe drives, and HDDs) across multiple Windows Server nodes in a cluster. Instead of relying on an external SAN, S2D uses the servers' local storage to create a resilient shared storage pool)
- Experience managing production-critical infrastructure environments with high availability requirements.
- Experience with incident management, problem management, RCA, and continuous operational improvement.
- Knowledge of monitoring, observability, alerting, and performance management.
Microsoft Hyper-V (Core Expertise)
- Deep hands-on expertise in Microsoft Hyper-V architecture, deployment, administration, troubleshooting, and optimization.
- Extensive experience in operating enterprise private cloud environments on Hyper-V.
- Strong experience supporting enterprise-scale VDI deployments on Hyper-V.
- Hyper-V Failover Clustering and high-availability architecture.
- Storage integration including SAN, NAS, Storage Spaces Direct (S2D), Cluster Shared Volumes (CSV), and storage optimization.
- Networking within Hyper-V environments including virtual switches, VLANs, NIC Teaming, QoS, and network performance tuning.
- System Center Virtual Machine Manager (SCVMM).
Automation & Platform Engineering
- Strong PowerShell scripting and automation experience.
- Experience automating infrastructure deployment, operational tasks, health checks, and reporting.
- Familiarity with Infrastructure as Code concepts and configuration management.
- Experience developing reusable operational tooling to improve reliability and reduce manual effort.
Preferred Skills
- Windows Server 2016/2019/2022 administration.
- Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
- Exposure to hybrid cloud and private cloud platforms.
- Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
- Experience supporting enterprise VDI environments.
- Understanding of ITIL Incident, Problem, Change, and Release Management.
- Experience working in regulated industries such as Banking or Financial Services.
Experience & Qualifications
- 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
- Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
- Proven experience implementing automation to reduce operational overhead and improve service reliability.
- Experience supporting enterprise private cloud and VDI environments.
- Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
- Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
- Experience in Banking or Financial Services environments is advantageous.
Key Responsibilities
- Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
- Optimize, and support highly available VDI environments on Hyper-V.
- Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
- Disaster recovery, backup, patch management, and business continuity strategies.
- Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
- Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
- Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
- Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
- Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
- Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
- Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
- Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
- Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
- Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.
- Jack Henry & Associates, Inc. is seeking a Senior Site Reliability Engineer to drive modernization across a large-scale hybrid cloud footprint, with emphasis on re-architecting on-prem workloads to Google Cloud Platform. The role involves implementing SRE practices, IaC...Senior
$140k - $150k
WORK OPTION: Remote_________________The NBA is hiring a Senior Site Reliability Engineer (SRE) - Messaging & Collaboration to ensure the availability, performance, and reliability of enterprise messaging and collaboration platforms, including Microsoft Exchange Online...SeniorFull timeTemporary workLocal areaRemote workWeekend work$120k - $175k
...level of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's...SeniorFull timeRemote workWork visaFlexible hours- ...responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include... ...continuous improvement. Position Summary: The Senior Azure Site Reliability Engineer acts as an advanced senior...SeniorWork at officeShift workDay shift
$185k - $227k
...professionals. If the opportunity to build your career is compelling, read on for more details. ROLE AND RESPONSIBILITIES: A Senior Site Reliability Engineer (SRE) is expected to own the operational stability and performance ofJuul’s hybrid cloud infrastructure (Nutanix, AWS/...SeniorRemote work- ...exceptional professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of... ...and position yourself among the top echelon in site reliability. As a Sr Lead Site Reliability Engineer at JPMorgan Chase within...Senior
- Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability.As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Chief Data & Analytics...SeniorWork at office
- UiPath, Inc. is seeking a Senior Software Engineer for our Site Reliability Engineering organization. You will design, build, and operate SRE platform systems, leveraging AI to improve reliability and performance across critical services. You will work on live-site monitoring...Senior
- ...Overview We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise...
- ...Goldman Sachs is seeking a Vice President in Compliance Engineering SRE for Dallas. The role combines software and systems engineering... ...monitoring, and collaborate with cross-functional teams to deliver reliable, compliant platforms for regulatory risk management. #J-18808-...
$155k - $175k
...Next! Summary We are seeking a highly skilled and experienced Site Reliability Manager to join our team to ensure the reliability,... ...performance of our systems and services. You will lead a team of engineers focusing on three core pillars: Application Reliability, DevSecOps...Work experience placementH1bWork at officeLocal area$134.6k - $210.1k
Position SummaryThe Senior Principal Engineer, DevOps plays an integral role in implementing and executing cloud practices for build management, product release and operation processes. The role is responsible for managing and automating the build and deployment process...SeniorTemporary workWork at officeImmediate startFlexible hoursNight shift$170k - $240k
...thrive.The RoleAre you a Platform Engineer looking to make a big impact... ...looking for a self-motivated Senior Platform Engineer to join our... ...improve the security and reliability of our applications running on... ...years experience in DevOps, Site Reliability Engineering, Platform...SeniorRemote work$117.5k - $150.4k
Position SummaryThe Senior Platform Engineer - Platform Services & Data Center is responsible for the engineering, implementation, modernization... ...infrastructure, including Domain Controllers, replication, Sites and Services, FSMO roles, Global Catalogs, and Group Policy...SeniorTemporary workWork at officeImmediate startRemote workFlexible hoursNight shift- ...Vercel, the agentic infrastructure company, is hiring a senior software engineer to shape the next generation of AI-enabled products. You will... ...across UI, APIs, data stores, and production systems to deliver reliable, scalable experiences for millions of developers. We value...SeniorWork from homeFlexible hours
- ...Corporation, a leading US-based global power company, is seeking a Senior Solutions Engineer - Systems Integration to own how BESS subsystems come... ...engineering, define requirements, and ensure safe, reliable operation across diverse project environments. Responsibilities...Senior
- Muon Space, Inc. seeks a Systems Engineer, Mission Platforms to join our Mission Engineering team in San Jose, CA. You will serve as the primary systems integrator for end-to-end mission capability baselines, translating roadmaps into unified baselines spanning hardware...Senior
- Anduril Industries in Costa Mesa, California is seeking a Systems Engineer to drive the technical direction for Tactical Recon & Strike programs, including Altius, Anvil, Bolt, and Ghost. You will lead cross-functional teams to deliver robust, safety-critical capabilities...Senior
- ...Senior Software Engineer – Frontend We are seeking a Frontend Engineer to design and scale AI-powered applications that automate complex professional workflows. You will work closely with the leadership team and domain experts to deploy systems used by specialized...SeniorH1bRelocationVisa sponsorship
- Dormont Manufacturing Co is seeking a Software Developer to champion the design of innovative information security products. This pivotal role involves delivering exceptional quality in software development and ensuring the robustness of product design and architecture....Senior
- DocuSign is seeking a Lead AI Solutions Delivery Engineer to engage with enterprise customers to deliver AI-driven solutions. You will design workflows, integrate an array of Docusign’s capabilities, and optimize enterprise processes. The ideal candidate has over 12 years...SeniorRemote job
- ...Trilyon, Inc. seeks a Senior Consultant CTRM to drive enterprise SAP Commodity Management and CTRM transformations, leading large-scale... ...strong leadership, governance, and communication skills. This is an on-site, 12-month contract based in Maumee, OH. #J-18808-Ljbffr...SeniorContract work
- ...loyalty in one product. The platform serves thousands of businesses and supports a remote-first culture across the US. Boulevard seeks a Senior Product Manager to own the developer ecosystem, shape platform strategy, and coordinate partnerships with external teams to drive...SeniorRemote job
- Databricks in San Francisco Bay Area is seeking a Sr. Staff Technical Solutions Engineer to partner with Field and Engineering teams, delivering high-touch, tailored technical solutions for the largest Databricks customers in the DNB segment. You will leverage Apache Spark...Senior
- Base-2 Solutions is seeking a seasoned software engineer to advance analytics and large-scale systems for national defense. You will work on diverse tech stacks, contribute to real-time processing, and collaborate across teams to deliver robust software solutions. The...Senior
- The Google for Education Solutions Engineer at SHI International Corp. is a senior, customer-facing technical leader responsible for managing the full lifecycle of Google for Education engagements—from pre-sales discovery through deployment and long-term success. This role...Senior
$110k - $115k
Planned Parenthood of the Pacific Southwest is hiring a Senior Philanthropy Officer for its Louisville or Indianapolis locations. This role is pivotal in managing relationships with principal donors and overseeing planned giving initiatives across Indiana and Kentucky....Senior- Compliance Engineering, Site Reliability Engineering, Vice President, Dallas location_on Dallas, TX, United States We are Compliance Engineering, a global team of more than 300 engineers and scientists who work on the most complex, mission-critical problems. We build and...Full timeTemporary workWork at office
- Mirantis is seeking a Senior Software Engineer (Storage) to design and build high-performance storage provisioning, integration, and tooling for GPU-accelerated AI platforms. The role focuses on Go-based control-plane services, CSI drivers, and declarative workflows across...SeniorRemote work
- Alpaca, a US-headquartered brokerage infrastructure provider, is seeking a Senior Engineer to design, implement, and maintain its core systems and services that empower millions of users trading billions of dollars in assets. You will mentor regional engineers, lead incident...SeniorRemote job
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- on-site clinical research associate (traveling/remote) Ridgewood, NY
- site reliability engineer remote
- site reliability engineer sre
- site reliability engineering manager
- site reliability engineer
- lead site reliability engineer
- junior site reliability engineer
- senior workforce manager
- senior wedding planner
- senior water technician


