Senior Site Reliability Engineer
2T Consulting
We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.
The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.
Required Technical Skills
- Strong understanding of Site Reliability Engineering principles and operational excellence.
- Experience with infrastructure reliability, service availability, resiliency, and performance optimization.
- Storage Space Direct and failover clustering technical expertise. (Storage Spaces Direct enables you to build highly available, software-defined storage by pooling local disks (SSDs, NVMe drives, and HDDs) across multiple Windows Server nodes in a cluster. Instead of relying on an external SAN, S2D uses the servers' local storage to create a resilient shared storage pool)
- Experience managing production-critical infrastructure environments with high availability requirements.
- Experience with incident management, problem management, RCA, and continuous operational improvement.
- Knowledge of monitoring, observability, alerting, and performance management.
Microsoft Hyper-V (Core Expertise)
- Deep hands-on expertise in Microsoft Hyper-V architecture, deployment, administration, troubleshooting, and optimization.
- Extensive experience in operating enterprise private cloud environments on Hyper-V.
- Strong experience supporting enterprise-scale VDI deployments on Hyper-V.
- Hyper-V Failover Clustering and high-availability architecture.
- Storage integration including SAN, NAS, Storage Spaces Direct (S2D), Cluster Shared Volumes (CSV), and storage optimization.
- Networking within Hyper-V environments including virtual switches, VLANs, NIC Teaming, QoS, and network performance tuning.
- System Center Virtual Machine Manager (SCVMM).
Automation & Platform Engineering
- Strong PowerShell scripting and automation experience.
- Experience automating infrastructure deployment, operational tasks, health checks, and reporting.
- Familiarity with Infrastructure as Code concepts and configuration management.
- Experience developing reusable operational tooling to improve reliability and reduce manual effort.
Preferred Skills
- Windows Server 2016/2019/2022 administration.
- Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
- Exposure to hybrid cloud and private cloud platforms.
- Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
- Experience supporting enterprise VDI environments.
- Understanding of ITIL Incident, Problem, Change, and Release Management.
- Experience working in regulated industries such as Banking or Financial Services.
Experience & Qualifications
- 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
- Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
- Proven experience implementing automation to reduce operational overhead and improve service reliability.
- Experience supporting enterprise private cloud and VDI environments.
- Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
- Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
- Experience in Banking or Financial Services environments is advantageous.
Key Responsibilities
- Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
- Optimize, and support highly available VDI environments on Hyper-V.
- Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
- Disaster recovery, backup, patch management, and business continuity strategies.
- Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
- Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
- Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
- Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
- Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
- Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
- Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
- Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
- Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
- Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.
- ...We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern...SeniorLocal area
- LiveRamp, a data collaboration platform leader, is seeking a Senior Site Reliability Engineer in San Francisco with 5+ years of SRE/DevOps experience. The role focuses on deployments, 24/7 support across regions, and establishing SRE best practices. Strong skills in Terraform...Senior
$174k - $252k
Senior Software Engineer, Site Reliability Engineering corporate_fare Google place Seattle, WA, USA ; Kirkland, WA, USA Mid Experience driving progress, solving problems, and mentoring more junior team members; deeper expertise and applied knowledge within relevant area...SeniorTemporary work$152.6k - $191.5k
...is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include... ...continuous improvement. Position Summary: The Senior Site Reliability Engineer acts as an advanced senior...SeniorFull timeWork at officeShift workDay shift- Google is seeking a Senior Software Engineer in Site Reliability Engineering to strengthen the reliability of Google's public services from the Seattle/Kirkland area. You will design, build, and operate scalable systems with emphasis on availability and performance. The...Senior
- Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability.As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Chief Data & Analytics...SeniorWork at office
- Rocket Lab in Long Beach, CA, seeks a Senior Principal Software Engineer (TS/SCI) to lead the architecture and delivery of software for space systems. You will own flight and ground software, drive testing, and mentor teams to ensure mission success in a fast-paced aerospace...Senior
- ...Mount Thor, Inc. is seeking an on-site engineer to design and improve the data center systems behind our compute fleet. You will own the technical direction for data center architecture and operations, define standards for power, cooling, cabling, and network integration...
- ...Overview We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise...
- ...First Due seeks a Director of Platform Engineering to lead DevOps, SRE, and DBRE, turning fragmented practices into a centralized, disciplined... ...hands-on depth and executive presence to represent platform reliability to the ELT. Reporting to SVP of Engineering, this leader will...
$132.23k - $176.31k
...shape the future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem....SeniorFull timeTemporary workRemote work- ...Collins Aerospace MiS in Cedar Rapids, IA seeks a Senior Principal Systems Engineer to lead analysis, architecture, and design across total system products, including concept, design, modeling, testing, and disposal. You will decompose customer specs and develop verification...Senior
$253k - $336k
...TEAM: CorpTech Platform is the internal engineering force multiplier behind Anduril's... ...products. ABOUT THE JOB: The Director of Site Reliability Engineering owns the reliability system... ...structure required to execute it, and give senior leaders a clear view of operational...Full timeWork experience placement$170k - $240k
...thrive.The RoleAre you a Platform Engineer looking to make a big impact... ...looking for a self-motivated Senior Platform Engineer to join our... ...improve the security and reliability of our applications running on... ...years experience in DevOps, Site Reliability Engineering, Platform...SeniorRemote work- ...NVIDIA Corporation in Santa Clara, CA seeks a Senior Software Engineer to build and enhance CI/CD, testing, and delivery workflows for NVIDIA's AI software stack. You will enable developers and researchers to push boundaries in AI by stabilizing the software across the...SeniorWorldwide
- ...security platform used by thousands of companies worldwide. As a Senior Software Engineer on the Test Core team, you will own core backend systems,... ...growing data and demand. You will mentor engineers, design reliable architectures, and collaborate across teams to improve...SeniorWorldwide
- ...applications and next steps. Our partner is looking for a Senior Software Engineer I - Platform Enablement based in United States. This role sits... ...opportunity to raise engineering standards while building reliable, secure infrastructure that enables teams to move faster....Senior
$119.3k - $196.6k
...Ordinary and NIOD, and BALMAIN Beauty.DescriptionThe Senior Lead, Finance Platform Engineering provides engineering leadership and is accountable for... ...technology partners to deliver secure, scalable, and reliable Finance capabilities that support global business operations...SeniorFull timeLocal area- Everforth Apex in Charlotte, NC (Hybrid with Newark, DE) is seeking a Spring Boot Developer to design and deliver complex middleware solutions that meet functional and compliance requirements. You will implement REST and SOAP web services, apply design patterns, contribute...Senior
- ...Bank of America’s Enterprise Payments Technology team is seeking a senior Android engineer to design, develop and deliver complex payment solutions. You will ensure software meets functional, non-functional and compliance requirements, with maintainable architecture and...SeniorWork at office
- Muon Space, Inc. seeks a Systems Engineer, Mission Platforms to join our Mission Engineering team in San Jose, CA. You will serve as the primary systems integrator for end-to-end mission capability baselines, translating roadmaps into unified baselines spanning hardware...Senior
- ...Rocket Lab is seeking a Senior Software Engineer I - Data Engineering to design and scale core data platforms powering operations. Based onsite in Long Beach, CA, you’ll build pipelines, storage, APIs, and processing tools used by engineers across the company. You’ll...Senior
- Iridium Communications Inc. seeks an experienced Senior Software Engineer to join the team developing a state-of-the-art satellite ground system and user equipment for the company’s PNT services. You will lead the design and architecture for PNT software products and work...Senior
- Cloudflare is seeking a Senior Solution Engineer to drive technical discovery and adoption across enterprise customers. This remote US role focuses on architecting solutions, leading demos/PoCs, and partnering with Account Executives to expand relationships and value realization...SeniorRemote job
- EnCharge AI is seeking a highly experienced, hands-on Principal Solutions Engineer to lead customer and partner solution development for our AI acceleration platform. This senior role bridges AI hardware, software, systems, and customers. You will own strategic customer...Senior
- Klaviyo in Boston, Massachusetts, is seeking a Senior Platform Engineer for its Asynchronous Processing team. You will architect, build, and... ...Kafka, and SQS running in AWS and Kubernetes. You will own reliability, scalability, and observability, mentor engineers, participate...Senior
- Valkai, Inc. is seeking an experienced infrastructure engineer to build and operate secure, scalable systems underpinning our agents and products. You will own deployment, observability, security, and performance across cloud and self-hosted environments while collaborating...Senior
- Triwill Group is seeking a Senior Sales Engineer based in the United States to own the technical side of complex cybersecurity deals. The role focuses on guiding discovery and solution design through validation, demonstrations, and close. You will work closely with Account...Senior
- ...Senior Software Engineer – FrontendWe are seeking a Frontend Engineer to design and scale AI-powered applications that automate complex professional workflows. You will work closely with the leadership team and domain experts to deploy systems used by specialized professionals...SeniorH1bRelocationVisa sponsorship
- ...loyalty in one product. The platform serves thousands of businesses and supports a remote-first culture across the US. Boulevard seeks a Senior Product Manager to own the developer ecosystem, shape platform strategy, and coordinate partnerships with external teams to drive...SeniorRemote job
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- on-site clinical research associate (traveling/remote) Ridgewood, NY
- site reliability engineer
- site reliability engineering manager
- junior site reliability engineer
- site reliability engineer sre
- site reliability engineer remote
- lead site reliability engineer
- senior warehouse specialist
- senior textile designer
- senior wellbeing practitioner


