Site Reliability Engineer
2T Consulting
We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.
The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.
Required Technical Skills
- Strong understanding of Site Reliability Engineering principles and operational excellence.
- Experience with infrastructure reliability, service availability, resiliency, and performance optimization.
- Storage Space Direct and failover clustering technical expertise. (Storage Spaces Direct enables you to build highly available, software-defined storage by pooling local disks (SSDs, NVMe drives, and HDDs) across multiple Windows Server nodes in a cluster. Instead of relying on an external SAN, S2D uses the servers' local storage to create a resilient shared storage pool)
- Experience managing production-critical infrastructure environments with high availability requirements.
- Experience with incident management, problem management, RCA, and continuous operational improvement.
- Knowledge of monitoring, observability, alerting, and performance management.
Microsoft Hyper-V (Core Expertise)
- Deep hands-on expertise in Microsoft Hyper-V architecture, deployment, administration, troubleshooting, and optimization.
- Extensive experience in operating enterprise private cloud environments on Hyper-V.
- Strong experience supporting enterprise-scale VDI deployments on Hyper-V.
- Hyper-V Failover Clustering and high-availability architecture.
- Storage integration including SAN, NAS, Storage Spaces Direct (S2D), Cluster Shared Volumes (CSV), and storage optimization.
- Networking within Hyper-V environments including virtual switches, VLANs, NIC Teaming, QoS, and network performance tuning.
- System Center Virtual Machine Manager (SCVMM).
Automation & Platform Engineering
- Strong PowerShell scripting and automation experience.
- Experience automating infrastructure deployment, operational tasks, health checks, and reporting.
- Familiarity with Infrastructure as Code concepts and configuration management.
- Experience developing reusable operational tooling to improve reliability and reduce manual effort.
Preferred Skills
- Windows Server 2016/2019/2022 administration.
- Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
- Exposure to hybrid cloud and private cloud platforms.
- Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
- Experience supporting enterprise VDI environments.
- Understanding of ITIL Incident, Problem, Change, and Release Management.
- Experience working in regulated industries such as Banking or Financial Services.
Experience & Qualifications
- 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
- Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
- Proven experience implementing automation to reduce operational overhead and improve service reliability.
- Experience supporting enterprise private cloud and VDI environments.
- Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
- Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
- Experience in Banking or Financial Services environments is advantageous.
Key Responsibilities
- Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
- Optimize, and support highly available VDI environments on Hyper-V.
- Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
- Disaster recovery, backup, patch management, and business continuity strategies.
- Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
- Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
- Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
- Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
- Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
- Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
- Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
- Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
- Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
- Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.
- ...We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern...SuggestedLocal area
- ...Overview We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise...Suggested
- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team,, you will solve complex...SuggestedWork at office
$132.23k - $176.31k
...future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role...SuggestedFull timeTemporary workRemote work- ...The selected colleague will work at an MUFG office or client sites four days per week and work remotely one day. A member of... ...MUFG is seeking a highly motivated Certified Sr. Cloud Site Reliability Engineer to build a robust, scalable, and reliable web application environment...SuggestedFull timeWork at officeLocal areaRemote work
$83.54k - $137.24k
...progress and enhancing lives by providing reliable, high-speed connectivity solutions that... ...you! Job Summary The Role DNS Engineer – SRE is a high-impact role responsible for... ...views infrastructure through the lens of Site Reliability Engineering (SRE)...Local areaRemote work- ...Mount Thor, Inc. is seeking an on-site engineer to design and improve the data center systems behind our compute fleet. You will own the technical direction for data center architecture and operations, define standards for power, cooling, cabling, and network integration...
- ...First Due seeks a Director of Platform Engineering to lead DevOps, SRE, and DBRE, turning fragmented practices into a centralized, disciplined... ...hands-on depth and executive presence to represent platform reliability to the ELT. Reporting to SVP of Engineering, this leader will...
- ...core SRE principles across hybrid environments. You will shape strategy, architecture, and roadmaps for internal observability and reliability tooling, spanning telemetry standards, pipeline integration, and SRE maturity. The role blends hands-on delivery with leadership,...
- ...Principal Site Reliability Engineer (SRE)Are you ready to make an impact at DTCC? Do you want to work on innovative projects, collaborate with a dynamic and supportive team, and receive investment in your professional development? At DTCC, we are at the forefront of innovation...
$152.6k - $191.5k
...is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include composing... ...improvement. Position Summary: The Senior Site Reliability Engineer acts as an advanced senior...Full timeWork at officeShift workDay shift- ...contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology Team, you draw upon your advanced knowledge to identify new opportunities...
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Cloud Foundational Services team, you hold a leadership role in your team, demonstrate...
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer focussed on Operations Excellence, you will have the opportunity to shape how we respond to, learn from, and...
- ...Job Description Job Description BCforward is currently seeking a highly motivated SRE Software Engineer. Job Title: SRE Software Engineer Location: Jersey City, NJ Duration: Temp - 12 months Job Description We are seeking a Software Engineer-Other...Temporary work
- ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Commercial Investment Banking team of Fraud Prevention , you will solve complex and broad...
- LiveRamp, a data collaboration platform leader, is seeking a Senior Site Reliability Engineer in San Francisco with 5+ years of SRE/DevOps experience. The role focuses on deployments, 24/7 support across regions, and establishing SRE best practices. Strong skills in Terraform...
$253k - $336k
...not years. ABOUT THE TEAM: CorpTech Platform is the internal engineering force multiplier behind Anduril's corporate systems. It... ...its mission critical products. ABOUT THE JOB: The Director of Site Reliability Engineering owns the reliability system for the software that...Full timeWork experience placement$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week$174k - $252k
Senior Software Engineer, Site Reliability Engineering corporate_fare Google place Seattle, WA, USA ; Kirkland, WA, USA Mid Experience driving progress, solving problems, and mentoring more junior team members; deeper expertise and applied knowledge within relevant area...Temporary work- Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability.As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Chief Data & Analytics...Work at office
$61k - $101k
...Salary: $61,000 - 101,000 per year Requirements: Formal training or certification in site reliability engineering concepts, with 5+ years of applied experience Formal training or certification in site reliability engineering concepts, with advanced practical experience...Full time- ...Job Description Job Description The primary responsibility of the System Engineer is to design, build, and maintain the Bank’s infrastructure such as internal systems, network and various banking-related applications. ESSENTIAL FUNCTIONS AND RESPONSIBILITIES...Bank staffLocal area
$119.3k - $196.6k
...and BALMAIN Beauty.DescriptionThe Senior Lead, Finance Platform Engineering provides engineering leadership and is accountable for the... ...strategic technology partners to deliver secure, scalable, and reliable Finance capabilities that support global business operations.This...Full timeLocal area- ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Asset and Wealth Management team, you will solve complex and broad business problems...Work experience placement
$95.6k - $138.7k
...this job involves:The Critical Systems Engineer provides onsite field coordination, supervision... ...the highest standards of operational reliability and safety compliance.Your day-to-day... ...facility operationsLocation & Schedule:On-site position with 24/7 on-call availability...Full timeFor contractorsFor subcontractorLocal area- Job Title: Lead Software Engineer - Full StackAbout the OpportunityA rapidly growing life sciences... ....Key ResponsibilitiesDevelop scalable, reliable, testable, and maintainable software... ...York City area.Must be willing to work on-site in Long Island City, NY.Preferred...Live inRelocation
$60 - $67 per hour
Site Reliability EngineerColumbus, OH - onsite from day oneContract to HireOnsite Interview 60% SRE & 40% EngineeringTop skills: Public Cloud... ...Required:Formal training or certification in software engineering /Site Reliability Engineering concepts and 6+ years of applied...Full time$105.6k - $150.4k
Position SummaryThe Lead Engineer Full Stack leads development teams in designing, testing, and implementing responsive web applications. In addition to leading and delegating work to other developers, the Lead Engineer will be responsible for standardization of documentation...Temporary workWork at officeImmediate startFlexible hoursNight shift$120k - $180k
Lead Software Engineer, Full StackLong Island City, NYCompensation: $120,000 to $180,000 per... ...the industry.ScheduleThis is an on-site position based in Long Island City, NY. Candidates... ...to day, this person develops scalable, reliable, testable, and maintainable software...Live inRelocation
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- on-site clinical research associate (traveling/remote) Corona, NY
- site safety Corona, NY
- construction site safety Corona, NY
- site reliability engineer
- site reliability engineering manager
- junior site reliability engineer
- site reliability engineer sre
- site reliability engineer remote
- lead site reliability engineer
- savannah river site




