Senior Site Reliability Engineer
2T Consulting
We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.
The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.
Required Technical Skills
- Strong understanding of Site Reliability Engineering principles and operational excellence.
- Experience with infrastructure reliability, service availability, resiliency, and performance optimization.
- Storage Space Direct and failover clustering technical expertise. (Storage Spaces Direct enables you to build highly available, software-defined storage by pooling local disks (SSDs, NVMe drives, and HDDs) across multiple Windows Server nodes in a cluster. Instead of relying on an external SAN, S2D uses the servers' local storage to create a resilient shared storage pool)
- Experience managing production-critical infrastructure environments with high availability requirements.
- Experience with incident management, problem management, RCA, and continuous operational improvement.
- Knowledge of monitoring, observability, alerting, and performance management.
Microsoft Hyper-V (Core Expertise)
- Deep hands-on expertise in Microsoft Hyper-V architecture, deployment, administration, troubleshooting, and optimization.
- Extensive experience in operating enterprise private cloud environments on Hyper-V.
- Strong experience supporting enterprise-scale VDI deployments on Hyper-V.
- Hyper-V Failover Clustering and high-availability architecture.
- Storage integration including SAN, NAS, Storage Spaces Direct (S2D), Cluster Shared Volumes (CSV), and storage optimization.
- Networking within Hyper-V environments including virtual switches, VLANs, NIC Teaming, QoS, and network performance tuning.
- System Center Virtual Machine Manager (SCVMM).
Automation & Platform Engineering
- Strong PowerShell scripting and automation experience.
- Experience automating infrastructure deployment, operational tasks, health checks, and reporting.
- Familiarity with Infrastructure as Code concepts and configuration management.
- Experience developing reusable operational tooling to improve reliability and reduce manual effort.
Preferred Skills
- Windows Server 2016/2019/2022 administration.
- Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
- Exposure to hybrid cloud and private cloud platforms.
- Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
- Experience supporting enterprise VDI environments.
- Understanding of ITIL Incident, Problem, Change, and Release Management.
- Experience working in regulated industries such as Banking or Financial Services.
Experience & Qualifications
- 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
- Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
- Proven experience implementing automation to reduce operational overhead and improve service reliability.
- Experience supporting enterprise private cloud and VDI environments.
- Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
- Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
- Experience in Banking or Financial Services environments is advantageous.
Key Responsibilities
- Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
- Optimize, and support highly available VDI environments on Hyper-V.
- Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
- Disaster recovery, backup, patch management, and business continuity strategies.
- Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
- Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
- Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
- Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
- Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
- Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
- Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
- Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
- Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
- Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.
$140k - $150k
WORK OPTION: Remote_________________The NBA is hiring a Senior Site Reliability Engineer (SRE) - Messaging & Collaboration to ensure the availability, performance, and reliability of enterprise messaging and collaboration platforms, including Microsoft Exchange Online...SeniorFull timeTemporary workLocal areaRemote workWeekend work- Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability.As a Sr Lead Site Reliability Engineer at JPMorgan Chase within the Consumer & Community Banking...Senior
- ...PVH (Tommy Hilfiger/Calvin Klein) seeks a Senior Software Engineer to own the reliability and performance of our Kubernetes-based data platform across multi-region deployments. You will design scalable infrastructure, optimize deployment pipelines, and strengthen security...Senior
- ...are looking for people just like you. Join our team and help us develop game-changing, high-quality solutions.As a Senior Lead Site Reliability Engineer at JPMorganChase within the Core Engineering Solutions team of Consumer and Community Banking, you are an integral part...Senior
- Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability.As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the enterprise technology...Senior
$113.3k
...is important to maintain our strong culture, achieve our goals, and thrive as #OneJamf. What you'll do at Jamf: As a Senior Site Reliability Engineer, you'll help us balance development velocity with the reliability our customers depend on. You'll partner with engineering...SeniorWork at officeRemote workWorldwideFlexible hours- ...meaningful products that make a real impact on children's education and literacy. About the Role We're looking for a Senior Site Reliability Engineer to drive the stability, observability, and reliability of Epic's platform as we grow. You are an experienced engineer...SeniorRemote work
$175k - $185k
...together. Come join our team as we develop new ways to improve the lives of working Americans. About the role: This Senior Database Reliability Engineer is responsible for supporting all production database systems so that they run smoothly. We run MySQL 8.4 on Google...SeniorDaily paidRemote workHome officeFlexible hours- As a Performance II-Epic, your role is to provide reliability engineering services through observability and performance engineering techniques.... ...passion for optimizing operational efficiency. You will use Site Reliability Engineering practices to deliver a seamless user...Full timePart timeWork experience placementRemote workFlexible hours
- ...Site Reliability Engineer As a Site Reliability Engineer, your role is to provide reliability engineering services through observability and performance engineering techniques. Using monitoring and performance tools to deliver detailed feedback to product owners and...Work experience placement
$180k - $215k
...youth and family development, and health-related causes.Position SummaryThe Senior Engineering Lead leads the Broadcast Platform & Systems Support Engineering team, responsible for the reliability, support, and continuous evolution of the NBA’s mission-critical broadcast...SeniorFull timeTemporary workLocal areaRemote workFlexible hoursNight shift1 day per week$121.2k - $181.8k
...activities.Role Description:We are looking for a Senior Developer in application development... ...leadership and guidance to junior Dev engineers.Mentor team members on best practices... ....Optimize CI/CD pipelines for speed and reliability.Containerization and Orchestration:Lead...SeniorFull time$120k - $202.5k
...Analytics (CyberDNA) team is seeking a Sr.Platform Operations Engineer to help shape the next generation of cybersecurity data, analytics... ...environment. The ideal candidate support the operational reliability, administration, monitoring, troubleshooting, and continuous improvement...SeniorFull timeTemporary workFlexible hours- ...responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include... ...continuous improvement. Position Summary: The Senior GCP Site Reliability Engineer acts as an advanced senior...SeniorWork at officeShift workDay shift
- ...Senior Java Engineer Job Location: New Jersey Job Type: Contract Required Skillset: Design and development experience with a strong command of Java J2EE and custom frameworks. Design & development experience on Spark, Elastic Search, Kafka. Expert...SeniorContract work
- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the within the Consumer & Community Banking Data and Analytics team, you will solve complex and...
- ...Senior React JS Developer Location: Dallas, TX OR Rutherford, NJ (ONSITE) Full Time ONLY Roles and Responsibilities: Lead front-end development efforts using React to build responsive and dynamic user interfaces. Lead front-end architectural decisions and...SeniorFull time
- ...contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology and Enterprise Technology Team, you draw upon your advanced knowledge to...Local area
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability.As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate Technology team, you hold a leadership role in your team, demonstrate...Local area
- Responsibilities Kforce has a client in Secaucus, NJ that is seeking a Senior Software Engineer (Storage). Summary: The ideal candidate will possess deep... ..., develop, and maintain high-performance, scalable, and reliable backend services supporting distributed storage platforms...SeniorHourly payContract work
$120k - $202.5k
Who We Are Looking ForWe are looking for a Senior Cloud Security Architect. You will be responsible for defining, implementing, and... ...private cloud, and SaaS environments. You will partner with cloud engineering, application development, infrastructure, cybersecurity, and...SeniorFull timeTemporary workFlexible hoursShift work- ...QualificationsBA degree or higherComputer Science or related disciplines.7+ years working with technical teams to define software business requirementsSummaryFunction: Information TechnologyExperience level: Mid-Senior LevelIndustry: Information Technology And ServicesSenior
$121.2k - $181.8k
...Category: Technology, Applications DevelopmentCompany: CitiJob SummaryWe are seeking a highly motivated and experienced Principal Engineer to join our Retail and Wealth Risk Engineering team under the Enterprise Risk Technology platform. This is an intermediate-level position...SeniorFull time$114k - $148k
...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based...Full timeTemporary workWork experience placementRemote work$189.59k - $228k
Pershing LLC seeks Senior Vice President, Release Train Engineer in Jersey City, NJ, to perform Agile Release Train (ART) or product. Facilitate development events and processes and assist the teams in delivering value. Communicate with stakeholders, escalate impediments...SeniorTemporary workWork at officeRemote workWorldwideFlexible hours- ...exceptional professionals for this role. JOB DESCRIPTION Elevate your engineering prowess to unprecedented levels by joining a team of... ...and position yourself among the top echelon in site reliability. As an Associate Site Reliability Engineer at JPMorgan Chase...Worldwide
$130k - $180k
...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and... ...an in-house AI R&D team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to...Temporary workWork at officeImmediate startRemote workFlexible hours- ...Job Title : Site Reliability Engineer (AWS) (SRE)- Location : Jersey city ,NJ -( 3 days WFO, 2 days WFH) Duration : 6 +Months Position Responsibilities: Site Reliability Engineer (AWS) (SRE) Work Location: Jersey city New Jersey Only near by...Work from homeWeekend work
$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will... ...Professional development From entry-level employees to senior leaders, we believe there’s always room to learn. We offer...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week- ...selected colleague will work at an MUFG office or client sites four days per week and work remotely one day. A... ...recruitment team will provide more details.Job Summary:The Senior Software Deployment and Patching Engineer is responsible for managing and optimizing the...SeniorFull timeWork at officeLocal areaRemote work1 day per week
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- senior operations technician Carlstadt, NJ
- construction site safety Carlstadt, NJ
- on-site clinical research associate (traveling/remote) Carlstadt, NJ
- site reliability engineering manager
- junior site reliability engineer
- site reliability engineer
- site reliability engineer sre
- site reliability engineer remote
- lead site reliability engineer
- senior travel

