Site Reliability Engineer
2T Consulting
We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.
The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.
Required Technical Skills
- Strong understanding of Site Reliability Engineering principles and operational excellence.
- Experience with infrastructure reliability, service availability, resiliency, and performance optimization.
- Storage Space Direct and failover clustering technical expertise. (Storage Spaces Direct enables you to build highly available, software-defined storage by pooling local disks (SSDs, NVMe drives, and HDDs) across multiple Windows Server nodes in a cluster. Instead of relying on an external SAN, S2D uses the servers' local storage to create a resilient shared storage pool)
- Experience managing production-critical infrastructure environments with high availability requirements.
- Experience with incident management, problem management, RCA, and continuous operational improvement.
- Knowledge of monitoring, observability, alerting, and performance management.
Microsoft Hyper-V (Core Expertise)
- Deep hands-on expertise in Microsoft Hyper-V architecture, deployment, administration, troubleshooting, and optimization.
- Extensive experience in operating enterprise private cloud environments on Hyper-V.
- Strong experience supporting enterprise-scale VDI deployments on Hyper-V.
- Hyper-V Failover Clustering and high-availability architecture.
- Storage integration including SAN, NAS, Storage Spaces Direct (S2D), Cluster Shared Volumes (CSV), and storage optimization.
- Networking within Hyper-V environments including virtual switches, VLANs, NIC Teaming, QoS, and network performance tuning.
- System Center Virtual Machine Manager (SCVMM).
Automation & Platform Engineering
- Strong PowerShell scripting and automation experience.
- Experience automating infrastructure deployment, operational tasks, health checks, and reporting.
- Familiarity with Infrastructure as Code concepts and configuration management.
- Experience developing reusable operational tooling to improve reliability and reduce manual effort.
Preferred Skills
- Windows Server 2016/2019/2022 administration.
- Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
- Exposure to hybrid cloud and private cloud platforms.
- Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
- Experience supporting enterprise VDI environments.
- Understanding of ITIL Incident, Problem, Change, and Release Management.
- Experience working in regulated industries such as Banking or Financial Services.
Experience & Qualifications
- 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
- Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
- Proven experience implementing automation to reduce operational overhead and improve service reliability.
- Experience supporting enterprise private cloud and VDI environments.
- Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
- Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
- Experience in Banking or Financial Services environments is advantageous.
Key Responsibilities
- Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
- Optimize, and support highly available VDI environments on Hyper-V.
- Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
- Disaster recovery, backup, patch management, and business continuity strategies.
- Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
- Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
- Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
- Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
- Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
- Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
- Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
- Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
- Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
- Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.
- ...Overview We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise...Suggested
- ...We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern...SuggestedLocal area
- ...Versant Media is seeking a hands-on System Reliability Operations Engineer to scale reliability practices across its enterprise platforms. You will help create standards, runbooks, and governance with collaboration across Software, Platform, Infrastructure, and Security...Suggested
$140k - $150k
WORK OPTION: Remote_________________The NBA is hiring a Senior Site Reliability Engineer (SRE) - Messaging & Collaboration to ensure the availability, performance, and reliability of enterprise messaging and collaboration platforms, including Microsoft Exchange Online (...SuggestedFull timeTemporary workLocal areaRemote workWeekend work$90k - $120k
As a Performance II-Epic, your role is to provide reliability engineering services through observability and performance engineering techniques.... ...passion for optimizing operational efficiency. You will use Site Reliability Engineering practices to deliver a seamless user...SuggestedFull timePart timeWork experience placementRemote workFlexible hours- ...Site Reliability Engineer As a Site Reliability Engineer, your role is to provide reliability engineering services through observability and performance engineering techniques. Using monitoring and performance tools to deliver detailed feedback to product owners and...Work experience placement
- ...Tomatoes, GolfNow and GolfPass. Job Description The Cloud Reliability Engineer is responsible for ensuring the availability, performance,... ...practical experience. ~3–7 years of experience in Site Reliability Engineering, Cloud Engineering, DevOps, Infrastructure...Work at officeLocal area
$46 - $63 per hour
...into user-centric solutions* Lead UI design discussions, provide technical feasibility assessments, and drive alignment between engineering and UX teams* Present UI solutions, feature demonstrations, and technical recommendations to customers, stakeholders, and leadership...Remote work$50 - $65 per hour
DescriptionKforce has a client that is seeking an experienced AI Platform Ops Engineer in Englewood, CO.Summary:The AI Platform Ops Engineer will help... ...solutions.You will be responsible for maintaining a secure, reliable, and highly automated platform while partnering with...- ...ABOUT THE ROLE Fandango is looking for a SENIOR PLATFORM ENGINEER to build our next big thing in platform growth across Fandango,... ...reusable patterns, and automation solutions that enable rapid, reliable, and secure deployments Design and implement standardized, reusable...Local area
$60k - $80k
...Solution Engineer At UVeye, we're on a mission to redefine vehicle safety and reliability on a global scale. Founded in 2016, we have pioneered the world's first fully automated suite of vehicle inspection systems. At the heart of this innovation lies our advanced AI...$125k
...Overview This position has an On-Site Requirement in Teaneck NJ Interstate Waste Services is the most progressive and innovative provider of solid waste and recycling services in the greater New York, New Jersey and Connecticut markets with a rail-served landfill...Local area- ...Description We are looking for a Senior Full-Stack Software Engineer to join our growing engineering team. You will play a key role... ...Englewood Cliffs office have access to a variety of convenient on-site services and amenities, including free employee parking,...Work at officeLocal area
- ...Job Description Versant’s Software Engineering team provides core services to our business... ..., and DevOps teams to ensure scalable, reliable platform operations. ~ Work closely... ...access to a variety of convenient on-site services and amenities, including free employee...Work at officeLocal area
- ...Senior Software Engineer – Full-Time Direct Employment Cognizant is seeking an experienced Senior Software Engineer for a full-time, long-term remote position. This position is being posted by Cognizant for direct employment with the company. It is not a staffing-...Full timeRemote work
- ...GolfPass. Job Description We are looking for a Software Engineer to join our growing engineering team. You will contribute to the... ...Englewood Cliffs office have access to a variety of convenient on-site services and amenities, including free employee parking,...Work at officeLocal area
- ...Neflix OSS • Experience even driven large scale enterprise applications using Kafka or RabbitMQ. • Demonstrable experience in Engineering best practices like Code Reviews, Code Re-factoring, Security audits, Performance tuning, building Operational tools and...
- ...and GolfPass. Job Description Versant’s Software Engineering team provides core services to our business, underpinning our... ...across our platform landscape. Drive improvements in latency, reliability, and scalability of streaming pipelines. Authentication,...Full timeLocal area
$60 - $75 per hour
...Englewood, CO that is seeking a Senior Agentic AI Platform Ops Engineer.Summary:The Senior Agentic AI Platform Ops Engineer will lead the... ...-functional engineering teams to deliver scalable, secure, and reliable AI platform solutionsRequirements* 5+ years of experience in...- ...description:The client's Digital Technology team is seeking a Software Engineer to manage and build software solutions across client's Digital... ...practices.Experience working on large scale, high traffic web sites / applications.Experience working in financial, media domain....
$150k - $210k
...Embodied AI EngineerWe are seeking an exceptional Embodied AI Engineer to build the foundation of LG's vision of theZero-labor Home. This... ...matching policies, world models) taking them from prototype to reliable and reproducible systems.Build and maintain scalable training...Full timeTemporary workFor contractorsLocal areaImmediate start- ...documentation, cost, and planning. The role requires managing construction plans, schedules, and cost curves, collaborating with engineers and managers to resolve issues, and drafting control-related communications. The position is full-time and onsite at the Haworth WTP...Full time
- ...real advantage. You will be joining a team of exceptional engineers, analysts, and investors working at the intersection of AI and... ...engineers who enjoy turning machine learning and AI capabilities into reliable product systems used by real customers at scale. It is...Full timeWork at officeLocal area
- ...Android applications with a focus on performance, usability, and video streaming. Collaborate with product managers, designers, and engineers to define and deliver new features. Optimize app performance, memory usage, and network efficiency to support a smooth user...Monday to FridayShift workDay shift3 days per week
- ..., Git). Excellent problem-solving skills and attention to detail. Strong communication and teamwork abilities. Qualifications: ~ Bachelor's degree in computer science, Engineering, or a related field. ~5+ years of experience in Java development....
- ...programming language. The ideal candidate will have a strong knowledge of software engineering principles and will be responsible for designing, developing, and maintaining high-performance and reliable code using Java and related technologies. The role also requires...
- ...linking, and App Store deployment Prior experience working in a high-paced, client-facing environment Education: ~ Bachelor's degree in Computer Science, Engineering, or related technical field (or equivalent experience) Skills: IOS,Swift,Kotlin
- Job Title Need only local candidates who can attend F2F interview. • Strong in core Java and Spring, Spring boot • Graphql, Python, AWS • Knowledge of Nifi, JOLT ( JSON to JSON ) and XSLT is a plus • Min 8 years' experience • Excellent communication skillsLocal area
- ...technologies and a track record of building high-performance, reliable apps. Key Responsibilities: Develop and maintain Android... ...Education: ~ Bachelor's degree in Computer Science, Engineering, or related technical field (or equivalent experience)...Work experience placement
$221.2k - $387.1k
...Job Description Principal Software Engineer The engineering organization is a dynamic group of builders, thinkers, and problem... ...Every engineer plays an important role in shaping the quality, reliability, and long-term success of our products. About the Team:...Work at officeRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer remote
- site reliability engineer sre
- site reliability engineering manager
- site reliability engineer
- lead site reliability engineer
- junior site reliability engineer
- site activation specialist
- website development
- on-site clinical research associate (traveling/remote)
- solar site auditor


