Site Reliability Engineer
Longfinch Technologies
Overview
We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.
The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.
Key Responsibilities
- Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
- Optimize, and support highly available VDI environments on Hyper-V.
- Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
- Disaster recovery, backup, patch management, and business continuity strategies.
- Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
- Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
- Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
- Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
- Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
- Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
- Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
- Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
- Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
- Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.
Experience & Qualifications
- 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
- Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
- Proven experience implementing automation to reduce operational overhead and improve service reliability.
- Experience supporting enterprise private cloud and VDI environments.
- Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
- Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
- Experience in Banking or Financial Services environments is advantageous.
Preferred Skills
- Windows Server 2016/2019/2022 administration.
- Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
- Exposure to hybrid cloud and private cloud platforms.
- Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
- Experience supporting enterprise VDI environments.
- Understanding of ITIL Incident, Problem, Change, and Release Management.
- Experience working in regulated industries such as Banking or Financial Services.
$140k - $150k
WORK OPTION: Remote_________________The NBA is hiring a Senior Site Reliability Engineer (SRE) - Messaging & Collaboration to ensure the availability, performance, and reliability of enterprise messaging and collaboration platforms, including Microsoft Exchange Online (...SuggestedFull timeTemporary workLocal areaRemote workWeekend work$90k - $120k
As a Performance II-Epic, your role is to provide reliability engineering services through observability and performance engineering techniques.... ...passion for optimizing operational efficiency. You will use Site Reliability Engineering practices to deliver a seamless user...SuggestedFull timePart timeWork experience placementRemote workFlexible hours- ...We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern...SuggestedLocal area
- ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team,, you will solve complex...SuggestedWork at office
$175k - $185k
...together. Come join our team as we develop new ways to improve the lives of working Americans. About the role: As the Senior Site Reliability Engineer, you will lead Branch’s effort to achieve greater reliability, performance, scalability, capacity and observability of our...SuggestedDaily paidRemote workHome officeFlexible hours- ...meaningful products that make a real impact on children's education and literacy. About the Role We're looking for a Senior Site Reliability Engineer to drive the stability, observability, and reliability of Epic's platform as we grow. You are an experienced engineer who...Remote work
$140k - $170k
...experience) Eligibility: U.S. citizenship required (customer badging requirement); no security clearance required The Role As a Site Reliability Engineer on Blitzy's Public Sector team, you will be the backbone of our platform's reliability and operational excellence for a...Remote work- ...PVH (Tommy Hilfiger/Calvin Klein) seeks a Senior Software Engineer to own the reliability and performance of our Kubernetes-based data platform across multi-region deployments. You will design scalable infrastructure, optimize deployment pipelines, and strengthen security...
$155k - $175k
...Next! Summary We are seeking a highly skilled and experienced Site Reliability Manager to join our team to ensure the reliability,... ...performance of our systems and services. You will lead a team of engineers focusing on three core pillars: Application Reliability, DevSecOps...Work experience placementH1bWork at officeLocal area- ...resolution or escalation (ServiceNow/Jira). Perform physical DC tasks: cable installation, hardware swap-outs, rack and stack, labeling (on-site roles). Execute structured shift handoffs at 8AM and 8PM PST with the APAC operations team. Maintain and update operational runbooks...Shift workNight shift
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer focussed on Operations Excellence, you will have the opportunity to shape how we respond to, learn from, and...
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Cloud Foundational Services team, you hold a leadership role in your team, demonstrate...
- ...contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology Team, you draw upon your advanced knowledge to identify new opportunities...
- ...readiness Collaborate with development teams to identify reliability risks and improve system architecture Contribute to... ...infrastructure security controls and best practices Certifications in SRE, DevOps, or Performance Engineering are a plus...Local area
- ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Consumer and Investment Banking team, you will solve complex and broad business problems...Local area
- ...is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include composing... .... Position Summary: The Senior Azure Site Reliability Engineer acts as an advanced senior individual...Work at officeShift workDay shift
$114k - $148k
...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based...Full timeTemporary workWork experience placementRemote work- ...scale, we invite you to bring your talents to Zscaler to help shape the future of cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or Bellevue WA) to join our Zero Trust Exchange team. This is a remote role based in San...InternshipWork at officeLocal areaRemote workWorldwide
$160k - $240k
...passionate about building unified IT solutions that simplify the way IT organizations work. We are currently looking for a Senior Site Reliability Engineer to join our SRE team in the Platform Engineering organization and help us scale our products to millions of end-users. We...Permanent employmentFull timeRemote workWork from homeRelocationFlexible hours$150k - $200k
...our CEO's funding announcement: . The Reliability team owns the availability, performance,... ...enforcing reliability standards across engineering Designing incident response processes and... ...strong ownership of production systems. As a Site Reliability Engineer on the Reliability...Remote workVisa sponsorshipWork visaFlexible hours$253k - $336k
...not years. ABOUT THE TEAM: CorpTech Platform is the internal engineering force multiplier behind Anduril's corporate systems. It... ...its mission critical products. ABOUT THE JOB: The Director of Site Reliability Engineering owns the reliability system for the software that...Full timeWork experience placement- ...are looking for people just like you. Join our team and help us develop game-changing, high-quality solutions. As a Lead Site Reliability Engineer at JPMorganChase within the Corporate sector, Enterprise Technology team, you are an integral part of a team that develops...Work at office
$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week$174k - $252k
Senior Software Engineer, Site Reliability Engineering corporate_fare Google place Seattle, WA, USA ; Kirkland, WA, USA Mid Experience driving progress, solving problems, and mentoring more junior team members; deeper expertise and applied knowledge within relevant area...Temporary work- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the AI Machine Learning and Data platform team, you hold a leadership role in...
- Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability.As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Chief Data & Analytics...Work at office
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Cloud Foundational Services team, you hold a leadership role in your team, demonstrate...
- ...we serve.The Information Technology group delivers secure, reliable technology solutions that enable DTCC to be the trusted... ...scalability, and performance of enterprise platforms.As a Principal Site Reliability Engineer (SRE), you will drive operational excellence across...Remote workFlexible hours
- Mars, Incorporated and its Affiliates is seeking a Site OT/IT Specialist to ensure delivery of OT/IT services at the site and to support... ...for Factory OT/IT initiatives. Join Mars to help ensure reliable operations and continuous improvement. #J-18808-Ljbffr Mars, Incorporated...Local area
$124k - $186k
...impact.With interesting opportunities in engineering, marketing, sales, supply chain,... ...special for you.POSITION SUMMARYThe Test and Reliability Engineer is responsible for ensuring we... ...and departments.Experience working with site certifications such as ISO9001 and ISO14...Full timeWork experience placementLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- on-site clinical research associate (traveling/remote) Secaucus, NJ
- site safety Secaucus, NJ
- construction site safety Secaucus, NJ
- site reliability engineer
- lead site reliability engineer
- junior site reliability engineer
- site reliability engineering manager
- site reliability engineer remote
- site reliability engineer sre
- site paramedic


