Site Reliability Engineer (SRE)
Longfinch Technologies
Overview
We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.
The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.
Key Responsibilities
- Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
- Optimize, and support highly available VDI environments on Hyper-V.
- Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
- Disaster recovery, backup, patch management, and business continuity strategies.
- Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
- Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
- Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
- Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
- Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
- Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
- Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
- Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
- Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
- Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.
Experience & Qualifications
- 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
- Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
- Proven experience implementing automation to reduce operational overhead and improve service reliability.
- Experience supporting enterprise private cloud and VDI environments.
- Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
- Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
- Experience in Banking or Financial Services environments is advantageous.
Preferred Skills
- Windows Server 2016/2019/2022 administration.
- Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
- Exposure to hybrid cloud and private cloud platforms.
- Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
- Experience supporting enterprise VDI environments.
- Understanding of ITIL Incident, Problem, Change, and Release Management.
- Experience working in regulated industries such as Banking or Financial Services.
- ...Overview We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise...Suggested
- ...First Due seeks a Director of Platform Engineering to lead DevOps, SRE, and DBRE, turning fragmented practices into a centralized, disciplined... ...hands-on depth and executive presence to represent platform reliability to the ELT. Reporting to SVP of Engineering, this...SuggestedRemote work
$55 - $60 per hour
...Overview We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise...SuggestedContract work$60 - $67 per hour
Site Reliability EngineerColumbus, OH - onsite from day oneContract to HireOnsite Interview 60% SRE & 40% EngineeringTop skills: Public Cloud exposure, Terraform, CICD: Jenkins/... ...training or certification in software engineering /Site Reliability Engineering concepts...SuggestedFull time- ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office... ...AI capabilities within the work environment to support SRE workflows with strong validation habits and awareness of data...SuggestedWork at office
$88k - $168k
Do you bring deep full-stack software engineering experience and the judgment to identify, design and deliver the right solutions? Are you... ...the globe. Your Mission As a Principal Software Engineer in our SRE & Architecture team within Tech, reporting to the Head of...Permanent employmentContract workTemporary workWork at officeLocal areaWork from homeWorldwideWeekend work- ...the world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Asset and Wealth Management... ...and 3 years applied experience Foundational understanding of SRE culture and principles, including Service Level Indicators (...Work experience placement
- ...Principal Site Reliability Engineer (SRE)Are you ready to make an impact at DTCC? Do you want to work on innovative projects, collaborate with a dynamic and supportive team, and receive investment in your professional development? At DTCC, we are at the forefront of innovation...
$116.25k - $155k
...Ops Strategy serves as the technical navigator and quantitative engine within the Global Operations Readiness team. Reporting into... ...years with a Master's degree). Demonstrated experience in high-reliability, high-tech hardware, server assembly, electronics manufacturing...Permanent employmentContract workTemporary workWork at officeLocal areaShift workDay shift- ...Site Reliability Engineer III There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your... ...experience Exposure to or hands-on experience in supporting SRE practices for Data management/migration platforms and...
- ...most complex and mission-critical systems. As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Commercial Investment Banking team... ...Kubernetes-based environments and AWS. You will apply SRE principles to drive measurable improvements in availability...
- ...significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Cloud Foundational... ...capabilities within the work environment to improve SRE workflows (e.g., incident investigation support and knowledge...
- ...world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Commercial Investment Banking... ...Kubernetes-based environments and AWS. You will apply SRE principles to drive measurable improvements in availability...
- ...Job Description Job Description BCforward is currently seeking a highly motivated SRE Software Engineer. Job Title: SRE Software Engineer Location: Jersey City, NJ Duration: Temp - 12 months Job Description We are seeking a Software Engineer-Other...Temporary work
- ...for partnering with leaders across engineering and technology to define objective reliability goals for services. Key... ...Position Summary: The Senior Site Reliability Engineer acts as an advanced... ...they support Champions modern SRE, observability, and AIOps practices...Work at officeShift workDay shift
- ...help healthcare organizations streamline operations through reliable, customizable technology that improves efficiency and... ...the Labgen environment . We are looking for a hands-on Site Reliability Engineer / Infrastructure Engineer to own the reliability, security...Contract workWork at officeTrial period
- ...help healthcare organizations streamline operations through reliable, customizable technology that improves efficiency and... ...the Medgen environment . We are looking for a hands-on Site Reliability Engineer / Infrastructure Engineer to own the reliability, security...Work at officeTrial period
- ...serve.The Information Technology group delivers secure, reliable technology solutions that enable DTCC to be the trusted infrastructure... ..., and performance of enterprise platforms.As a Principal Site Reliability Engineer (SRE), you will drive operational excellence across mission-...Remote workFlexible hours
$61k - $101k
...year Requirements: Formal training or certification in site reliability engineering concepts, with 5+ years of applied experience Formal training... ...enterprise system architecture, toil reduction, and other SRE best practices Strong programming skills in at least one...Full time- ...The selected colleague will work at an MUFG office or client sites four days per week and work remotely one day. A member of... ...MUFG is seeking a highly motivated Certified Sr. Cloud Site Reliability Engineer to build a robust, scalable, and reliable web application environment...Full timeWork at officeLocal areaRemote work
- ...best work together.What's the role?We are looking for a Software Engineer to help improve Etsy's development workflows and environments.... ...toolingWe work closely with other infrastructure, product, and SRE teams to ensure consistency between development and productionWe...Full timeWork at officeLocal areaRemote workVisa sponsorship
- ...contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology Team, you draw upon your advanced knowledge to identify new opportunities...
- ...appropriate Collaborates with other software engineers and teams to design and implement... ..., test, and implement availability, reliability, scalability, and solutions in their applications... ...customers Supports the adoption of site reliability engineering best practices...
- Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability.As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Chief Data & Analytics...Work at office
- ...and forward-thinking Senior DevOps Engineer to drive the modernization, scalability... ...GitHub) to ensure fast, secure, and reliable software deployments.Architect and... ...8+ years of experience in a DevOps, Site Reliability Engineering (SRE), or Systems Engineering role.Hands-...Permanent employmentFull timeWork at officeShift workNight shiftWeekend workAfternoon shift
- ...We are seeking a Secure Enterprise Browser DevOps Engineer with strong production support and reliability engineering experience. The role focuses on maintaining... ...for enterprise environments. Collaborate with SRE, security, engineering, and incident response teams...Permanent employment
- ...As an Enterprise Browser DevOps Engineer, you will be responsible for the reliability, availability, and performance of the firm's secure enterprise browser platform... ...share production ownership and on-call duties with SRE and incident response teams, removing single points...Permanent employment
$90.3k - $189.6k
...Job Title: Release Train Engineer Job Category: Information Technology Time Type: Full time Minimum Clearance Required to Start: Secret Employee Type: Regular Percentage of Travel Required: Up to 10% Type of Travel: Local * * * **The Opportunity:...Full timeContract workWork experience placementLocal areaFlexible hours- ...reporting to the Chief Mechanical Officer. This position serves as the lead authority for fleet maintenance, technical compliance, reliability, and rolling stock performance for the agency. The role develops, implements, and oversees maintenance programs, inspection...
$265k - $310k
...around slow iteration speed and memory safety by building Arc and Dia using Swift. A small team of language compiler and systems engineers have implemented the protocol that allows us to run our Swift code across MacOS, Windows, iOS and Android. You can read more about...Full timeLive inWork at officeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer (SRE). Be the first to apply!


