Site Reliability Engineer (SRE)
Longfinch Technologies
Overview
We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.
The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.
Key Responsibilities
- Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
- Optimize, and support highly available VDI environments on Hyper-V.
- Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
- Disaster recovery, backup, patch management, and business continuity strategies.
- Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
- Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
- Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
- Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
- Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
- Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
- Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
- Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
- Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
- Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.
Experience & Qualifications
- 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
- Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
- Proven experience implementing automation to reduce operational overhead and improve service reliability.
- Experience supporting enterprise private cloud and VDI environments.
- Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
- Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
- Experience in Banking or Financial Services environments is advantageous.
Preferred Skills
- Windows Server 2016/2019/2022 administration.
- Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
- Exposure to hybrid cloud and private cloud platforms.
- Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
- Experience supporting enterprise VDI environments.
- Understanding of ITIL Incident, Problem, Change, and Release Management.
- Experience working in regulated industries such as Banking or Financial Services.
- ...Overview We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise...Suggested
- Site Reliability Engineer (SRE) Location: Englewood, NJ Work Arrangement: On-site Experience: 5+ Years Job Description We are looking for an experienced Site Reliability Engineer to support and improve the reliability, performance, and scalability...Suggested
$140k - $150k
WORK OPTION: Remote_________________The NBA is hiring a Senior Site Reliability Engineer (SRE) - Messaging & Collaboration to ensure the availability, performance, and reliability of enterprise messaging and collaboration platforms, including Microsoft Exchange Online...SuggestedFull timeTemporary workLocal areaRemote workWeekend work- ...We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern...SuggestedLocal area
- ServiceNow in Santa Clara, CA, is seeking a Senior Software Engineer - SRE & AIOps to strengthen reliability across hybrid cloud and data center operations. You will implement automation-first systems to reduce toil and accelerate incident remediation for global engineering...Suggested
- ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office... ...AI capabilities within the work environment to support SRE workflows with strong validation habits and awareness of data...Work at office
$60 - $67 per hour
Site Reliability Engineer Columbus, OH – onsite from day one Contract to Hire Onsite Interview 60% SRE & 40% Engineering Top skills: Public Cloud exposure, Terraform, CICD: Jenkins/ Drools, Java/ Python, Scripting for automation Required: ~Formal training or certification...Full timeContract workTemporary work$132.23k - $176.31k
...future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role...Full timeTemporary workRemote work- ...First Due seeks a Director of Platform Engineering to lead DevOps, SRE, and DBRE, turning fragmented practices into a centralized, disciplined... ...hands-on depth and executive presence to represent platform reliability to the ELT. Reporting to SVP of Engineering, this leader...
- ...significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Cloud Foundational... ...capabilities within the work environment to improve SRE workflows (e.g., incident investigation support and knowledge...
- ...Job Description Job Description BCforward is currently seeking a highly motivated SRE Software Engineer. Job Title: SRE Software Engineer Location: Jersey City, NJ Duration: Temp - 12 months Job Description We are seeking a Software Engineer-Other...Temporary work
- ...for partnering with leaders across engineering and technology to define objective reliability goals for services. Key... ...Position Summary: The Senior Azure Site Reliability Engineer acts as an advanced... ...production adoption. Mentor SRE engineers and raise the technical...Work at officeShift workDay shift
- LiveRamp, a data collaboration platform leader, is seeking a Senior Site Reliability Engineer in San Francisco with 5+ years of SRE/DevOps experience. The role focuses on deployments, 24/7 support across regions, and establishing SRE best practices. Strong skills in Terraform...
$253k - $336k
...TEAM: CorpTech Platform is the internal engineering force multiplier behind Anduril's... ...products. ABOUT THE JOB: The Director of Site Reliability Engineering owns the reliability system... ...manufacturing operations. This role leads the SRE organization and partners with software...Full timeWork experience placement$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively... ...using methodologies & tools such as XP, Lean, DevSecOps, SRE, ADO, GitHub, SonarQube, MLflow, and agentic AI frameworks (...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week$95k - $124k
...hybrid and will require 3 days on site at one of the following Quest... ...with Dynatrace3 plus years SRE experienceExperience in software... .../Azure/GCP CertificationsChaos Engineering CertificationsAgile CertificationsKnowledge: Site Reliability Engineering PrinciplesDevSecOps...Full timePart timeWork experience placementFlexible hours$174k - $252k
Senior Software Engineer, Site Reliability Engineering corporate_fare Google place Seattle, WA, USA ; Kirkland, WA, USA Mid Experience driving progress... ...Engineering. About the job Site Reliability Engineering (SRE) is what you get when you treat operations as if it’s a...Temporary work- Google is seeking a Senior Software Engineer in Site Reliability Engineering to strengthen the reliability of Google's public services from the Seattle/Kirkland area. You will design, build, and operate scalable systems with emphasis on availability and performance. The...
- Google is seeking a Software Engineering Manager for Site Reliability Engineering in Durham, NC. You will lead a team responsible for uptime, availability, and performance of key services, and drive automation to prevent problems. The role requires deep experience in programming...
- ...serve.The Information Technology group delivers secure, reliable technology solutions that enable DTCC to be the trusted infrastructure... ..., and performance of enterprise platforms.As a Principal Site Reliability Engineer (SRE), you will drive operational excellence across mission-...Remote workFlexible hours
$61k - $101k
...year Requirements: Formal training or certification in site reliability engineering concepts, with 5+ years of applied experience Formal training... ...enterprise system architecture, toil reduction, and other SRE best practices Strong programming skills in at least one...Full time$88k - $168k
Do you bring deep full-stack software engineering experience and the judgment to identify, design and deliver the right solutions? Are you... ...the globe. Your Mission As a Principal Software Engineer in our SRE & Architecture team within Tech, reporting to the Head of...Permanent employmentContract workTemporary workWork at officeLocal areaWork from homeWorldwideWeekend work- ...The selected colleague will work at an MUFG office or client sites four days per week and work remotely one day. A member of... ...MUFG is seeking a highly motivated Certified Sr. Cloud Site Reliability Engineer to build a robust, scalable, and reliable web application environment...Full timeWork at officeLocal areaRemote work
- 3Core Systems, Inc. is seeking a Senior Splunk/Data Management Engineer to design, implement, and optimize Splunk infrastructure to support scalable, robust data solutions. The role emphasizes SRE practices, incident response, and cross-functional leadership across multiple...
- ...Mount Thor, Inc. is seeking an on-site engineer to design and improve the data center systems behind our compute fleet. You will own the technical direction for data center architecture and operations, define standards for power, cooling, cabling, and network integration...
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer focussed on Operations Excellence, you will have the opportunity to shape how we respond to, learn from, and...
- ...contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology Team, you draw upon your advanced knowledge to identify new opportunities...
- ...overworked staff, and outdated infrastructure have compromised reliable access to essential medications. We are addressing these... ...on — and you set the standard for how infrastructure is built. Engineers write application code; you make sure it deploys reliably,...Full time
- ...world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Commercial Investment Banking... ...Kubernetes-based environments and AWS. You will apply SRE principles to drive measurable improvements in availability...
- Ford is seeking a Director of Cloud SRE to lead a global team focused on federating core SRE principles across hybrid environments... ...strategy, architecture, and roadmaps for internal observability and reliability tooling, spanning telemetry standards, pipeline integration,...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer (SRE). Be the first to apply!
- on-site clinical research associate (traveling/remote) Weehawken, NJ
- site reliability engineer
- site reliability engineering manager
- junior site reliability engineer
- site reliability engineer sre
- site reliability engineer remote
- lead site reliability engineer
- savannah river site
- site inspector
- website coordinator



