Site Reliability Engineer (SRE)
Longfinch Technologies
Overview
We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.
The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.
Key Responsibilities
- Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
- Optimize, and support highly available VDI environments on Hyper-V.
- Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
- Disaster recovery, backup, patch management, and business continuity strategies.
- Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
- Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
- Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
- Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
- Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
- Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
- Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
- Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
- Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
- Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.
Experience & Qualifications
- 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
- Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
- Proven experience implementing automation to reduce operational overhead and improve service reliability.
- Experience supporting enterprise private cloud and VDI environments.
- Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
- Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
- Experience in Banking or Financial Services environments is advantageous.
Preferred Skills
- Windows Server 2016/2019/2022 administration.
- Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
- Exposure to hybrid cloud and private cloud platforms.
- Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
- Experience supporting enterprise VDI environments.
- Understanding of ITIL Incident, Problem, Change, and Release Management.
- Experience working in regulated industries such as Banking or Financial Services.
- ...Overview We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise...Suggested
- Site Reliability Engineer (SRE) Location: Englewood, NJ Work Arrangement: On-site Experience: 5+ Years Job Description We are looking for an experienced Site Reliability Engineer to support and improve the reliability, performance, and scalability...Suggested
- ...We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern...SuggestedLocal area
- ...the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office... ...AI capabilities within the work environment to support SRE workflows with strong validation habits and awareness of data...SuggestedWork at office
$140k - $150k
WORK OPTION: Remote_________________The NBA is hiring a Senior Site Reliability Engineer (SRE) - Messaging & Collaboration to ensure the availability, performance, and reliability of enterprise messaging and collaboration platforms, including Microsoft Exchange Online...SuggestedFull timeTemporary workLocal areaRemote workWeekend work$60 - $67 per hour
Site Reliability Engineer Columbus, OH – onsite from day one Contract to Hire Onsite Interview 60% SRE & 40% Engineering Top skills: Public Cloud exposure, Terraform, CICD: Jenkins/ Drools, Java/ Python, Scripting for automation Required: ~Formal training or certification...Full timeContract workTemporary work- ...significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Cloud Foundational... ...capabilities within the work environment to improve SRE workflows (e.g., incident investigation support and knowledge...
- ...Job Description Job Description BCforward is currently seeking a highly motivated SRE Software Engineer. Job Title: SRE Software Engineer Location: Jersey City, NJ Duration: Temp - 12 months Job Description We are seeking a Software Engineer-Other...Temporary work
- ...for partnering with leaders across engineering and technology to define objective reliability goals for services. Key... ...Position Summary: The Senior Azure Site Reliability Engineer acts as an advanced... ...production adoption. Mentor SRE engineers and raise the technical...Work at officeShift workDay shift
$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively... ...using methodologies & tools such as XP, Lean, DevSecOps, SRE, ADO, GitHub, SonarQube, MLflow, and agentic AI frameworks (...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week- ...Intellectt Inc. Please find the below requriement and let me know if you are interested in applying for this role. Job Title: Site Reliability Engineer Location: Englewood, NJ Work Arrangement: On-site Pay Rate: $56/hr on W2 or $65/hr ON c2c Note: Need local...Local area
- ...serve.The Information Technology group delivers secure, reliable technology solutions that enable DTCC to be the trusted infrastructure... ..., and performance of enterprise platforms.As a Principal Site Reliability Engineer (SRE), you will drive operational excellence across mission-...Remote workFlexible hours
$61k - $101k
...year Requirements: Formal training or certification in site reliability engineering concepts, with 5+ years of applied experience Formal training... ...enterprise system architecture, toil reduction, and other SRE best practices Strong programming skills in at least one...Full time$132.23k - $176.31k
...future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role...Full timeTemporary workRemote work- ...The selected colleague will work at an MUFG office or client sites four days per week and work remotely one day. A member of... ...MUFG is seeking a highly motivated Certified Sr. Cloud Site Reliability Engineer to build a robust, scalable, and reliable web application environment...Full timeWork at officeLocal areaRemote work
- ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer focussed on Operations Excellence, you will have the opportunity to shape how we respond to, learn from, and...
- ...contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology Team, you draw upon your advanced knowledge to identify new opportunities...
- ...world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Commercial Investment Banking... ...Kubernetes-based environments and AWS. You will apply SRE principles to drive measurable improvements in availability...
- ...world's most complex and mission-critical systems. As a Site Reliability Engineer III at JPMorgan Chase within the Asset and Wealth Management... ...+ years applied experience Foundational understanding of SRE culture and principles, including Service Level Indicators (...Work experience placement
$95k - $124k
...hybrid and will require 3 days on site at one of the following Quest... ...with Dynatrace3 plus years SRE experienceExperience in software... .../Azure/GCP CertificationsChaos Engineering CertificationsAgile CertificationsKnowledge: Site Reliability Engineering PrinciplesDevSecOps...Full timePart timeWork experience placementFlexible hours- Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability.As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Chief Data & Analytics...Work at office
- ...Job Description Versant’s Software Engineering team provides core services to our business... .... Drive improvements in latency, reliability, and scalability of streaming pipelines.... ...through data-driven insights. Partner with SRE and data teams to build unified...Local area
- ...opportunity for you to take your software engineering career to the next level. The Chief Data... ...operational support for the platform to SRE and app teams.Performs platform design, set... ...comprehensive health care coverage, on-site health and wellness centers, a retirement...Work at office
- ...focused on improving the security, release reliability, maintainability, and operational... ...critical banking applications.Forward Deployed Engineers will conduct targeted, hands-on... ...combines senior software engineering, DevOps, SRE, test automation, application security,...
- ...effectively and responsibly.As a Lead Software Engineer at JPMorganChase within the Chief Data... ...operational support for the platform to SRE and app teams. Performs platform design, set... ...comprehensive health care coverage, on-site health and wellness centers, a retirement...Work at office
- ...opportunity for you to take your software engineering career to the next level. As a Software... ...architectureFollow best practices in SRE, DevOps, and observability to enhance... ...skillsFormal training or certification on site reliability engineering concepts and 5+ years...
$130k - $170k
...continuously leverage advances in AI technologies for the benefit of the business.What We're Looking ForWe are looking for a Software Engineer II to focus on solutions for our AI Transformation team. You will work in-office at least three days per week at our Fort Lee, New...Temporary workWork at office3 days per week- ...Tomatoes, GolfNow and GolfPass. Job Description The Cloud Reliability Engineer is responsible for ensuring the availability, performance,... ...practical experience. ~3–7 years of experience in Site Reliability Engineering, Cloud Engineering, DevOps, Infrastructure...Work at officeLocal area
- ...focused on improving the security, release reliability, maintainability, and operational... ...critical banking applications.Forward Deployed Engineers will conduct targeted, hands-on... ...combines senior software engineering, DevOps, SRE, test automation, application security,...Contract work
- We are looking for a skilled DevOps Engineer 2 to support the implementation,... ...deployments, and contribute to the reliability, security, and performance of production... ...in DevOps, Cloud Engineering, Site Reliability Engineering (SRE), Infrastructure Engineering, or a...Flexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer (SRE). Be the first to apply!
- on-site clinical research associate (traveling/remote) Englewood, NJ
- site reliability engineer
- site reliability engineering manager
- junior site reliability engineer
- site reliability engineer sre
- site reliability engineer remote
- lead site reliability engineer
- savannah river site
- site inspector
- website coordinator



