Site Reliability Engineer (SRE)
Longfinch Technologies
Overview
We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.
The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.
Key Responsibilities
- Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
- Optimize, and support highly available VDI environments on Hyper-V.
- Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
- Disaster recovery, backup, patch management, and business continuity strategies.
- Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
- Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
- Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
- Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
- Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
- Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
- Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
- Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
- Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
- Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.
Experience & Qualifications
- 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
- Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
- Proven experience implementing automation to reduce operational overhead and improve service reliability.
- Experience supporting enterprise private cloud and VDI environments.
- Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
- Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
- Experience in Banking or Financial Services environments is advantageous.
Preferred Skills
- Windows Server 2016/2019/2022 administration.
- Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
- Exposure to hybrid cloud and private cloud platforms.
- Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
- Experience supporting enterprise VDI environments.
- Understanding of ITIL Incident, Problem, Change, and Release Management.
- Experience working in regulated industries such as Banking or Financial Services.
- Site Reliability Engineer (SRE) Location: Englewood, NJ Work Arrangement: On-site Experience: 5+ Years Job Description We are looking for an experienced Site Reliability Engineer to support and improve the reliability, performance, and scalability...Suggested
$132.23k - $176.31k
...future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role...SuggestedFull timeTemporary workRemote work$95k - $124k
...hybrid and will require 3 days on site at one of the following Quest... ...with Dynatrace3 plus years SRE experienceExperience in software... .../Azure/GCP CertificationsChaos Engineering CertificationsAgile CertificationsKnowledge: Site Reliability Engineering PrinciplesDevSecOps...SuggestedFull timePart timeWork experience placementFlexible hours- ...As an Enterprise Browser DevOps Engineer, you will be responsible for the reliability, availability, and performance of the firm's secure enterprise browser platform... ...share production ownership and on-call duties with SRE and incident response teams, removing single points...SuggestedPermanent employment
- ...We are seeking a Secure Enterprise Browser DevOps Engineer with strong production support and reliability engineering experience. The role focuses on maintaining... ...for enterprise environments. Collaborate with SRE, security, engineering, and incident response teams...SuggestedPermanent employment
- ...Position Description We are seeking an experienced Senior Software Engineer in our Core Brokerage group, who will be responsible for the... ...solutioning. Responsibilities Design and implement highly reliable, scale‑able, extensible, maintainable, global, and operable products...Work at office
$100k
...Senior Software Engineer Application Deadline: 7 September 2026 Department: Tech 3PL - R&D - Dev Employment Type: Full Time... ...only way Solid understanding of how to develop responsive sites which are cross-browser and device compatible, using modern web...Full timeLocal areaRemote work- ...platform ecosystem. This role carries a dual mandate: ensuring the reliability, security, and continuity of production systems that underpin... ...bizhub One i-Series. Pour en savoir plus, rendez-vous sur le site de Konica Minolta et suivez l’entreprise sur Facebook, YouTube,...For contractorsWork at office
- ...We are seeking a Python Agentic AI Engineer to design, build, and deploy scalable AI agent and backend solutions. The role combines Python, agentic AI, LLMs, RAG, AWS, APIs, microservices, and event-driven architectures to deliver production-ready AI applications. Roles...
- ...Senior Software Engineer – Full-Time Direct Employment Cognizant is seeking an experienced Senior Software Engineer for a full-time, long-term remote position. This position is being posted by Cognizant for direct employment with the company. It is not a staffing-...Full timeRemote work
- ...Job Description Versant’s Software Engineering team provides core services to our business... .... Drive improvements in latency, reliability, and scalability of streaming pipelines.... ...through data-driven insights. Partner with SRE and data teams to build unified...Local area
- A technology services provider based in New Jersey is seeking a skilled technician for system integration responsibilities. The role involves connecting various security components, installing cabling systems, and customizing system functionalities. Candidates should have...
$125k
This position has an On-Site Requirement in Teaneck NJ Interstate Waste Services is the most progressive and innovative provider of solid waste and recycling services in the greater New York, New Jersey and Connecticut markets with a rail-served landfill in Ohio. IWS...Local area- A tech consulting company based in New Jersey is seeking a highly experienced Senior Software Engineer to join its Core Brokerage group. The ideal candidate will have over 10 years of software development experience, strong proficiency in Go or Java, and excellent problem...
- ALL US VISA'S ARE ACCEPTED NEED ONLY LOCAL PROFILES TO NY / NJ WHO CAN GO FOR AN IN-PERSON INTERVIEW Below are the must-haves: Experience: 10-12 years of hands-on minimum Location: NYC, In-Office (3 days a week, Mon, Tue, Wed) Should be comfortable with using...Work at officeLocal area3 days per week
$70k - $130k
...of industries and companies from startups to Fortune 500 organizations. If this sounds exciting to you, apply to be an Application Engineer to be part of something special!Your role will be to strategically identify and develop critical thin film processes and new...$227k - $303k
...developer-facing capabilities that enable every engineer at CoreWeave to build and ship software... ..., heterogeneous workloads, and the reliability demands of an infrastructure platform... ...Partner with infrastructure, security, SRE, and product engineering teams to ensure...Permanent employmentFull timeTemporary workCasual workWork at officeRemote workFlexible hours- ...experience with the future. To learn more aboutour solutions and innovations, visit our website here.Role PurposeAs Lead Systems Engineer, you'll hold the highest level of technical ownership on our Systems Engineering track joining a dispersed, cross-disciplinary team...Full timeLocal area
$130k - $140k
...least 5 years of experience in systems administration, software engineering, or IT support, with a minimum of 3 years directly managing and... ...corporate IT office environment. May require occasional on-site visits to operational facilities during hardware rollouts.Required...Full timePart timeWork experience placementWork at officeWork from homeFlexible hours$95.2k - $112k
Job TitleReliability Engineer- Production MaintenanceJob Description SummaryAs a Reliability Engineer, you will support maintenance operations at a high-volume grocery... ...implementation of permanent corrective action at the site. Identify opportunities to increase equipment...Minimum wagePermanent employmentFull timeContract workTemporary workFor contractorsFor subcontractorLocal areaFlexible hours- ...Reliability EngineerLocation: Bronx, NY, US, 10457 At Perrigo, we are driven by our mission to Makes Lives Better Through Trusted... ...to win in self-care.Description Overview The Reliability Engineer serves as the site reliability leader and technical subject matter expert...For contractors
- ...AMD Rack Level Reliability EngineerAdvance your career. Advance the world.At AMD, we believe technology can change lives for the better... ...reliability testing, coupled with a passion for reliability engineering principles. Excellent communication and collaboration skills...
- We are looking for an Analytics Application Engineer to support and advance an internal reporting platform serving teams in River Edge... ...engineering, data integration, and analytics development to deliver reliable dashboards and actionable operational insights. The ideal...
- Job Description Job Description Benefits: ~401(k) matching ~ Dental insurance ~ Health insurance Core Responsibilities System Integration: Connecting and configuring individual components (e.g., cameras, motion sensors, and access control readers) so...
- Our team is seeking a Senior Software Engineer to play a crucial role in advancing our core healthcare IT systems through the development... ...-scale healthcare data exchange systemsDeliver extremely reliable, scalable, and observable software capable of supporting massive...Permanent employment
- OverviewAs the Sr. Software Engineer-Front End, your primary responsibility is to build systems and functionality that support the Vitamin... ...tools (grunt, gulp, node.js)Make continuous improvements to site performance and SEOCreating self-contained, reusable, and testable...Local areaFlexible hours
- ...Wehmiller Companies Inc. is seeking a Manager, Controls System Integration in Nanuet, NY. In this full-time role, you will lead a team of engineers through multiple system integration projects, overseeing project life cycles from conception to start-up. A minimum of nine years...Full time
- ...we are, join our team.KPMG is currently seeking a Sr. Software Engineer to join our Tax Ignition Team. This is a hybrid work opportunity... ...benefits can be found towards the bottom of our KPMG US Careers site at Benefits & How We Work.Follow this link to obtain salary ranges...Local area
$130k - $170k
...continuously leverage advances in AI technologies for the benefit of the business.What We're Looking ForWe are looking for a Software Engineer II to focus on solutions for our AI Transformation team. You will work in-office at least three days per week at our Fort Lee, New...Temporary workWork at office3 days per week- ...Technical Skills ~7–8+ years of experience in DevOps, Cloud Engineering, or related roles. ~ Strong hands-on experience with Azure... ...automation frameworks. Implement DevOps best practices for reliability, security, scalability, and operational efficiency....
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer (SRE). Be the first to apply!



