Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

2T Consulting

We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.

The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.

Required Technical Skills
  • Strong understanding of Site Reliability Engineering principles and operational excellence.
  • Experience with infrastructure reliability, service availability, resiliency, and performance optimization.
  • Storage Space Direct and failover clustering technical expertise. (Storage Spaces Direct enables you to build highly available, software-defined storage by pooling local disks (SSDs, NVMe drives, and HDDs) across multiple Windows Server nodes in a cluster. Instead of relying on an external SAN, S2D uses the servers' local storage to create a resilient shared storage pool)
  • Experience managing production-critical infrastructure environments with high availability requirements.
  • Experience with incident management, problem management, RCA, and continuous operational improvement.
  • Knowledge of monitoring, observability, alerting, and performance management.
Microsoft Hyper-V (Core Expertise)
  • Deep hands-on expertise in Microsoft Hyper-V architecture, deployment, administration, troubleshooting, and optimization.
  • Extensive experience in operating enterprise private cloud environments on Hyper-V.
  • Strong experience supporting enterprise-scale VDI deployments on Hyper-V.
  • Hyper-V Failover Clustering and high-availability architecture.
  • Storage integration including SAN, NAS, Storage Spaces Direct (S2D), Cluster Shared Volumes (CSV), and storage optimization.
  • Networking within Hyper-V environments including virtual switches, VLANs, NIC Teaming, QoS, and network performance tuning.
  • System Center Virtual Machine Manager (SCVMM).
Automation & Platform Engineering
  • Strong PowerShell scripting and automation experience.
  • Experience automating infrastructure deployment, operational tasks, health checks, and reporting.
  • Familiarity with Infrastructure as Code concepts and configuration management.
  • Experience developing reusable operational tooling to improve reliability and reduce manual effort.
Preferred Skills
  • Windows Server 2016/2019/2022 administration.
  • Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
  • Exposure to hybrid cloud and private cloud platforms.
  • Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
  • Experience supporting enterprise VDI environments.
  • Understanding of ITIL Incident, Problem, Change, and Release Management.
  • Experience working in regulated industries such as Banking or Financial Services.
Experience & Qualifications
  • 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
  • Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
  • Proven experience implementing automation to reduce operational overhead and improve service reliability.
  • Experience supporting enterprise private cloud and VDI environments.
  • Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
  • Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
  • Experience in Banking or Financial Services environments is advantageous.
Key Responsibilities
  • Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
  • Optimize, and support highly available VDI environments on Hyper-V.
  • Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
  • Disaster recovery, backup, patch management, and business continuity strategies.
  • Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
  • Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
  • Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
  • Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
  • Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
  • Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
  • Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
  • Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
  • Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
  • Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Corona, NY vacancy
  •  ...We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern... 
    Suggested
    Local area

    2T Consulting

    Ozone Park, NY
    5 days ago
  •  ...Overview We are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise... 
    Suggested

    Longfinch Technologies

    Astoria, NY
    4 days ago
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Chief Data & Analytics Office (CDAO) AI/ML & Data Platforms team,, you will solve complex... 
    Suggested
    Work at office

    JP Morgan Chase

    Jersey City, NJ
    1 day ago
  • $132.23k - $176.31k

     ...future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Senior Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role... 
    Suggested
    Full time
    Temporary work
    Remote work

    Lumen

    Long Island City, NY
    9 hours ago
  •  ...The selected colleague will work at an MUFG office or client sites four days per week and work remotely one day. A member of...  ...MUFG is seeking a highly motivated Certified Sr. Cloud Site Reliability Engineer to build a robust, scalable, and reliable web application environment... 
    Suggested
    Full time
    Work at office
    Local area
    Remote work

    MUFG

    Jersey City, NJ
    3 days ago
  • $83.54k - $137.24k

     ...progress and enhancing lives by providing reliable, high-speed connectivity solutions that...  ...you! Job Summary The Role DNS Engineer – SRE is a high-impact role responsible for...  ...views infrastructure through the lens of Site Reliability Engineering (SRE)... 
    Local area
    Remote work

    Optimum Communications Corp

    Bethpage, NY
    2 days ago
  •  ...Mount Thor, Inc. is seeking an on-site engineer to design and improve the data center systems behind our compute fleet. You will own the technical direction for data center architecture and operations, define standards for power, cooling, cabling, and network integration... 

    Mount Thor, Inc.

    Brooklyn, NY
    3 days ago
  •  ...First Due seeks a Director of Platform Engineering to lead DevOps, SRE, and DBRE, turning fragmented practices into a centralized, disciplined...  ...hands-on depth and executive presence to represent platform reliability to the ELT. Reporting to SVP of Engineering, this leader will... 

    First Due

    Brooklyn, NY
    4 days ago
  •  ...core SRE principles across hybrid environments. You will shape strategy, architecture, and roadmaps for internal observability and reliability tooling, spanning telemetry standards, pipeline integration, and SRE maturity. The role blends hands-on delivery with leadership,... 

    Ford Motor Company

    Brooklyn, NY
    5 days ago
  •  ...Principal Site Reliability Engineer (SRE)Are you ready to make an impact at DTCC? Do you want to work on innovative projects, collaborate with a dynamic and supportive team, and receive investment in your professional development? At DTCC, we are at the forefront of innovation... 

    Dtcc

    Jersey City, NJ
    5 days ago
  • $152.6k - $191.5k

     ...is responsible for partnering with leaders across engineering and technology to define objective reliability goals for services. Key responsibilities include composing...  ...improvement. Position Summary: The Senior Site Reliability Engineer acts as an advanced senior... 
    Full time
    Work at office
    Shift work
    Day shift

    Bank of America

    Jersey City, NJ
    5 days ago
  •  ...contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology Team, you draw upon your advanced knowledge to identify new opportunities... 

    JP Morgan Chase

    Jersey City, NJ
    9 hours ago
  •  ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Cloud Foundational Services team, you hold a leadership role in your team, demonstrate... 

    JP Morgan Chase

    Jersey City, NJ
    9 hours ago
  •  ...globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer focussed on Operations Excellence, you will have the opportunity to shape how we respond to, learn from, and... 

    JP Morgan Chase

    Jersey City, NJ
    1 day ago
  •  ...Job Description Job Description BCforward is currently seeking a highly motivated SRE Software Engineer. Job Title: SRE Software Engineer Location: Jersey City, NJ Duration: Temp - 12 months   Job Description We are seeking a  Software Engineer-Other... 
    Temporary work

    BCForward

    Jersey City, NJ
    12 days ago
  •  ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.  As a Site Reliability Engineer III at JPMorgan Chase within the Commercial Investment Banking team of Fraud Prevention , you will solve complex and broad... 

    JPMorgan Chase & Co.

    Jersey City, NJ
    4 days ago
  • LiveRamp, a data collaboration platform leader, is seeking a Senior Site Reliability Engineer in San Francisco with 5+ years of SRE/DevOps experience. The role focuses on deployments, 24/7 support across regions, and establishing SRE best practices. Strong skills in Terraform... 

    FinOps Weekly

    Brooklyn, NY
    1 day ago
  • $253k - $336k

     ...not years. ABOUT THE TEAM: CorpTech Platform is the internal engineering force multiplier behind Anduril's corporate systems. It...  ...its mission critical products. ABOUT THE JOB: The Director of Site Reliability Engineering owns the reliability system for the software that... 
    Full time
    Work experience placement

    AI Chopping Block

    Brooklyn, NY
    1 day ago
  • $113.1k - $232.3k

    Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied AI Site Reliability Engineer II, you will actively engage in your engineering craft, taking a hands-on approach to the reliability, performance, and operational integrity... 
    Work at office
    Local area
    Visa sponsorship
    Flexible hours
    3 days per week

    Deloitte

    Jersey City, NJ
    4 days ago
  • $174k - $252k

    Senior Software Engineer, Site Reliability Engineering corporate_fare Google place Seattle, WA, USA ; Kirkland, WA, USA Mid Experience driving progress, solving problems, and mentoring more junior team members; deeper expertise and applied knowledge within relevant area... 
    Temporary work

    Google Inc.

    Brooklyn, NY
    2 days ago
  • Elevate your engineering prowess to unprecedented levels by joining a team of exceptionally gifted professionals and position yourself among the top echelon in site reliability.As a Senior Lead Site Reliability Engineer at JPMorgan Chase within the Chief Data & Analytics... 
    Work at office

    JP Morgan Chase

    Jersey City, NJ
    1 day ago
  • $61k - $101k

     ...Salary: $61,000 - 101,000 per year Requirements: Formal training or certification in site reliability engineering concepts, with 5+ years of applied experience Formal training or certification in site reliability engineering concepts, with advanced practical experience... 
    Full time

    J.P. Morgan

    Jersey City, NJ
    3 days ago
  •  ...Job Description Job Description The primary responsibility of the System Engineer is to design, build, and maintain the Bank’s infrastructure such as internal systems, network and various banking-related applications.   ESSENTIAL FUNCTIONS AND RESPONSIBILITIES... 
    Bank staff
    Local area

    Preferred Bank

    Flushing, NY
    7 days ago
  • $119.3k - $196.6k

     ...and BALMAIN Beauty.DescriptionThe Senior Lead, Finance Platform Engineering provides engineering leadership and is accountable for the...  ...strategic technology partners to deliver secure, scalable, and reliable Finance capabilities that support global business operations.This... 
    Full time
    Local area

    Estée Lauder

    Long Island City, NY
    1 day ago
  •  ...applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.  As a Site Reliability Engineer III at JPMorgan Chase within the Asset and Wealth Management team, you will solve complex and broad business problems... 
    Work experience placement

    JPMorgan Chase & Co.

    Jersey City, NJ
    7 days ago
  • $95.6k - $138.7k

     ...this job involves:The Critical Systems Engineer provides onsite field coordination, supervision...  ...the highest standards of operational reliability and safety compliance.Your day-to-day...  ...facility operationsLocation & Schedule:On-site position with 24/7 on-call availability... 
    Full time
    For contractors
    For subcontractor
    Local area

    Jones Lang LaSalle

    Flushing, NY
    4 days ago
  • Job Title: Lead Software Engineer - Full StackAbout the OpportunityA rapidly growing life sciences...  ....Key ResponsibilitiesDevelop scalable, reliable, testable, and maintainable software...  ...York City area.Must be willing to work on-site in Long Island City, NY.Preferred... 
    Live in
    Relocation

    Career Development Partners

    Long Island City, NY
    4 days ago
  • $60 - $67 per hour

    Site Reliability EngineerColumbus, OH - onsite from day oneContract to HireOnsite Interview 60% SRE & 40% EngineeringTop skills: Public Cloud...  ...Required:Formal training or certification in software engineering /Site Reliability Engineering concepts and 6+ years of applied... 
    Full time

    Pinnacle Group

    Jersey City, NJ
    2 days ago
  • $105.6k - $150.4k

    Position SummaryThe Lead Engineer Full Stack leads development teams in designing, testing, and implementing responsive web applications. In addition to leading and delegating work to other developers, the Lead Engineer will be responsible for standardization of documentation... 
    Temporary work
    Work at office
    Immediate start
    Flexible hours
    Night shift

    JetBlue Airways

    Long Island City, NY
    3 days ago
  • $120k - $180k

    Lead Software Engineer, Full StackLong Island City, NYCompensation: $120,000 to $180,000 per...  ...the industry.ScheduleThis is an on-site position based in Long Island City, NY. Candidates...  ...to day, this person develops scalable, reliable, testable, and maintainable software... 
    Live in
    Relocation

    iLocatum Recruiting

    Long Island City, NY
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!