Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer - Remote

DivIHN Integration

For further inquiries about this opportunity, please contact our Talent Specialist, Marshelin at View phone number on us.fitly.work

Title: Site Reliability Engineer - Remote
Location: Corning, NY -
Remote
Duration: 6 Months with possible extension based on business need
Schedule: Candidates can live anywhere in the US but must be able to work 8 AM 5 PM/9 AM 6 PM EST

Only W2 candidates are eligible for this position. Third-party or C2C candidates will not be considered.

Job Description:

Primary Purpose of the Role

The Site Reliability Engineer will support and manage Clients Kubernetes platform infrastructure that supports scientific and engineering applications.

Role Overview

  • Join Client's Model Operations and Deployment Engineering team at their flagship research facility, where your work will directly support groundbreaking materials science innovations. In this role, you will help maintain, enhance, and evolve the Kubernetes platforms that enable scientific and engineering teams to deploy, operate, and scale critical applications across both on-premises and cloud environments.
  • Client is looking for an experienced contract Site Reliability Engineer who can strengthen our team's platform engineering and operational capabilities. You will play a key role in supporting Kubernetes infrastructure managed through Rancher, improving system reliability and automation, and advancing infrastructure-as-code and GitOps practices across our environment.

This role focuses on:

  • Managing and maintaining Kubernetes platforms
  • Ensuring platform reliability, stability, and uptime
  • Supporting both on-premises and cloud-based Kubernetes environments
  • Managing infrastructure that hosts critical business and research applications
  • Supporting platform operations rather than application development
  • Helping the team expand its cloud migration initiatives
  • This is NOT a Software Developer role.

Key Responsibilities

  • Platform Operations: Maintain and enhance Kubernetes platforms across on-premises and cloud environments, ensuring reliability, scalability, and operational efficiency.
  • Cluster Management: Support provisioning, upgrades, troubleshooting, and lifecycle management of Kubernetes clusters managed through Rancher.
  • Linux Systems Administration: Provide deep technical expertise in Linux-based systems, including performance tuning, troubleshooting, automation, and operational support.
  • Infrastructure as Code: Develop and maintain infrastructure-as-code solutions to standardize and automate platform deployment and management, with a preference for Cluster API (CAPI)-based approaches.
  • GitOps and Deployment Automation: Support and improve GitOps workflows using ArgoCD to manage cluster and application configuration in a consistent, auditable manner.
  • Collaboration: Work closely with developers, scientists, and infrastructure teams to deliver reliable platform services and translate operational needs into sustainable engineering solutions.
  • Continuous Improvement: Identify opportunities to improve platform resilience, observability, security, and maintainability through automation and modern SRE practices.

Top requirements:

  • Bachelors is preferred, but not required.
  • Minimum of 5 years professional experience in site reliability engineering, platform engineering, DevOps, or systems engineering roles.
  • Candidates must have 5 years strong system admininstration with Linux! Rancher for Kubernetes experience is a must.

Top Required Skills (Must-Have)

1. Linux Administration (Highest Priority)

Experience Level

  • Preferred: 10+ years Linux experience
  • Acceptable: Minimum 5+ years strong Linux administration experience

2. Kubernetes Administration

Experience

  • Minimum 3-5 years Kubernetes administration
  • Production environment experience required

3. Rancher

4. Infrastructure as Code

5. GitOps

Qualifications Required:

Education prefrred :

  • BS in Computer Science, Software Engineering, Information Technology, or related field preferred; or equivalent professional experience.

Experience:

  • 5+ years of professional experience in site reliability engineering, platform engineering, DevOps, or systems engineering roles.
  • Hands-on experience operating and supporting Kubernetes platforms in production environments.
  • Strong experience managing Kubernetes clusters in both on-premises and cloud-based environments.
  • Strong Linux systems administration skills, including troubleshooting, scripting, networking, and system performance analysis.
  • Experience with Rancher for Kubernetes cluster management and platform operations.
  • Experience implementing infrastructure-as-code solutions for platform provisioning and lifecycle management.
  • Demonstrated success working in Agile teams (Scrum, Kanban).

Technical Skills:

  • Kubernetes: Cluster operations, upgrades, networking, storage, troubleshooting, and workload support.
  • Platform Management: Rancher or similar Kubernetes management platforms.
  • Linux: Advanced administration of Linux/Unix systems.
  • Infrastructure as Code: Strong IaC experience; Cluster API (CAPI) preferred.
  • GitOps/CI-CD: ArgoCD, Git version control, and deployment automation practices.
  • Scripting/Automation: Bash, Python, or similar scripting languages for automation and operational tooling.

Preferred:

  • Experience with hybrid infrastructure spanning on-premises and public cloud platforms (AWS, Azure, GCP).
  • Experience with Kubernetes ecosystem tooling for observability, logging, monitoring, and alerting.
  • Familiarity with security best practices for Kubernetes and Linux platforms.
  • Experience supporting scientific research environments, high-performance computing, or computational science workflows.
  • Knowledge of CI/CD pipeline development and platform automation patterns.

Certifications Preferred

  • Certified Kubernetes Administrator (CKA)

This is the most valued certification for this role.

Candidates with strong experience can still be selected without certifications.

Skills That Are Not Priorities

  • Terraform

Interview Process:

  • Round 1: Teams Interview
  • Round 2: Teams Interview
  • Additional Team Members

About us:

DivIHN , the 'IT Asset Performance Services' organization, provides Professional Consulting, Custom Projects, and Professional Resource Augmentation services to clients in the Mid-West and beyond. The strategic characteristics of the organization are Standardization, Specialization, and Collaboration.

DivIHN is an equal opportunity employer. DivIHN does not and shall not discriminate against any employee or qualified applicant on the basis of race, color, religion (creed), gender, gender expression, age, national origin (ancestry), disability, marital status, sexual orientation, or military status.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer - Remote in Remote vacancy
  • We're seeking a skilled and proactive Site Reliability Engineer to join our team, ensuring the stability, security, and efficiency of our technological...  ...-edge AI solutions to the government. This is a fully remote position for candidates in the continental U.S., with work... 
    Remote work

    Knexus

    Vienna, VA
    3 days ago
  •  ...and Seattle). Squad as a whole is responsible for the core platform services, split into 2 teams: 1 is Platform Engineering, and the other is Site Reliability Engineering. This is for the Site Reliability Engineering team. • Need in-depth knowledge of Linux, Windows... 
    Remote work
    Local area

    My3Tech Inc

    United States
    4 days ago
  • $150k - $200k

     ...will depend on. We're a small, remote-first team. We take ownership...  ...funding announcement: . The Reliability team owns the availability,...  ...reliability standards across engineering Designing incident response processes...  ...of production systems. As a Site Reliability Engineer on the... 
    Remote work
    Visa sponsorship
    Work visa
    Flexible hours

    GrabJobs

    San Bernardino, CA
    8 hours ago
  • $130k - $180k

     ...experienced and innovative leaders and engineers in the field. Where we work Headquartered...  ...team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure...  ...systems. Working conditions: Primarily remote Occasional travel to data centers required... 
    Remote work
    Temporary work
    Work at office
    Immediate start
    Flexible hours

    GrabJobs

    Boston, MA
    14 hours ago
  • $114k - $148k

     ...Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 Additional variable compensation and benefits may apply. Total compensation is based... 
    Remote work
    Full time
    Temporary work
    Work experience placement

    GrabJobs

    Newark, NJ
    4 days ago
  •  ...navigate the crypto world. At Newton, you'll work with a remote team spread across Canada, but you'll never feel distant. Ready...  ..., and solutions. Role Overview: We're looking for a Site Reliability Engineer to improve the reliability, resilience, and operational readiness... 
    Remote work
    Full time
    Shift work

    Newton

    Ontario, CA
    4 days ago
  • $104.9k - $174.7k

     ...link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability...  ...you may work a hybrid schedule. If not, this role is fully remote. We do not restrict applicants based on job site or posting... 
    Remote work
    Full time
    Work at office
    Local area
    Work from home

    RELX Group

    Buford, GA
    2 days ago
  • $35 - $45 per hour

    DescriptionKforce has a client that is seeking a remote Site Reliability Engineer to join their team.Summary:The team consists of systems that can track lead management, job management and sales management. It is built on Salesforce but underpinned by a lot of Java/API... 
    Remote work

    KForce

    Atlanta, GA
    3 days ago
  • $125k - $185k

    Washington, D.C.Engineering /Full-time /HybridA World-Changing CompanyPalantir builds the...  ...children, and more.The RoleWe’re looking for Site Reliability Engineers who can help us build,...  ...there are a few roles that allow for “Remote” work on an exceptional basis. If you are... 
    Remote work
    Full time
    Work experience placement
    Work at office
    Work from home
    Relocation package

    Palantir Technologies

    Washington DC
    14 hours ago
  • $230k - $250k

    GovCIO is hiring a Site Reliability Engineer with an active Secret clearance to ensure reliability, scalability, performance, and availability of...  .... This role is based in Arlington, VA, as a hybrid/remote position.ResponsibilitiesResponsibilities:Design and maintain... 
    Remote work

    Govcio

    Arlington, VA
    3 days ago
  •  ...education. Client is currently seeking a talented Software Engineer who is able to work into the Site Reliability Engineer role. This candidate is expected to work...  ...in production and pre-production environments. Remote to start due to covid, but will go back on site Duration... 
    Remote work

    Intelliswift

    Durham, NC
    3 days ago
  • $125k - $185k

     ...children, and more.The RoleWe’re looking for Forward Deployed Site Reliability Engineers who can help us build, operate, and maintain high-...  ...Based on business need, there are a few roles that allow for “Remote” work on an exceptional basis. If you are applying for one... 
    Remote work
    Full time
    Work experience placement
    Work at office
    Work from home
    Relocation package

    Palantir Technologies

    Washington DC
    3 days ago
  •  ...GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8 to 15 yearsSkillsKubernetes...  ...are seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal candidate... 
    Remote work

    SRI Tech

    Plano, TX
    3 days ago
  • $65 - $75 per hour

    DescriptionKforce has a client seeking a remote Senior Site Reliability Engineer to be a l be a leading member of the team working with a diverse range of technologies. You will enjoy working in a friendly environment and benefit from our investment in staff. The role also... 
    Remote work

    KForce

    Boca Raton, FL
    2 days ago
  • $210k - $230k

    GovCIO is currently hiring for a Senior Site Reliability Engineer (SRE) to design, implement, and maintain highly available, scalable, and resilient...  ...This position is located in Arlington, VA and is a hybrid remote/onsite position.ResponsibilitiesKey Responsibilities:... 
    Remote work
    Currently hiring

    Govcio

    Arlington, VA
    4 days ago
  • $62k - $141k

    Site Reliability EngineerThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more capabilities. Whether...  ...expected to have their cameras on during meetings.Remote: If this position is listed as remote, there may still... 
    Remote work
    Full time
    Contract work
    Part time
    Work at office
    Local area

    Booz Allen Hamilton

    Aurora, CO
    2 days ago
  •  ...role (three days in the office/two days remote).Job Summary:With a "document first" approach...  ...the availability, scalability, and reliability of systems and applications.What will be...  ...Terraform or CloudFormation.Mentor junior engineers and provide technical guidance.Stay up-... 
    Remote work
    Work at office

    Interactive Brokers

    Greenwich, CT
    2 days ago
  • $104.43k - $156.65k

     ..., Comcast prefers to have employees on-site collaborating unless the team has been...  ...than 100 miles from the office for the remote option.)Job SummaryCOMCAST Technology Solutions...  ..., Paramount+, and many others.Our Site Reliability Engineering (SRE) team is at the heart of our... 
    Remote work
    Permanent employment
    Full time
    Work at office
    Worldwide
    Flexible hours

    Comcast

    Centennial, CO
    14 hours ago
  • $96.8k - $145.2k

     ...thinking organization, apply now.We are currently seeking a Site Reliability Engineer (Onsite Hybrid) to join our team in Plano, Texas (US-TX),...  ...tailored to each client’s needs. While many positions offer remote or hybrid work options, these arrangements are subject to change... 
    Remote work
    Full time
    Temporary work
    Work at office
    Flexible hours

    NTT DATA

    Plano, TX
    2 days ago
  • $90k - $180k

     ...people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale,...  ...performance, and operational excellence of Merlin.net — a remote monitoring platform designed to help doctors, cardiologists... 
    Remote work

    Abbott

    Sunnyvale, CA
    14 hours ago
  • $80k - $133k

     ...Minimum Four (4) years of experience in IT administration, software engineering, or platform engineering, with a focus on AWS cloud...  ...obligated to pay a placement fee.SummaryLocation: US - TX, San Antonio; US - VA, McLean; US - Remote (Any location)Type: Full time
    Remote work
    Permanent employment
    Full time
    Contract work
    Flexible hours

    Guidehouse

    San Antonio, TX
    14 hours ago
  • $15k

     ...office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster to...  ...law.Compensation Range: $205K - $235KLocationBerkeley, CA; Remote, United StatesEmployment TypeFull timeLocation TypeRemoteDepartmentSoftwareCompensationBase... 
    Remote work
    Work at office
    Local area

    The Voleon Group

    Berkeley, CA
    1 day ago
  • $146k - $194k

     ...Anduril as a lead provider of specialized engineering and products for Intelligence Community...  ...requirements.ABOUT THE JOBAs a Site Reliability Engineer, your primary mission is to ensure...  ...supporting edge. Patch/update management and remote management tooling.DevOps improvements:... 
    Remote work
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Reston, VA
    1 day ago
  • Reliability Engineering Design, implement, and operate scalable, resilient, and highly available systems...  ...Three or more years of experience in Site Reliability Engineering, platform engineering...  ...that integrate cloud platforms with remote sites, field equipment, industrial... 
    Remote work

    Patterson-UTI

    Houston, TX
    1 day ago
  • $117k - $209.33k

     ...Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable,...  ...(not on this external site).SummaryLocation: Idaho, USA - Remote; AMER - United States - Texas - PlanoType: Full time
    Remote work
    Full time
    For contractors

    Autodesk

    Plano, TX
    4 days ago
  •  ...cloud-native platforms to advanced release engineering practices, our teams are redefining how...  ...: Hybrid - 2 days onsite, 3 days remote per weekSponsorship Notice: At this time...  ...office#LI-KC1#GMFjobsAbout The Role: The Site Reliability Engineer under the general direction... 
    Remote work
    Work experience placement
    H1b
    Work at office
    Visa sponsorship
    Flexible hours
    Shift work
    2 days per week

    GM Financial

    Arlington, TX
    2 days ago
  • $174.92k - $209.91k

     ...access to data as simple and reliable as electricity. With Fivetran...  ...and ready to query, with no engineering or maintenance required. We’re...  ...our teams, systems, and career sites.About the RoleFivetran is...  ...work model offers a blend of remote flexibility and in-person collaboration... 
    Remote work
    Full time
    Work at office

    Fivetran

    Oakland, CA
    14 hours ago
  •  ...cloud-native platforms to advanced release engineering practices, our teams are redefining how...  ...: Hybrid - 2 days onsite, 3 days remote per weekSponsorship Notice: At this time...  ...that accelerate development and improve reliability. Your work will directly influence how... 
    Remote work
    H1b
    Work at office
    Visa sponsorship
    Flexible hours
    2 days per week

    GM Financial

    Arlington, TX
    3 days ago
  • $150k - $180k

     ...operates through three business units: Remote Sensing (the data), Space Systems (the...  ...are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale...  ...organization.This position is based on-site in either our Arlington, VA office, Reston... 
    Remote work
    Permanent employment
    Full time
    Work at office
    Local area
    Worldwide

    Umbra

    Arlington, VA
    3 days ago
  • $102.1k - $202.2k

     ...yearEmployment type: Full-TimeWork site: 3 days / week in-officeRole...  ...EngineeringDiscipline: Site Reliability EngineeringCompany:...  ...the cloud. Our work enables remote computing experiences that are...  ...team, you will collaborate with engineers across disciplines to deliver... 
    Remote work
    Ongoing contract
    Work experience placement
    Local area
    3 days per week

    Microsoft

    Redmond, WA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer - Remote. Be the first to apply!