Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Network Reliability Engineer

$135k
Full-time

Group 1001

Group 1001 is a consumer-centric, technology-driven family of insurance companies on a mission to deliver outstanding value and operational performance by combining financial strength and stability with deep insurance expertise and a can-do culture. Group1001’s culture emphasizes the importance of collaboration, communication, core business focus, risk management, and striving for outcomes. This goal extends to how we hire and onboard our most valuable assets – our employees.

Why This Role Matters:

The Platform Engineering Services team at Group 1001 is building a Site Reliability Engineering practice with a network scope. We're hiring an Sr. Network Reliability Engineer who embodies Innovation and Excellence, and will apply SRE principles — code-as-source-of-truth, SLOs and error budgets, alerting on symptoms rather than causes, failure-mode-first design, and the elimination of toil — to the firm's network platform from carrier edge through cloud fabric to Kubernetes pod boundary. This is not a "keep the lights on" role. You will systematically engineer the lights-on work out of existence, build the abstractions that let other engineering teams express network intent in code, and treat the network as a single engineered system rather than a collection of vendor consoles. You will operate inside a DevSecOps practice spanning multi-cloud, multi-region environments, and you will partner closely with Cloud and Data Platforms, the NOC/SOC, and Cyber Security to extend reliability practice across the firm.

How You’ll Contribute:

  • Treat reliability as an engineered property. Define SLOs and error budgets for the network platform — DNS resolution, edge availability, mesh ingress success, cross-region path health — and use them to gate changes, not just to color dashboards. Lead postmortems with a focus on permanent remediation, not pattern-recognition. Alert on symptoms users feel, not on causes that may or may not produce impact.

  • Move network state into code. Use Terraform (or Pulumi), Ansible, and Python to replace CLI-driven configuration with declarative, version-controlled, peer-reviewed change running through Infra CI/CD. This applies equally to the edge tier (Cloudflare), security platforms (Zscaler ZIA/ZPA, ZTNA policies, next-gen firewalls), the cloud network fabric (Transit Gateway, Cloud WAN, VPCs, Route53, IPAM), and increasingly the Kubernetes and service-mesh layer.

  • Build network policy as intent, not rule lists. Express what flows are permitted, what segments are isolated, what egress is inspected, what zones share DNS — and engineer the compilers that turn that intent into per-vendor configuration. Use Policy as Code (OPA/Rego, Sentinel, Cilium NetworkPolicy) to catch invariant violations at plan time, not apply time.

  • Infrastructure as Code (IaC): Design, deploy, and manage network infrastructure using Terraform or Ansible, moving the firm away from manual configuration to a code-first approach.

  • Engineer the cloud network platform. Operate and extend our multi-account AWS Landing Zone — Cloud WAN segmentation, Transit Gateway peering, IPAM-driven CIDR allocation, shared private DNS, cross-account telemetry pipelines. Build the platform abstractions that make a new account or service land correctly with policy and connectivity composed from declarative inputs.

  • Extend platform thinking into the container tier. Kubernetes networking, service mesh (Istio, Linkerd, Consul Connect), eBPF-based observability and policy (Cilium, Hubble), and the integration points where mesh-level authz meets cloud-tier identity. Recognize that an "internal" service is one logical hop on a chain of policy enforcement points and engineer for that explicitly.

  • Improve telemetry and observability with intent. Build alerts as structured payloads with runbook links, suspected blast radius, and dependency-aware suppression. Author both system-health dashboards for operators and end-user monitoring dashboards that reflect actual user experience. Use Grafana, Elastic, Open Telemetry where each fits.

  • Mentor and grow the team. Provide technical guidance to junior engineers, foster a culture of learning, and work out loud across Platform Engineering so the patterns you build cross-pollinate to adjacent domains.

  • Handle hardware when required. Provide maintenance and configuration support for routers, switches, and firewalls at data centers and offices when needed — bringing code-first practices to physical hardware where possible (templating, change validation, zero-touch provisioning) and direct hands-on competence where it isn't.

  • Incident Response: Serve as an escalation point for network issues, some complex and some basic but not yet covered by runbooks. Troubleshooting with a focus on root cause analysis and permanent remediation with a documentation-first mindset.

  • Reduce toil and hand off cleanly. Repetitive operational tasks are scoped engineering problems with measurable payoff. Author runbooks and SOPs that the NOC can execute confidently; package routine work for L1/L2 handoff so engineering interrupt drops over time. Coordinate across Data Platforms, NOC/SOC, and Cyber Security so reliability practices spread instead of staying siloed.






What We’re Looking For:

  • Network Engineering: Deep understanding of TCP/IP, BGP, OSPF, VPNs, and SD-WAN architecture.

  • Automation: Proven experience with Terraform (state management, modules) and Ansible (playbooks, roles) – or similar – in a production environment. Proficiency in Python for automation and API interaction, or similar.

  • Security Platforms: Hands-on experience with Cloudflare, zScaler, and/or enterprise firewalls.

  • Observability: Experience configuring monitoring tools (e.g., Datadog, Prometheus, Grafana) to create meaningful alerts and dashboards.


Nice to Have

  • Service mesh experience (Istio, Linkerd, Consul Connect, Cilium).

  • eBPF-based observability (Hubble, Pixie).

  • AWS Multi-account landing zone tooling experience (AFT, Control Tower, or equivalent).

  • Policy as Code experience (OPA/Rego, Sentinel, Cilium NetworkPolicy).

  • Professional Attributes

  • Documentation First: A strong belief that a job isn't done until the documentation in written.

  • Toil Reduction: A mindset that actively seeks to automate repetitive tasks.

  • Hybrid Capability: Willingness to handle physical hardware tasks when required while maintaining a software-centric engineering mindset.

Compensation:




Our compensation reflects the cost of labor across several U.S. geographic markets. The base pay for this position ranges from $135,000/year in our lowest geographic market up to $190,000/year in our highest geographic market. Pay is based on factors such as market location, job-related skills, and experience.

Benefits Highlights:

Employees who meet benefit eligibility guidelines and work 30 hours or more weekly, have the ability to enroll in Group 1001’s benefits package. Employees (and their families) are eligible to participate in the Company’s comprehensive health, dental, and vision insurance plan options. Employees are also eligible for Basic and Supplemental Life Insurance, Short and Long-Term Disability. All employees (regardless of hours worked) have immediate access to the Company’s Employee Assistance Program and wellness programs—no enrollment is required. Employees may also participate in the Company’s 401K plan, with matching contributions by the Company.

Group 1001 , and its affiliated companies, is strongly committed to providing a supportive work environment where employee differences are valued. Diversity is an essential ingredient in making Group 1001 a welcoming place to work and is fundamental in building a high-performance team. Diversity embodies all the differences that make us unique individuals. All employees share the responsibility for maintaining a workplace culture of dignity, respect, understanding and appreciation of individual and group differences.

#LI-AS1 #LI-REMOTE
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Senior Network Reliability Engineer in Remote vacancy
  • $136k - $224.25k

    NVIDIA is looking for a Senior Network Reliability Engineer to support and maintain our cloud and datacenter network infrastructures. This network serves the needs across the whole software stack for NVIDIA, from Graphics Drivers to Autonomous Vehicles and Artificial Intelligence... 
    Senior
    Full time
    Remote work
    Shift work

    Nvidia

    Santa Clara, CA
    2 days ago
  •  ...across our physical data centers. We are looking for a Senior Site Reliability Engineer to improve the reliability, scalability, and operational...  ...postmortems, and durable corrective actions.Partner with Compute, Networking, Storage, Security, and Support teams.Participate in on-... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    4 days ago
  •  ...demonstrates proficiency across software engineering practices with advanced specialization...  ...contributing to Azure cloud platform reliability. This position is primarily responsible...  ...Excellence (CoE).Background in Virtual Network support, including VNet-injected environments... 
    Senior

    Walgreens Boots Alliance

    Deerfield, IL
    4 days ago
  •  ...Job Title: SRE - Platform - Kubernetes (Senior SRE Engineer) Job Location:Remote, prefer PST...  ...powers our cloud. We focus on delivering reliable, scalable, and simple platforms that...  ...working knowledge of Linux systems, networking, and distributed systems fundamentals... 
    Senior
    Remote work

    Apptad Inc

    Georgia
    2 days ago
  •  ...Job Title: Senior Site Reliability Engineer (SRE) AWS GovCloud Location-Remote Job Description: We are looking for a Senior SRE with...  ...infrastructure-as-code patterns. Support cloud networking, IAM, account provisioning, and compliance automation.... 
    Senior
    Full time
    Remote work

    Siri InfoSolutions Inc

    New York, NY
    2 days ago
  •  ...complex challenges across multiple high-impact projects, driving efficiency and innovation for global initiatives. Join EPAM to engineer solutions that matter. From AI to cloud transformation, you'll collaborate with top-tier innovators, gain autonomy to explore your... 
    Senior
    H1b
    Remote work
    Shift work

    EPAM Systems

    United States
    2 days ago
  •  ...Job Title : Senior Site Reliability Engineer Kubernetes Platform Location: Remote Job Type : Fulltime Job Summary Must...  ...of cloud infrastructure, Linux systems, and networking fundamentals Experience with Infrastructure as Code... 
    Senior
    Full time
    Remote work

    Siri InfoSolutions Inc

    Remote
    2 days ago
  • $80 - $90 per hour

     ...LaSalle Network is hiring for a Senior Site Reliability Engineer (Compute Platform) with a leading infrastructure and platform engineering firm known for innovation and cutting-edge technological solutions. This opportunity is a remote role focused on deep infrastructure... 
    Senior
    Hourly pay
    Contract work
    Temporary work
    Remote work

    LaSalle Network

    Chicago, IL
    8 days ago
  •  ...Job : Senior SRE - Platform - Private Cloud Engineer Location : Remote Fulltime Only Job Description Must Have Technical/Functional...  ...Monitor, troubleshoot, and optimize platform performance, reliability, and security. Develop and maintain... 
    Senior
    Full time
    Remote work

    Siri InfoSolutions Inc

    Remote
    2 days ago
  • $80 - $90 per hour

     ...LaSalle Network is hiring for a Senior Site Reliability Engineer (Storage Platforms) with a storage-focused, enterprise-leading company known for innovation and impactful cloud infrastructure. Join a dedicated team managing software-defined storage solutions and enterprise... 
    Senior
    Hourly pay
    Contract work
    Temporary work
    Remote work

    LaSalle Network

    Chicago, IL
    8 days ago
  • $210k - $220k

     ...trusted by security, compliance, and engineering teams to automate complex architectures...  ...-fluent, and infrastructure-minded Senior Site Reliability Engineer - Government Cloud to join our...  ...experience who handles distributed network topologies fluidly natively using DevOps... 
    Senior
    Permanent employment
    Full time
    Remote work
    Shift work

    Tines

    Remote
    a month ago
  • $149.4k - $202k

     ...Noctua Technology is seeking a Senior Software Engineer specializing in Site Reliability Engineering to join their team. This role focuses on the reliability and performance of cloud-native applications, emphasizing Infrastructure as Code and automation. The ideal candidate... 
    Senior
    Remote work

    Noctua Technology

    Virginia, MN
    1 day ago
  •  ...Description We are looking for an SRE / Cloud Engineer with strong hands-on experience...  ...role is ideal for someone who understands reliability engineering, cloud operations, incident...  ...infrastructure, applications, databases, queues, and network dependencies. ~Tune alerts to reduce... 
    Senior
    Full time
    Internship
    Remote work
    Relocation

    Miratech

    Remote
    2 days ago
  • $175k - $195k

    Role Description As a Site Reliability Engineer, you will be embedded with a cross-functional team responsible for key portions of our systems. Your role will involve: ~Owning platform reliability, availability, and performance. ~Building AWS infrastructure using... 
    Senior
    Full time
    Temporary work
    Remote work

    Filevine

    Remote
    10 hours ago
  • $82.3k - $228.8k

     ...authentic selves. We are seeking a highly experienced Senior Site Reliability Engineer – Compute Platforms to design, implement, and support...  ...troubleshooting across storage, Kubernetes, hypervisors, networking, and Linux systems Partner with operations, network,... 
    Senior
    Temporary work
    Work at office
    Remote work
    Worldwide
    3 days per week

    Five9

    San Ramon, CA
    more than 2 months ago
  •  ...efficiency and effectiveness of the team.Promoting the WSP brand, through active involvement with industry partners and professional engineering institutions. This may include sitting on councils, committees and working groups which promote and drive technical excellence in... 
    Senior
    Part time
    For subcontractor
    Local area
    Work from home
    Flexible hours
    2 days per week

    WSP Group

    Colorado
    4 days ago
  • About the RoleAs Senior Reliability Engineer, you will ensure the resilience and availability of Kohl’s systems and applications, collaborate closely...  ...of systems architecture, operating system internals and network fundamentals In-depth knowledge of application design... 
    Senior
    Work experience placement
    Remote work

    Kohl's

    Menomonee Falls, WI
    1 day ago
  • $132.4k - $251.6k

     ...than 100 years of experience and renowned engineering expertise to meet the needs of today’s...  ...responsible for ensuring our products are Safe, Reliable, Maintainable and delivered on time....  ..., System Safety and Supportability.The Senior Principal Reliability Engineer will... 
    Senior
    Temporary work
    Work experience placement
    Interim role
    Remote work
    Flexible hours

    Raytheon

    Tucson, AZ
    3 days ago
  • $110k - $150k

    DescriptionKforce has a client in Phoenix, AZ that is seeking a few Senior Reliability Engineers. We are partnering directly with the hiring manager on these key hires. This position is Hybrid Remote.Key Responsibilities:* Drive equipment reliability and continuous improvement... 
    Senior
    Work at office
    Remote work

    KForce

    Phoenix, AZ
    2 days ago
  • Are you ready to strengthen the reliability backbone of a high-impact manufacturing operation?Leola, PA, will soon be home to Dart’s new...  ...for paper-based food and beverage packaging.As a Reliability Engineer, you will play a critical role in ensuring our manufacturing equipment... 
    Senior
    Full time
    Temporary work

    Dart Container Corporation

    Leola, PA
    10 hours ago
  • $86.8k - $165.2k

     ...strength of more than 100 years of experience and renowned engineering expertise to meet the needs of today’s mission and stay ahead...  ...(LCE) Long Range Radars department is seeking a Senior Reliability Engineer, focusing on the Reliability Engineer discipline.... 
    Senior
    Temporary work
    Work experience placement
    Work at office
    Remote work
    Relocation package
    Flexible hours

    Raytheon

    Woburn, MA
    2 days ago
  •  ...next step to an altogether life-changing career.Learn about the Danaher Business System which makes everything possible.The Senior Reliability Engineer is responsible for improving equipment reliability, reducing unplanned downtime, and supporting maintenance and... 
    Senior
    Hourly pay
    Full time
    Local area
    Remote work
    Work from home
    Flexible hours

    Danaher Corporation

    Pensacola, FL
    4 days ago
  • $139.4k - $205k

     ...technologies that support and augment human networks, aiming to improve efficiency for...  ...on business impact. We are a highly senior team composed of former pioneers from...  ...seeking a highly motivated Senior Reliability & Test Engineer to join our team. This individual will... 
    Senior
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    4 days ago
  • Summary: VIAVI (NASDAQ: VIAV) is a global provider of network test, monitoring and assurance solutions for telecommunications...  ...customers.Duties & Responsibilities: Job Summary:The Senior Test / Reliability Engineer will be responsible for developing test plans and... 
    Senior
    Full time
    Contract work
    Work from home

    Viavi

    Germantown, MD
    3 days ago
  • $116.35k - $210.33k

    Leidos has a new and exciting opportunity for a Senior Network Engineer in our Intelligence Sector, Cyber & Analytics Business Area (CABA). Our...  ..., program management, and engineering teams to ensure reliable network performance in DoD/IC environments.Primary ResponsibilitiesOptimize... 
    Senior
    Full time
    Immediate start
    Remote work
    Flexible hours

    Leidos

    Columbia, MD
    2 days ago
  •  ...high performance cloud networkWork on deploying and configuring networking Hardware for new and existing clustersEnsure high...  ...be part of day2 operations and on-call rotation for Network Engineering teamYouHave 10+ years of experience in IT and networking spaceHave... 
    Senior
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    2 days ago
  • $131.3k - $237.35k

    Leidos is looking for a Senior Network Engineer to be part of a team responsible for the design, migration, implementation, enhancement, optimization and maintenance of LAN/DATACENTER/WAN.This position will work with engineers on various technical projects. This could include... 
    Senior
    Full time
    Remote work
    Relocation

    Leidos

    Chantilly, Loudoun County, VA
    2 days ago
  • $92.3k - $166.85k

    Leidos is seeking an experienced and motivated Senior Network Engineer to support the Defense Threat Reduction Agency (DTRA) on the Integrated...  ...firewalls, load balancers, and VPN solutions to ensure secure and reliable connectivity.Manage and support Unified Communications,... 
    Senior
    Full time
    Contract work
    Remote work

    Leidos

    Fort Belvoir, VA
    2 days ago
  • $104k - $166k

    ResponsibilitiesPosition Overview We are seeking a highly skilled and motivated Senior Network Engineer to support a federal government customer in designing,...  ..., systems engineering, and facilities teams to ensure reliable and secure network operations. QualificationsRequired... 
    Senior
    Contract work
    Remote work
    Shift work

    Peraton Corporation

    Washington DC
    1 day ago
  • Position SummaryOdyssey Systems is seeking a Senior Network Engineer to support advanced situational awareness and multi-domain operations (MDO) programs. This role focuses on the design, analysis, and integration of RF and signal processing systems that enable next-generation... 
    Senior
    Full time
    Contract work
    Live out
    Remote work

    Odyssey Systems Consulting Group

    Tampa, FL
    10 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Network Reliability Engineer. Be the first to apply!