Remote Site Reliability Engineer in Network Infrastructure
Nebius
About Nebius:
Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.
Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.
Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.
The Role
We’re looking for a Site Reliability Engineer to help build and run the fundamental part of Nebius – the Network – the infrastructure everything else depends on. This is an engineering-first SRE role: you’ll set clear reliability targets, build the tooling and automation to meet them, and make the network safer to operate as we scale quickly.
Your responsibilities will include:
- Define and own reliability goals for network services and critical paths (SLIs/SLOs, availability targets, error budgets where it makes sense)
- Drive reliability improvements across the whole network: not only services, but also site readiness, inter-site connectivity (DCI), and operational standards
- Own incident response for your areas, lead investigations/postmortems, and turn failures into durable fixes (not repeated firefighting)
- Build and evolve observability: actionable metrics/logs/traces, alerting, and faster debug loops during and after incidents
- Design safer change workflows: automation, CI/CD, test/staging environments, canarying, rollbacks, and auditability for network changes
- Work closely with network engineers and platform teams to embed operability into designs and keep operations practical and fast
We expect you to have:
- Strong production Linux fundamentals and a structured approach to debugging complex systems
- Solid understanding of networking basics and how real networks fail (control plane vs data plane, latency/loss, failure domains, etc.)
- Hands-on experience operating high-availability systems and improving them over time (not just “keeping lights on”)
- Ability to write and maintain software/automation (Go is common for us; Python is also welcome)
- Experience with modern infrastructure tooling (e.g., IaC, CI/CD, container platforms) and comfort automating operational workflows
It will be an added bonus if you have:
- Experience with high-throughput traffic processing: load balancers, tunneling/decap, NAT64, or similar datapath-heavy systems
- Low-level networking performance/debug background (eBPF/XDP, DPDK, perf/ftrace, kernel networking internals)
- Experience building network-safe delivery pipelines (testing labs, staged rollouts, automated verification, drift detection)
- Background with large-scale network observability/telemetry (e.g., routing/flow telemetry, regression detection at scale)
Benefits & Perks:
- Competitive compensation
- Career growth and learning opportunities
- Flexibility and ownership
- Collaborative and innovative culture
- Opportunity to work on impactful AI projects
- International environment and talented teams
What’s it like to work at Nebius:
Fast moving – Bold thinking – Constant growth – Meaningful impact – Trust and real ownership – Opportunity to shape the future of AI
Equal Opportunity Statement:
Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law.
Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire.
If you need accommodations during the application process, please let us know.
Jobicy JobID: 150130$159k - $272k
...Role SummaryIn this role as Principal Site Reliability Engineer, Infrastructure Observability you will help... ...operateMaintains a broad internal professional network and knows when to engage/activate... ...Maryland, Colorado, Washington and remote workers$175,000.00 - $299,000.00...Remote workFull timePrivate practiceLocal areaWork from home3 days per week- ...experienced Senior or Staff Engineer for our SRE, InfraSec team,... ...security of our cloud-based infrastructure. As a Staff SRE, you will be... ...hybrid basis, or it can be fully remote while working from a... ...AWS, Azure, GCP), including network and compute security, identity...Remote workFull timeWorldwide
- ...of enabling human life on Mars.SITE RELIABILITY ENGINEER (MANUFACTURING INFRASTRUCTURE)The application software team is... ...hardware. The compute, storage, and networking that run our factories must be... ...role requires you to be onsite. Remote and/or hybrid work will not be...Remote workPermanent employmentWeekend work
$135k - $200k
...The Role As a Senior Software Engineer on Network Infrastructure you will be joining a team whose mission... ...(architecture, design patterns, reliability and scaling) of new and existing systems... ...are a few roles that allow for “Remote” work on an exceptional basis. If you...Remote workFull timeWork experience placementWork at officeWork from homeRelocation package- ...Job Title: Software Engineer - Senior Level (IE Platform Infrastructure) Location: Arlington, VA Clearance: Top... .... This role focuses on designing reliable integration environments, managing... ...Development Training Reimbursement ~ Flexible/remote work schedule....Remote workFull timeFlexible hours
$165k - $225k
...high-performance AI infrastructure for organizations running... ...enterprise-grade reliability and compliance. Your... ...storage, and networking fabric – with the systems... ...enable researchers and engineering teams to programmatically... ...and success as we grow together. #li-remote...Remote workFull timeImmediate startFlexible hours$280k - $380k
...scale. We focus on reliability and automation, engineering systems that perform... ...turning complex infrastructure into reliable, well... ...experienced DevOps/SRE (Site Reliability... ...Solid understanding of networking, security, and compliance... ...are flexible for remote work except for...Remote workFull timeWork at officeLocal areaMonday to ThursdayFlexible hours- ...looking for a Software Engineer - Platform / Core Infrastructure who’s passionate about building... ...systems that scale reliably and securely, and can be... ...across compute, storage, networking, and observability to drive... ...Know # Location: Remote in the US or EMEA (all team...Remote workFull timeContract workImmediate start
$168k - $193.73k
...hiring a Senior Software Engineer to help modernize, scale,... ...and operate the shared cloud infrastructure that powers YipitData’s data... ..., you will drive forward Site Reliability Engineering and DevOps initiatives... ...may be performed fully remotely within the United States....Remote jobFull timeWork at officeVisa sponsorshipFlexible hours$180.5k - $236.91k
...hiring a Senior Software Engineer, Cloud Infrastructure / SRE to join our Engineering... ...Location: This is a remote position, open to candidates... ...domains such as DevOps, site reliability, and cloud best practices... ...similar. Security & Networking: Knowledge of cloud-native...Remote jobFull timeWork at office- ...Role Summary: As an Infrastructure Engineer, you will have the opportunity... ...improvement of our infrastructure reliability and security while... ...Demonstrable experience with networks, security, load balancers,... ...least three days a week. For remote employees, occasional travel...Remote workFull timeWork experience placementWork at office3 days per week
$213k - $263k
...looking for experienced data-minded software engineers and data scientists to help us improve... ...Design and implement robust tools and infrastructure for data mining, exploration, and... ...location or, if the role can be performed remote, the specific salary range for your...Remote workFull time- ...stage is exactly why this role exists. The Staff AI Engineer owns ShiftKey’s AI infrastructure layer: the organizational knowledge platform, the infrastructure... ...leadership. Where you’ll work This is a fully remote position based in the United States. You should be...Remote workFull timeWork experience placementWork at officeFlexible hoursShift work
- ...Building and operating robust network infrastructure, the full-time Senior Network Infrastructure Engineer will design and manage public and private networks for a remote organization, focusing on high-performance Linux networking, BGP routing, and automation using AI...Remote workFull time
$128k - $145k
...environments. We’relooking for aCloud Infrastructure Engineerto helpoperateand evolve... ...,you’llwork closely with senior engineers to support the reliability, security, and performance of a production... ...with core AWS services (compute, networking, storage) Basic experience with...Remote workWork from home$100k - $150k
...Platform Infrastructure Engineer - Remote Bright Vision Technologies is a technology consulting and software development company delivering... ...OpenShift internals, including operators, CRDs, RBAC, and networking. Strong Linux administration skills, including...Remote workFull timeH1bLocal areaImmediate startVisa sponsorship$140k - $160k
...Connect Centric Network Engineer At Connect Centric, we... ...enhance the network infrastructure across physical data... ...Your work will ensure reliable, secure, and high-performing... .... Support secure remote access and VPN tunnel... ...Minimal. On-site presence required several...Remote workFull timeRelocation package$90.58k
...Description of Duties The IT Network Engineer is responsible for the... ...university’s enterprise-wide network infrastructure. This includes the... ...Support VPN and secure remote access solutions. Ensure... ...able to provide support on site as needed to 6 respective UDC...Remote workFull timeH1bWork at officeLocal areaImmediate start$65.7k - $118.3k
...Network Infrastructure Engineer Would you enjoy improving stability and safety of one of the largest... ...Partnering across teams to ensure the reliability, scalability and usability of our... ...Broadway, Cambridge, MA, 02142, US (Remote) Citizenship required Yes Is Security...Remote workWork experience placementWork at office- ...Passionate about cloud infrastructure operations, the full-... ...Cloud Infrastructure Engineer will guide NVIDIA... ...observability, while working remotely or onsite in Santa... ..., CPU, storage, and network health across AI... ...infrastructure engineering, Site Reliability Engineering, or...Remote workFull time
- ...Owning the real-time cloud infrastructure for connecting software... ...hardware, the full-time remote Senior/Staff Software Engineer will manage data paths, AWS... ...while ensuring system reliability and observability. Key Responsibilities... ...and Terraform Hands-on site reliability experience...Remote workFull time
- ...Cloud Infrastructure Engineer Location: Texas Remote Contract: Longterm Note: Looking for W2 candidates only. Position Requirements: Experience... ...Good knowledge and experience with securing cloud networking, disaster recovery and integration with Azure, AWS,...Remote workContract workWork experience placement
$145k - $195k
...CAD environment with AI-powered engineering workflows, helping teams design,... ...Angeles, CA with both a local and remote team. We were founded and... ...Senior Software Engineer, Cloud Infrastructure who can take ownership of the reliability, automation, and evolution of our...Remote workLocal area- ...protective products, as well as engineered coated materials and... ...Description Title: Cloud Infrastructure Engineer Department: MIS... ...Exempt Salaried Location: Remote Position Purpose: Administer... ...Virtualization, Server, Storage, Network, Endpoints and Directory...Remote workImmediate startWeekend work
- ...Cloud Infrastructure Engineer – AWS Bright Vision Technologies is seeking a highly experienced... ..., DevOps, cloud security, and Site Reliability Engineering (SRE) while providing technical... ...practices. Design resilient AWS networking architectures, including VPCs,...Remote work
- ...are seeking a DevOps Engineer with strong hands-on experience... ..., cloud and hybrid infrastructure, CI/CD, GitOps,... ...support customers through reliable, repeatable, and... ...maintain Kubernetes RBAC, network policies, ingress controllers... ...client meetings, on-site engagements, internal...Remote workCasual workFlexible hoursAfternoon shift
$190k - $260k
...team of researchers, engineers, designers, and... ...performance, scalable and reliable machine learning... ...are looking for a Site Reliability... ...and influence the Infrastructure team’s roadmap based... ...compute/storage/network resource and cost... ...offices if you are remote, plus an annual company...Remote workFull timeWork experience placementWork at officeLocal areaHome office- ...Do Own and evolve how compute infrastructure is provisioned and operates across... ...behavior across CPU, memory, storage, and network, including how workloads are scheduled... ...infrastructure. Collaborate with control plane engineering to establish a clear and stable...Remote workPermanent employmentWork at officeFlexible hours
- ...meaningful impact. As a Senior Lead Infrastructure Engineer at JPMorganChase within the... ..., deploys, and configures network solutions/fabrics for the... ...of on-premises and remote Data Center technologies and... ...comprehensive health care coverage, on-site health and wellness centers,...Remote workShift work
- Senior Site Reliability Engineer - Network Operations (Remote) Fastly 15 August 2025 SRE DevOps Automation Networking BGP Fastly is seeking a Senior Site Reliability... ...be responsible for building and operating the infrastructure that powers the Fastly Edge Cloud Platform, a...Remote jobLocal areaFlexible hoursNight shift
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Remote Site Reliability Engineer in Network Infrastructure. Be the first to apply!
- site reliability engineer sre Remote
- site reliability engineer Remote
- site reliability engineer remote Remote
- lead network engineer Remote
- junior network engineer no experience Remote
- cisco network engineer Remote
- junior network engineer Remote
- production network engineer Remote
- network engineer Remote
- network engineer full time Remote



