Site Reliability Engineer
$130k - $180kNebius
About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to work in our office in Amsterdam. Hardware Infrastructure team designs, develops and supports systems involved in the data-centers lifecycle: - Serving functional and load testing system. - Monitoring of engineering equipment located in our data centers (power supply, air and water cooling, etc. ) - Monitoring of IT equipment: racks, servers, JBODs, JBOGs, power shelves, network devices, etc. - Asset tracking. - Hardware repairs tasks tracking. - Server production.
In this position, your responsibility will be to : - Ensure fault-tolerance, scale and uninterrupted operations for our services. - Use cutting-edge technology to solve a variety of infrastructure problems. - Implement and improve CI/CD processes. We expect you to have : - Proficiency in Linux systems, with expertise in Python and Bash scripting for automation. - Demonstrated ability to troubleshoot complex system issues, including hardware, software and networking problems. - Strong analytical and problem-solving skills, with a focus on optimizing system performance. - Working proficiency in English. It would be an added bonus if you had : - Desire to be involved in backend development. - Experience designing, developing and running high-load distributed systems. Working conditions: - Primarily remote - Occasional travel to data centers required, especially if not located near one - Collaboration with globally distributed engineering and operations teams Key employee benefits: - Health insurance: 100% company-paid medical, dental, and vision coverage for employees and families - 401(k) plan: up to 4% company match with immediate vesting - Parental leave: 20 weeks paid for primary caregivers, 12 weeks for secondary caregivers - Remote work reimbursement: up to $85/month for mobile and internet - Disability & life insurance: company-paid short-term, long-term, and life insurance coverage Compensation - We offer competitive salaries, ranging from $130k- $180k base + quarterly performance bonuses. Join Nebius and help operate the systems that power next-generation AI infrastructure. Benefits & Perks: - Competitive compensation - Career growth and learning opportunities - Flexibility and ownership - Collaborative and innovative culture - Opportunity to work on impactful AI projects - International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law.
- ...to meet you. Our Enterprise Information Technology (EIT) organization is expanding, and we are seeking a Senior Site Reliability Engineer to help drive a major architectural modernization. In this role, you will move beyond traditional infrastructure maintenance...SuggestedPermanent employmentFull timeH1bLocal areaRemote workShift work
- ...and performance our customers have come to expect, and help raise the reliability bar as we grow. What you would do: Design, build, and operate the shared platform foundations engineers ship on every day: GCP infrastructure, Kubernetes, networking, routing,...SuggestedRemote workWorldwideFlexible hours
$147k - $168k
...Inc. as one of the most innovative and fastest-growing technology companies in the country. Role Summary As a Site Reliability Engineer at Filevine, you will improve the reliability, scalability, and operational maturity of the Filevine platform. You’ll...SuggestedFull timeTemporary workWork experience placementWork at officeRemote work2 days per week3 days per week- ...Site Reliability Engineer Company: Milestone Systems Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Golang, Python, Linux, Shell scripting, Kubernetes, Docker, Terraform, CI/CD, GitOps, ArgoCD, Spinnaker, Prometheus, Datadog,...SuggestedFull timeRemote work
- ...grow, adapt and use your skills consistently. Our customers rely on us in the moments that matter. Engineering delivers on that promise. The Senior Site Reliability Engineer is responsible for ensuring our SaaS products are fast, stable and optimized for our customers...SuggestedWork experience placementRemote workFlexible hours
$141.8k - $195k
...work, grow fast, and bring their full selves to the herd. Why You’ll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in the U.S. Cribl provides...Temporary workRemote work- ...encourage you to apply. The Role As a Senior Platform Engineer, you are a champion for DevOps and SRE culture and industry... ...met. \n What You Will Be Doing Improving production reliability and system resilience within an SRE scoped team Championing...Remote workFlexible hours
- ...Site Reliability Engineer Company: GitLab Work Type: Remote Employment: Full Time Location: CA, US Seniority: Senior Level Technologies: Terraform, Ansible, Kubernetes, Go, Ruby, Jsonnet, Prometheus, ELK, Grafana Requirements: Senior-level SRE with strong Terraform/IaC...Full timeRemote work
- ...Senior Site Reliability Engineer Company: CyberArk Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: AWS, Kubernetes, Terraform, CloudFormation, Ansible, CloudWatch, Grafana, Datadog, OpenSearch, PagerDuty Requirements: Senior SRE...Full timeRemote work
- ...Site Reliability Engineer OXIO is the first NeoTelco. We arebuilding the world’s largest, most accessible, and insightful Telecom network. Our platform empowers anyone to spin up their own carrier from a browser, scaling and supporting you as you scale your network...Remote work
- ...Job Title: Site Reliability Engineer (Azure Government & Infrastructure) Pay Type : SALARIED EXEMPT Location: Remote Citizenship Requirement: U.S. Citizen (Required) Summary of Position Role/Responsibilities The Site Reliability Engineer (SRE) for...Full timeRemote workMonday to Friday
$7.5k
...investment management. We have become a multibillion-dollar asset manager, and we have ambitious goals for the future. As a Site Reliability Engineer (SRE), you will work at the intersection of production operations and software development as you improve, manage, and...Local areaRemote work- ...healthcare organizations maintain accurate, compliant, and reliable provider networks at scale. Our vision is simple: One... ...patients. About the Role We're looking for a Senior Site Reliability Engineer who takes ownership seriously - someone who designs for...Remote work
$152k - $195k
...including Silver Lake Waterman, Moody’s, Sequoia Capital, GV and Riverwood Capital. About the Team: As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and optimization of our Kubernetes-based infrastructure and CI/CD systems...Remote work- ...About The Role: We're looking for a Senior Site Reliability Engineer to help us mature and scale the infrastructure behind our multi-cloud SaaS platform. Most of our footprint runs on Microsoft Azure, built from the ground up around cloud architecture principles:...Remote workFlexible hours
$145k - $193k
...entertainment, we want to talk to you. About the Role & Team The SRE team at PENN Entertainment is looking for a Senior Site Reliability Engineer to help build and operate the infrastructure behind a large-scale sports betting and media platform. You'll own critical...Remote work$160k - $180k
...big impact. See Arkestro in action at arkestro.com. About the Role Arkestro is hiring for a Senior SRE Engineer to manage our performance and reliability for our software platform and infrastructure. The right candidate will own and develop our infrastructural...Local areaRemote work- ...Site Reliability Engineer Company: Crunchafi Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Azure, AKS, Azure Kubernetes Service, Terraform, Bicep, ARM templates, GitHub Actions, Azure DevOps, Kubernetes, Docker, App Insights...Full timeRemote work
- ...Site Reliability Engineer Company: Quzara Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: Azure, Terraform, Bicep, Ansible, Azure Monitor, Azure Automation, Azure Policy, Azure Site Recovery, TLS/SSL Requirements: 4+ years in SRE...Full timeRemote work
- ...The Site Reliability Engineer (SRE) / Subject Matter Expert (SME) – Computer Systems Engineer/Architect will provide senior-level reach-back expertise to support the reliability, scalability, performance, and operational resilience of the GEOMAP platform in secure cloud...Full timeContract workFor contractorsFor subcontractorRemote work
- ...organization across multiple locations in the US, South America, and India. Location: Remote (US-Based Candidates Only) Site Reliability Engineer II (SRE) Position Overview We are seeking a Site Reliability Engineer II (SRE) to join our growing Site...Remote workFlexible hoursShift workWeekday work
$160k - $208k
...early diagnosis and longitudinal care management of chronic conditions. We are looking for an experienced Site Reliability and Infrastructure Engineer to join our engineering team. You will support Counterpart Health's existing technology infrastructure by...Work experience placementWork at officeRemote workFlexible hours$109.8k - $183k
...artifacts rather than getting direct access to environments from day one. This is a ground-up role — you'll help define how reliability engineering works here by mapping systems, writing runbooks, setting baselines, and building the practices this team will run on going...Base plus commissionLocal areaRemote workWorldwide$120k - $165k
...We provide tools, resources and support to enable users to reach their health goals. We are looking for a Software Engineer III - Site Reliability to join the MyFitnessPal PEAS team. As a member of the PEAS team, you'll have the opportunity to positively impact...Full timeRemote workFlexible hours- ...Senior Site Reliability Engineer Remote – Home Based Job Summary We’re partnering with a company in the SaaS space to find a Senior Site Reliability Engineer . In this role, you’ll be part of the IT Operations group responsible for maintaining all environments...Temporary workRemote workWork from homeFlexible hours
- ...Cloud Operations Engineer III Function: Engineering Reports to: Manager, Cloud... ...operations team, responsible for the reliability, observability, performance, and operational... ...experience in cloud operations, site reliability, platform, or DevOps engineering...Temporary workCasual workLocal areaRemote workWorldwideFlexible hours
- ...About the Role We are seeking a Senior Site Reliability Engineer to join our cloud engineering team. You will own the reliability, scalability, and observability of our critical financial SaaS applications and infrastructure, working across cloud platforms to ensure...Remote work
$180k - $230k
...Acceleration Job Description We're looking for a Senior SRE to own the reliability, scalability, and observability of our production systems. You'll work closely with platform and data engineering to keep high-throughput, data-intensive services running at the...Work at officeLocal areaImmediate startRemote work3 days per week$186.82k - $224.18k
...made, and we are not tiptoeing into it. We are rebuilding our engineering culture around a simple belief: AI changes everything. How... ...Is Babylist is looking for a Senior Software Engineer, Site Reliability to join our Platform team. In this position, you will play...Work at officeLocal areaImmediate startRemote workFlexible hoursShift work- ...Senior Site Reliability Engineer Company: Sphera Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Terraform, ARM templates, Kubernetes, Azure, SonarCloud, CheckPoint, Hadoop, Kafka, Presto, NewRelic, CI/CD, Linux, Windows, Redis...Full timeRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!


