Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

$130k - $180k
Full-time

Nebius

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to work in our office in Amsterdam. Hardware Infrastructure team designs, develops and supports systems involved in the data-centers lifecycle: - Serving functional and load testing system. - Monitoring of engineering equipment located in our data centers (power supply, air and water cooling, etc. ) - Monitoring of IT equipment: racks, servers, JBODs, JBOGs, power shelves, network devices, etc. - Asset tracking. - Hardware repairs tasks tracking. - Server production.

In this position, your responsibility will be to : - Ensure fault-tolerance, scale and uninterrupted operations for our services. - Use cutting-edge technology to solve a variety of infrastructure problems. - Implement and improve CI/CD processes. We expect you to have :  - Proficiency in Linux systems, with expertise in Python and Bash scripting for automation. - Demonstrated ability to troubleshoot complex system issues, including hardware, software and networking problems. - Strong analytical and problem-solving skills, with a focus on optimizing system performance. - Working proficiency in English. It would be an added bonus if you had : - Desire to be involved in backend development. - Experience designing, developing and running high-load distributed systems. Working conditions: - Primarily remote  - Occasional travel to data centers required, especially if not located near one - Collaboration with globally distributed engineering and operations teams Key employee benefits: - Health insurance: 100% company-paid medical, dental, and vision coverage for employees and families - 401(k) plan: up to 4% company match with immediate vesting - Parental leave: 20 weeks paid for primary caregivers, 12 weeks for secondary caregivers - Remote work reimbursement: up to $85/month for mobile and internet - Disability & life insurance: company-paid short-term, long-term, and life insurance coverage Compensation - We offer competitive salaries, ranging from $130k- $180k base + quarterly performance bonuses. Join Nebius and help operate the systems that power next-generation AI infrastructure. Benefits & Perks: - Competitive compensation - Career growth and learning opportunities - Flexibility and ownership - Collaborative and innovative culture - Opportunity to work on impactful AI projects - International environment and talented teams What's it like to work at Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI  Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law.

Vacancy posted 20 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Remote vacancy
  •  ...to meet you. Our Enterprise Information Technology (EIT) organization is expanding, and we are seeking a Senior Site Reliability Engineer to help drive a major architectural modernization. In this role, you will move beyond traditional infrastructure maintenance... 
    Suggested
    Permanent employment
    Full time
    H1b
    Local area
    Remote work
    Shift work

    Jack Henry & Associates

    New York, NY
    13 hours ago
  •  ...and performance our customers have come to expect, and help raise the reliability bar as we grow. What you would do: Design, build, and operate the shared platform foundations engineers ship on every day: GCP infrastructure, Kubernetes, networking, routing,... 
    Suggested
    Remote work
    Worldwide
    Flexible hours

    Sanity

    United States
    13 hours ago
  • $147k - $168k

     ...Inc. as one of the most innovative and fastest-growing technology companies in the country. Role Summary As a Site Reliability Engineer at Filevine, you will improve the reliability, scalability, and operational maturity of the Filevine platform. You’ll... 
    Suggested
    Full time
    Temporary work
    Work experience placement
    Work at office
    Remote work
    2 days per week
    3 days per week

    Filevine

    United States
    13 hours ago
  •  ...Site Reliability Engineer Company: Milestone Systems Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Golang, Python, Linux, Shell scripting, Kubernetes, Docker, Terraform, CI/CD, GitOps, ArgoCD, Spinnaker, Prometheus, Datadog,... 
    Suggested
    Full time
    Remote work

    Milestone Systems Inc

    United States
    13 hours ago
  •  ...grow, adapt and use your skills consistently. Our customers rely on us in the moments that matter. Engineering delivers on that promise.   The Senior Site Reliability Engineer is responsible for ensuring our SaaS products are fast, stable and optimized for our customers... 
    Suggested
    Work experience placement
    Remote work
    Flexible hours

    Donnelley Financial Solutions

    United States
    4 days ago
  • $141.8k - $195k

     ...work, grow fast, and bring their full selves to the herd. Why You’ll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in the U.S. Cribl provides... 
    Temporary work
    Remote work

    Cribl

    United States
    13 hours ago
  •  ...encourage you to apply. The Role  As a Senior Platform Engineer, you are a champion for DevOps and SRE culture and industry...  ...met. \n What You Will Be Doing Improving production reliability and system resilience within an SRE scoped team Championing... 
    Remote work
    Flexible hours

    Megaport

    United States
    4 days ago
  •  ...Site Reliability Engineer Company: GitLab Work Type: Remote Employment: Full Time Location: CA, US Seniority: Senior Level Technologies: Terraform, Ansible, Kubernetes, Go, Ruby, Jsonnet, Prometheus, ELK, Grafana Requirements: Senior-level SRE with strong Terraform/IaC... 
    Full time
    Remote work

    GitLab

    United States
    13 hours ago
  •  ...Senior Site Reliability Engineer Company: CyberArk Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: AWS, Kubernetes, Terraform, CloudFormation, Ansible, CloudWatch, Grafana, Datadog, OpenSearch, PagerDuty Requirements: Senior SRE... 
    Full time
    Remote work

    CyberArk

    United States
    13 hours ago
  •  ...Site Reliability Engineer OXIO is the first NeoTelco. We arebuilding the world’s largest, most accessible, and insightful Telecom network. Our platform empowers anyone to spin up their own carrier from a browser, scaling and supporting you as you scale your network... 
    Remote work

    OXIO

    United States
    13 hours ago
  •  ...Job Title:  Site Reliability Engineer (Azure Government & Infrastructure) Pay Type : SALARIED EXEMPT  Location:  Remote Citizenship Requirement: U.S. Citizen (Required) Summary of Position Role/Responsibilities The Site Reliability Engineer (SRE) for... 
    Full time
    Remote work
    Monday to Friday

    Quzara LLC

    United States
    13 hours ago
  • $7.5k

     ...investment management. We have become a multibillion-dollar asset manager, and we have ambitious goals for the future. As a Site Reliability Engineer (SRE), you will work at the intersection of production operations and software development as you improve, manage, and... 
    Local area
    Remote work

    The Voleon Group

    United States
    13 hours ago
  •  ...healthcare organizations maintain accurate, compliant, and reliable provider networks at scale. Our vision is simple: One...  ...patients. About the Role We're looking for a Senior Site Reliability Engineer who takes ownership seriously - someone who designs for... 
    Remote work

    CertifyOS

    United States
    13 hours ago
  • $152k - $195k

     ...including Silver Lake Waterman, Moody’s, Sequoia Capital, GV and Riverwood Capital. About the Team: As a Senior Site Reliability Engineer, you will be a key technical leader driving the design and optimization of our Kubernetes-based infrastructure and CI/CD systems... 
    Remote work

    SecurityScorecard

    United States
    1 day ago
  •  ...About The Role: We're looking for a Senior Site Reliability Engineer to help us mature and scale the infrastructure behind our multi-cloud SaaS platform. Most of our footprint runs on Microsoft Azure, built from the ground up around cloud architecture principles:... 
    Remote work
    Flexible hours

    Dental Intelligence

    United States
    4 days ago
  • $145k - $193k

     ...entertainment, we want to talk to you. About the Role & Team The SRE team at PENN Entertainment is looking for a Senior Site Reliability Engineer to help build and operate the infrastructure behind a large-scale sports betting and media platform. You'll own critical... 
    Remote work

    Penn Interactive

    United States
    1 day ago
  • $160k - $180k

     ...big impact. See Arkestro in action at arkestro.com. About the Role Arkestro is hiring for a Senior SRE Engineer to manage our performance and reliability for our software platform and infrastructure. The right candidate will own and develop our infrastructural... 
    Local area
    Remote work

    Arkestro

    United States
    13 hours ago
  •  ...Site Reliability Engineer Company: Crunchafi Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Azure, AKS, Azure Kubernetes Service, Terraform, Bicep, ARM templates, GitHub Actions, Azure DevOps, Kubernetes, Docker, App Insights... 
    Full time
    Remote work

    Crunchafi

    United States
    13 hours ago
  •  ...Site Reliability Engineer Company: Quzara Work Type: Remote Employment: Full Time Location: US Seniority: Mid Level Technologies: Azure, Terraform, Bicep, Ansible, Azure Monitor, Azure Automation, Azure Policy, Azure Site Recovery, TLS/SSL Requirements: 4+ years in SRE... 
    Full time
    Remote work

    Quzara LLC

    United States
    13 hours ago
  •  ...The Site Reliability Engineer (SRE) / Subject Matter Expert (SME) – Computer Systems Engineer/Architect will provide senior-level reach-back expertise to support the reliability, scalability, performance, and operational resilience of the GEOMAP platform in secure cloud... 
    Full time
    Contract work
    For contractors
    For subcontractor
    Remote work

    Diné Development

    United States
    13 hours ago
  •  ...organization across multiple locations in the US, South America, and India. Location: Remote (US-Based Candidates Only) Site Reliability Engineer II (SRE) Position Overview We are seeking a Site Reliability Engineer II (SRE) to join our growing Site... 
    Remote work
    Flexible hours
    Shift work
    Weekday work

    NationsBenefits, LLC

    United States
    13 hours ago
  • $160k - $208k

     ...early diagnosis and longitudinal care management of chronic conditions. We are looking for an experienced Site Reliability and Infrastructure Engineer to join our engineering team. You will support Counterpart Health's existing technology infrastructure by... 
    Work experience placement
    Work at office
    Remote work
    Flexible hours

    Clover Health

    United States
    5 days ago
  • $109.8k - $183k

     ...artifacts rather than getting direct access to environments from day one. This is a ground-up role — you'll help define how reliability engineering works here by mapping systems, writing runbooks, setting baselines, and building the practices this team will run on going... 
    Base plus commission
    Local area
    Remote work
    Worldwide

    Veeam Software

    United States
    13 hours ago
  • $120k - $165k

     ...We provide tools, resources and support to enable users to reach their health goals. We are looking for a Software Engineer III - Site Reliability to join the MyFitnessPal PEAS team. As a member of the PEAS team, you'll have the opportunity to positively impact... 
    Full time
    Remote work
    Flexible hours

    MyFitnessPal

    United States
    13 hours ago
  •  ...Senior Site Reliability Engineer Remote – Home Based Job Summary We’re partnering with a company in the SaaS space to find a Senior Site Reliability Engineer . In this role, you’ll be part of the IT Operations group responsible for maintaining all environments... 
    Temporary work
    Remote work
    Work from home
    Flexible hours

    SourceDirect Talent

    United States
    2 days ago
  •  ...Cloud Operations Engineer III Function: Engineering  Reports to: Manager, Cloud...  ...operations team, responsible for the reliability, observability, performance, and operational...  ...experience in cloud operations, site reliability, platform, or DevOps engineering... 
    Temporary work
    Casual work
    Local area
    Remote work
    Worldwide
    Flexible hours

    On-Board Services

    United States
    13 hours ago
  •  ...About the Role We are seeking a Senior Site Reliability Engineer to join our cloud engineering team. You will own the reliability, scalability, and observability of our critical financial SaaS applications and infrastructure, working across cloud platforms to ensure... 
    Remote work

    MeridianLink

    United States
    13 hours ago
  • $180k - $230k

     ...Acceleration Job Description We're looking for a Senior SRE to own the reliability, scalability, and observability of our production systems. You'll work closely with platform and data engineering to keep high-throughput, data-intensive services running at the... 
    Work at office
    Local area
    Immediate start
    Remote work
    3 days per week

    GridCARE, Inc.

    Eastern, KY
    1 day ago
  • $186.82k - $224.18k

     ...made, and we are not tiptoeing into it. We are rebuilding our engineering culture around a simple belief: AI changes everything. How...  ...Is Babylist is looking for a Senior Software Engineer, Site Reliability to join our Platform team. In this position, you will play... 
    Work at office
    Local area
    Immediate start
    Remote work
    Flexible hours
    Shift work

    Babylist

    United States
    13 hours ago
  •  ...Senior Site Reliability Engineer Company: Sphera Work Type: Remote Employment: Full Time Location: US Seniority: Senior Level Technologies: Terraform, ARM templates, Kubernetes, Azure, SonarCloud, CheckPoint, Hadoop, Kafka, Presto, NewRelic, CI/CD, Linux, Windows, Redis... 
    Full time
    Remote work

    Sphera

    United States
    13 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!