Remote Senior Site Reliability Engineer
GrabJobs
About Us Epic Kids is the leading digital reading platform built for kids 12 and under, trusted by millions of children, educators, and families around the world. Our mission is to inspire a lifelong love of reading by providing unlimited access to thousands of high-quality books, videos, and educational content through a safe and engaging experience. We combine technology, storytelling, and learning innovation to help every child become a confident reader. At Epic, you'll join a collaborative and fast-paced global team passionate about building meaningful products that make a real impact on children's education and literacy. About the Role We're looking for a Senior Site Reliability Engineer to drive the stability, observability, and reliability of Epic's platform as we grow. You are an experienced engineer who works independently on complex infrastructure problems, makes sound technical decisions, and helps raise the bar for the engineers around you. You will own pieces of our GCP infrastructure, container platform, CI/CD pipelines, and observability stack—setting reliability standards, hardening the systems behind them, and making sure issues are caught early and resolved fast. You will partner closely with both our product engineering and data engineering teams to keep the platforms that power their applications and workflows running reliably. This is a fully remote, US-based role working closely with a global engineering team. What You'll Do Drive the reliability of Epic's infrastructure—set and track SLOs/SLIs, reduce toil, and engineer out recurring instability. Build and operate the cloud infrastructure and container platform for high availability, scalability, and cost efficiency—including workload scheduling, autoscaling, networking, and graceful failure handling. Maintain and improve CI/CD pipelines for fast, safe delivery across engineering teams. Own and evolve the observability stack—metrics, logs, traces, dashboards, and alerts. Manage infrastructure as code across the organization, with a focus on consistency, change safety, and reproducibility. Own platform security practices—including secrets management, IAM policies, and network segmentation. Support compliance-aware infrastructure practices—including vulnerability management, access reviews, audit-evidence flows, and incident-response readiness. Participate in a frequent on-call rotation; drive incident response, blameless post-mortems, and follow-through on systemic fixes. Partner with product and data engineering teams to troubleshoot platform issues and guide developers on infrastructure best practices. What We're Looking For Required Qualifications Bachelor's degree or higher in Computer Science, Software Engineering, or a related field. 5+ years of experience in infrastructure, platform, DevOps, or a related engineering role, with a track record of measurably improving production reliability—including defining SLOs, reducing incident frequency or MTTR, and eliminating recurring failure modes. Hands-on experience with Google Cloud Platform (GCP), including GCE, GCS, VPC, IAM, Cloud Monitoring, and related services. Experience with Docker and Kubernetes (GKE), including containerizing workloads, Helm, and cluster fundamentals. Experience with CI/CD pipelines such as GitHub Actions, ArgoCD, Jenkins, or similar tools. Experience with an observability platform such as New Relic, including metrics, logging, alerting, and dashboards. Proficiency with Terraform for managing infrastructure as code. Scripting or programming experience with Python, Bash, or similar languages. Preferred Qualifications Experience operating workflow orchestration platforms such as Dagster or Airflow as a service for data or platform teams. Familiarity with PromRelay for metrics forwarding and alert routing. Familiarity with the operational footprint of data platforms, including warehouse infrastructure, job schedulers, and batch workloads. Experience working within distributed or global engineering teams. Working knowledge of compliance frameworks such as SOC 2, FERPA, and COPPA, as well as GRC tools. Proficiency in Mandarin Chinese is a plus. Why You'll Love Working at Epic Join a mission-driven company making a meaningful impact on children's literacy and education. Work alongside talented teammates in a collaborative, supportive, and global environment. Enjoy the flexibility of a fully remote, U.S.-based position. Help build and scale the infrastructure powering millions of young readers around the world. Salary - 160K to 200K (bonus included)
$65 - $75 per hour
DescriptionKforce has a client seeking a remote Senior Site Reliability Engineer to be a l be a leading member of the team working with a diverse range of technologies. You will enjoy working in a friendly environment and benefit from our investment in staff. The role also...Remote workSenior- ...Meghana GorusuCompany: SRI Tech SolutionsJob Title: Senior Site Reliability EngineerLocation: Plano , TX (remote)Years of Experience: 8 to 15 yearsSkillsKubernetes... ...seeking a highly skilled Senior Site Reliability Engineer (SRE) to join our dynamic team. The ideal...Remote workSenior
$104.9k - $174.7k
...the link below, About the Role:We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability... ...may work a hybrid schedule. If not, this role is fully remote. We do not restrict applicants based on job site or...Remote workSeniorFull timeWork at officeLocal areaWork from home$90k - $180k
...serve people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale... ..., and operational excellence of Merlin.net — a remote monitoring platform designed to help doctors, cardiologists...Remote workSenior- ...Job Title Location Remote - United States Job Category Information Technology, Platform Engineering, Site Reliability Engineering Industry Computer Software, SaaS, National Security Employee Type FT Exempt Manage Others No Minimum Experience 5 Years...Remote workSenior
$15k
...beautiful modern office, daily catered lunches, and more.As a Senior Cluster Site Reliability Engineer (SRE), you will help scale our research compute cluster... ...Compensation Range: $205K - $235KLocationBerkeley, CA; Remote, United StatesEmployment TypeFull timeLocation...Remote workSeniorWork at officeLocal area$150k - $180k
...operates through three business units: Remote Sensing (the data), Space Systems (the... ...are seeking an experienced SeniorSite Reliability Engineer to help design, build, operate, and scale... ...organization.This position is based on-site in either our Arlington, VA office, Reston...Remote workSeniorPermanent employmentFull timeWork at officeLocal areaWorldwide$86.9k - $198k
Site Reliability Engineer, SeniorThe Opportunity: Engineering to make a system more resilient and efficient frees up time and money to build more... ...expected to have their cameras on during meetings.Remote: If this position is listed as remote, there may still be occasions...Remote workSeniorFull timeContract workPart timeWork at officeLocal area$117k - $209.33k
...6Position OverviewWant to help make a better world? As a Senior Site Reliability Engineer at Autodesk, you can help us build and operate reliable,... ...internally (not on this external site).SummaryLocation: Idaho, USA - Remote; AMER - United States - Texas - PlanoType: Full timeRemote workSeniorFull timeFor contractors- Site Reliability Engineers are responsible for ensuring the availability, reliability, scalability, and performance of the firm’s most critical customer... ....This is an on-site position located in Springfield, MO. Remote work is not an option for this position.Primary...Remote workSeniorLocal areaFlexible hoursShift work
$140k - $150k
WORK OPTION: Remote_________________The NBA is hiring a Senior Site Reliability Engineer (SRE) - Messaging & Collaboration to ensure the availability, performance, and reliability of enterprise messaging and collaboration platforms, including Microsoft Exchange Online (...Remote workSeniorFull timeTemporary workLocal areaWeekend work- ...A leading livestream shopping platform is seeking a Senior Software Engineer for the Logistics Platform team. This role focuses on improving logistical... ...operational debugging. The position offers flexibility for remote work and benefits including health insurance and generous...Remote workSenior
$127k - $249k
...on a hybrid basis, or it can be fully remote while working from a location based in... ...zones. We are looking for an experienced Senior Engineer for our SRE, Atlas team to support,... ...Role OverviewWe are seeking a talented Site Reliability Engineer (SRE) with a strong...Remote workSeniorLocal areaWorldwideFlexible hours$130k - $180k
...belonging at iManage. Mondays and Fridays are reserved for (remote-friendly) focus time to get things done. Have the best of... ...belonging, collaboration, and accomplishment.Being a Senior Site Reliability Engineer at iManage Means… You are an engineer, a builder, and a systems...Remote workSeniorWork at officeLocal areaWorldwideMonday to FridayFlexible hours$118.6k - $195.68k
Job SummaryThe Red Hat IT OpenShift team is looking for a Senior Site Reliability Engineer (SRE) to design, develop, scale, and operate our Red Hat... ...for bonus, commission, and/or equity. For positions with Remote-US locations, the actual salary range for the position may...Remote workSeniorPermanent employmentFull timeContract workWork experience placementWork at officeFlexible hours$139k - $155k
...LinkedIn.**Travel to Office expectations**For Remote Roles: If this role is remote, there will be in... ...:The AI SRE team is a focused group of SRE engineers dedicated to making PointClickCare's AI and ML platforms reliable, secure, and operationally excellent — from data...Remote workSeniorFull timeWork at office$112.7k - $193.2k
...We are seeking a highly experienced Senior Software Engineering Manager to lead the strategy, architecture... ...'ll enjoy the flexibility to work remotely * from anywhere within the U.S. as... ...testing automation, observability, reliability, and operational readinessPartner...Remote workSeniorMinimum wageFull timeWork experience placementWork at officeLocal area$127k - $249k
We are looking for an experienced Senior or Staff Engineer for our SRE, InfraSec team, to guide the security of our cloud-based infrastructure.... ...San Francisco offices on a hybrid basis, or it can be fully remote while working from a location based in either Eastern or...Remote workSeniorLocal areaWorldwideFlexible hours- ...software developers, platform engineers, and IT staff to improve... ...requirements, service quality, reliability, security, and compliance needs... ...Required: 8+ years of experience in Site Reliability Engineering,... ...or equivalent experience REMOTE WORK NOTICE: This position may...Remote workSeniorWork at office
$175k - $250k
...Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or be willing to... ...ensuring scalability, performance, and reliability across environments. What You’ll Do Design...Remote workSeniorFull timeRelocationRelocation package- ...Cassandra, SQL Server, My SQL and Mongo DB Seniority level Seniority level Mid-Senior... ...in to set job alerts for “Senior Site Reliability Engineer” roles. Bellevue, WA $204,000.00-$259... ...ago Site Reliability Engineer (SRE, Remote US) Seattle, WA $120,000.00-$160,000....Remote workSeniorContract work
- ...Description The Senior Site Reliability Engineer (SRE) will implement, secure, and operate the cloud infrastructure that supports CenCore Group... ...role is primarily performed in a professional office or remote technology environment, depending on business needs and...Remote workSeniorWork at office
- ...security, user fund transparency, trading engine speed, deep liquidity, and an... ...around the world. We’re looking for a Senior Site Reliability Engineer Engineer to take ownership of... ...and performance. This is a full-time remote role , with a preference for candidates...Remote workSeniorFull timeWork from home
$120k - $175k
...together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering... ...applicants from anywhere in the U.S. and are willing to consider remote candidates. #LI-Remote Working at PrizePicks: The...Remote workSeniorFull timeWork visaFlexible hours$149.4k - $202k
...Senior Software Engineer- Site Reliability Engineering (SRE) DC, MD, VA, CA The Site Reliability Engineering discipline at Noctua Technology, LLC is... ...applications and infrastructure. Location : Primarily Remote. Candidates must be based in CA or DC Metro Area for...Remote workSenior$141.8k - $195k
...massive, fast‑moving market. With a global workforce, we’re remote‑first and grounded in a simple idea: software is a people... ...herd. Why You’ll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of...Remote workSeniorTemporary work- ...cryptographically secured identity, improving engineering velocity while maintaining security. We... ...to build and innovate with confidence. Remote-first and globally distributed, we work... ...customers to trust us for secure and reliable access to their infrastructure. Excellent...Remote workSeniorWork at officeLocal areaSleeping nights
- ...education and literacy. About the Role We're looking for a Senior Site Reliability Engineer to drive the stability, observability, and reliability of... ...and workflows running reliably. This is a fully remote, US-based role working closely with a global engineering...Remote workSenior
- ...Senior Site Reliability Engineer (Enterprise Platform) Location: Remote - US - Open to Europe if happy to overlap with EST Compensation: Competitive We are a high-growth software company supporting the development of a premier open-source, EVM-compatible public ledger...Remote workSeniorContract workCurrently hiring
- ...Senior Site Reliability Engineer - AI Infrastructure Location: Global Remote / San Francisco · Full-Time Andromeda Cluster was founded by Nat Friedman and Daniel Gross to give early-stage startups access to the kind of scaled AI infrastructure once reserved only...Remote workSeniorFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Remote Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Raleigh, NC
- site reliability engineer sre Raleigh, NC
- remote clinical Raleigh, NC
- junior front end developer remote Raleigh, NC
- remote team lead Raleigh, NC
- ai engineer remote Raleigh, NC
- remote sales director Raleigh, NC
- remote auto claims adjuster Raleigh, NC
- salesforce remote Raleigh, NC
- entry level project manager remote Raleigh, NC

