Site Reliability Engineer
GrabJobs
About Teleport Don't wait for the future of infrastructure. Be part of it. Teleport is the AI Infrastructure Identity Company. We're solving one of the hardest problems in security: giving every human, machine, workload, and AI agent a cryptographically secured identity, improving engineering velocity while maintaining security. We make trusted computing simple. This gives you the freedom, power, and autonomy to build and innovate with confidence. Remote-first and globally distributed, we work with companies like Nasdaq, IBM, and Elastic to secure infrastructure for an AI world. About the Role Teleport Cloud takes our traditionally open-source and enterprise access plane and provides a SaaS option for our customers to adopt. As such, our team is building our production and software as a service infrastructure from scratch. We tackle the hard problems that allow our customers to trust us for secure and reliable access to their infrastructure. Excellent security is table stakes; a security breach can compromise our customers' infrastructure. We must also balance security with maintaining productivity and building a compelling product offering for our customers. Most of the code you will write will be written in Go. We strongly encourage you to explore our GitHub Repo to get a taste of what we are building. Important information We require to be in-person in our Oakland, CA, office for the onboarding week. We conduct background checks. What You'll do Re-engineer the core teleport product to scale globally and optimize routing latency for teams distributed around the world Re-write portions of the core Teleport product to enable our goals for the cloud product Build out our monitoring and observability stack to alert us to production issues and minimize false positives so we can all get a good sleep at night Work on automation to tackle and eliminate the highest toil activities Execute on traditional operation challenges, such as patching, scaling, backup and restore, disaster recovery, and more Investigate the outages and incidents our customers experience with our product Participate in the on-call rotation to ensure 24/7/365 system uptime. What We're Looking For Willingness to collaboratively work with Teleports’ engineers on coding challenge in Go as part of the interview process. Strong experience in Linux systems, networking, containers, and troubleshooting. Have solid Go and Kubernetes development experience. Strong experience developing scripts, automation, or lightweight programs, submitting patches to the product codebase, or building tooling that incorporates AI agents into operational workflows. AWS Cloud experience is preferred, GCP experience is acceptable. Systems Observability tools: Prometheus, Grafana, Loki etc. Operate and support the observability platform to maintain visibility and reliability. Experience operate in a team where sound security choices are critical, and where reasoning about correctness and system invariants (e.g. formal or property-based methods) is valued. Intellectual curiosity and a willingness to master new technologies. Transparency, honesty, and a no-ego mindset. Excellent communication skills. How We Hire Our process is designed to be straightforward and respectful of your time. We skip performative rounds and focus on what matters: understanding how you think and what you can do. For this role, we use a take-home challenge that mirrors real work at Teleport — on your time, your way. You'll have support from the team throughout. Zoom meeting with a Teleport recruiter. You’ll learn about the company, our products, compensation philosophy, interview process, and key requirements. Zoom meeting with the Hiring Manager or a Lead Engineer. They will walk you through the coding challenge and answer your questions. Coding challenge collaboration. The day after your meeting with the Hiring Manager or Lead Engineer, you join a Slack channel and complete a coding challenge in Go using GitHub. The challenge usually takes about 2 weeks and ends with a Zoom meeting with the Hiring Manager and a member of the interview team to review your solution. If your challenge solution meets our bar, you'll receive an offer to join Teleport. Why Teleport You're joining a company where the problem is real, the team is small, and your work shows up directly in the product. We're not a big company, you won't get lost in a crowd. You'll have the freedom and autonomy to do what you're great at, alongside teammates who care about doing it right and want to see you succeed. The work is collaborative, there's real room to grow, and the mission is one worth showing up for. Remote-first and globally distributed, we're genuinely passionate about what we're building. The Benefits At Teleport, we believe your career is more important than a list of perks. That's why we focus on making your day-to-day the best it can be — giving you the autonomy, access, and support to do the best work of your career. - Extensive health coverage - Annual expense budget - Rest and recovery policies that maximize your ability to recharge - Investment in your future with retirement savings plans - Professional development opportunities Teleport is an equal opportunity employer and does not discriminate against any employee or applicant on the basis of age, color, disability, gender, national origin, race, religion, sexual orientation, veteran status, or any classifications protected by federal, state, or local law. Candidate Privacy Notice: For information about our collection and processing of job applicant personal data for this position, please see our Job Applicant Privacy Policy and Notice of Collection at goteleport.com/legal/apply/job-applicant/
- ...change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability and performance...SuggestedWork at officeRemote workFlexible hours
$115.5k - $164.8k
...mission that matters at a company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a significant... ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on-call rotations...SuggestedWork experience placementWork at officeRemote work$160k - $200k
...data, ideally using promQLKey Responsibilities:Mentor and evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage & remediation processes as a player / coachPerform incident...SuggestedTemporary workWork at officeLocal areaFlexible hours3 days per week$138.1k - $198.2k
...more intuitive with technology that simply works. The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments... ...Our customers are all engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter of our engineering...SuggestedPermanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours$93.6k - $170.64k
We currently have a career opportunity for a NOC AI-Ops Engineer to join our team located in Boston, MA. This is a hybrid role, 3 days... ....Job Overview:We are seeking a Senior AIOps and Incident/Site Reliability Engineer to lead incident management, operational resilience,...SuggestedWork at officeLocal areaNight shift3 days per week$130k - $150k
...systems and hybrid infrastructure, meaning experience with cloud technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are reliable, scalable, and performant across on-premises and cloud...Work at officeWork from home3 days per week$127k - $249k
The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper)....Work at officeLocal areaRemote workWorldwideFlexible hours- ...commuting distance of one of our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team, you will operate the... ...ideal candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based...Full time
$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours- ...and AI agent a cryptographically secured identity, improving engineering velocity while maintaining security. We make trusted computing... ...problems that allow our customers to trust us for secure and reliable access to their infrastructure. Excellent security is table stakes...Work at officeLocal areaRemote workSleeping nights
$140k - $205k
...Senior Technology Site Reliability Engineer Cooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operationsteam. Position summary: The Senior Technology Site Reliability Engineer("SRE") is responsible for ensuring the reliability...Full timeTemporary workWork at officeFlexible hoursWeekend work$121.4k - $218.6k
...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with... ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust...Work experience placementWork at office- ...scale, we invite you to bring your talents to Zscaler to help shape the future of cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or Bellevue WA) to join our Zero Trust Exchange team. This is a remote role based in San...InternshipWork at officeLocal areaRemote workWorldwide
- ...Site Reliability Engineer Cambridge, MA About Watershed Our vision is to become the leading biocomputing platform. The future of biology is in big data analysis, and we are on a mission to accelerate digital drug discovery with the Watershed platform. Watershed...
$140k - $210.9k
...States. The position will be primarily on-site with residency commutable to one of our... .../DevOps backgrounds or software engineering backgrounds (e.g., Java Python, Go) with... ...strong interest in operating and improving reliability of distributed production systems. Responsibilities...Full timeTemporary workPart timeWork at officeShift work$160k - $200k
...Senior Site Reliability Engineer This role is located in Somerville, MA - We are a hybrid work environment and are in the office 3+ days/per week. Tulip, the leader in AI-native frontline operations, is helping companies around the world equip their workforce with...Temporary workWork at officeLocal areaFlexible hours3 days per week- ...itD is seeking a Site Reliability Engineer to develop and enhance automation solutions that improve the reliability, scalability, and operational efficiency of large-scale cloud infrastructure. The ideal candidate will bring hands-on experience in site reliability engineering...Work experience placementRemote work
$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...Temporary workImmediate startFlexible hoursShift work$128k - $160k
...time. To those who see AI as a driver of progress, come build the future together. The Crown Is Yours As a Senior Site Reliability Engineer, you'll build and scale the critical infrastructure behind every product. In this role, you'll take on complex challenges...Full timeImmediate start$166k - $220k
...failure. As such, it is critical that Anduril services are reliable and maintainable. This means that all services &... ...ground systems & Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build & operate Anduril...Full timeWork experience placementImmediate start$160k - $200k
Role Overview We are looking for a Control System Engineer/Site Reliability Engineer (SRE) to integrate and maintain the hardware and software systems that enable QuEra’s quantum controls and software stack. You’ll work closely with software engineers, physicists, hardware...Local areaRemote work$127k - $249k
Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that... ...maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering...Local areaWorldwideFlexible hours$130k - $180k
...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and... ...an in-house AI R&D team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to...Temporary workWork at officeImmediate startRemote workFlexible hours$160k - $240k
...passionate about building unified IT solutions that simplify the way IT organizations work. We are currently looking for a Senior Site Reliability Engineer to join our SRE team in the Platform Engineering organization and help us scale our products to millions of end-users. We...Permanent employmentFull timeRemote workWork from homeRelocationFlexible hours- ...Information Technology group delivers secure, reliable technology solutions that enable DTCC to... ...This RoleAs a Senior Application Support Engineer, you will help power DTCC's global... ...trade processing and settlement.Leveraging Site Reliability Engineering (SRE) principles,...Remote workFlexible hours
$105.79k - $141.05k
...shape the future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role...Temporary workRemote work$160k - $225k
...Staff Site Reliability Engineer Manifold is the AI platform for life sciences, accelerating life-changing medicines to patients. Our products speed up workflows in areas from target identification and clinical development to market access and precision medicine in the...$130k - $140k
...mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionSite Reliability EngineerLocation(s): Waltham, MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to ensuring systems are reliable, scalable, and efficient...Ongoing contractFull timeTemporary workWork experience placement$166k - $220k
...through to the field. When something breaks in a deployed environment, we fix it. About the Role We're looking for a Site Reliability Engineer to join the Imaging team. This is not a product development role, and it isn't a traditional cloud-SRE role either. You...Full timeWork experience placementImmediate startRemote workWeekend workDay shift$139k - $257.55k
...Individual Contributor The Challenge The Adobe Creative Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning, autonomous AI workflows, and cloud-native infrastructure. Adobe Stock...Temporary workLocal areaRemote workRelocation
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer Boston, MA
- site reliability engineer sre Boston, MA
- site reliability engineer remote Boston, MA
- site services specialist Boston, MA
- construction site safety Boston, MA
- site leader Boston, MA
- official site Boston, MA
- website content developer Boston, MA
- on site coordinator Boston, MA
- IT site lead Boston, MA

