Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

GrabJobs

About Teleport Don't wait for the future of infrastructure. Be part of it. Teleport is the AI Infrastructure Identity Company. We're solving one of the hardest problems in security: giving every human, machine, workload, and AI agent a cryptographically secured identity, improving engineering velocity while maintaining security. We make trusted computing simple. This gives you the freedom, power, and autonomy to build and innovate with confidence. Remote-first and globally distributed, we work with companies like Nasdaq, IBM, and Elastic to secure infrastructure for an AI world. About the Role Teleport Cloud takes our traditionally open-source and enterprise access plane and provides a SaaS option for our customers to adopt. As such, our team is building our production and software as a service infrastructure from scratch. We tackle the hard problems that allow our customers to trust us for secure and reliable access to their infrastructure. Excellent security is table stakes; a security breach can compromise our customers' infrastructure. We must also balance security with maintaining productivity and building a compelling product offering for our customers. Most of the code you will write will be written in Go. We strongly encourage you to explore our GitHub Repo to get a taste of what we are building. Important information We require to be in-person in our Oakland, CA, office for the onboarding week. We conduct background checks. What You'll do Re-engineer the core teleport product to scale globally and optimize routing latency for teams distributed around the world Re-write portions of the core Teleport product to enable our goals for the cloud product Build out our monitoring and observability stack to alert us to production issues and minimize false positives so we can all get a good sleep at night Work on automation to tackle and eliminate the highest toil activities Execute on traditional operation challenges, such as patching, scaling, backup and restore, disaster recovery, and more Investigate the outages and incidents our customers experience with our product Participate in the on-call rotation to ensure 24/7/365 system uptime. What We're Looking For Willingness to collaboratively work with Teleports’ engineers on coding challenge in Go as part of the interview process. Strong experience in Linux systems, networking, containers, and troubleshooting. Have solid Go and Kubernetes development experience. Strong experience developing scripts, automation, or lightweight programs, submitting patches to the product codebase, or building tooling that incorporates AI agents into operational workflows. AWS Cloud experience is preferred, GCP experience is acceptable. Systems Observability tools: Prometheus, Grafana, Loki etc. Operate and support the observability platform to maintain visibility and reliability. Experience operate in a team where sound security choices are critical, and where reasoning about correctness and system invariants (e.g. formal or property-based methods) is valued. Intellectual curiosity and a willingness to master new technologies. Transparency, honesty, and a no-ego mindset. Excellent communication skills. How We Hire Our process is designed to be straightforward and respectful of your time. We skip performative rounds and focus on what matters: understanding how you think and what you can do. For this role, we use a take-home challenge that mirrors real work at Teleport — on your time, your way. You'll have support from the team throughout. Zoom meeting with a Teleport recruiter. You’ll learn about the company, our products, compensation philosophy, interview process, and key requirements. Zoom meeting with the Hiring Manager or a Lead Engineer. They will walk you through the coding challenge and answer your questions. Coding challenge collaboration. The day after your meeting with the Hiring Manager or Lead Engineer, you join a Slack channel and complete a coding challenge in Go using GitHub. The challenge usually takes about 2 weeks and ends with a Zoom meeting with the Hiring Manager and a member of the interview team to review your solution. If your challenge solution meets our bar, you'll receive an offer to join Teleport. Why Teleport You're joining a company where the problem is real, the team is small, and your work shows up directly in the product. We're not a big company, you won't get lost in a crowd. You'll have the freedom and autonomy to do what you're great at, alongside teammates who care about doing it right and want to see you succeed. The work is collaborative, there's real room to grow, and the mission is one worth showing up for. Remote-first and globally distributed, we're genuinely passionate about what we're building. The Benefits At Teleport, we believe your career is more important than a list of perks. That's why we focus on making your day-to-day the best it can be — giving you the autonomy, access, and support to do the best work of your career. - Extensive health coverage - Annual expense budget - Rest and recovery policies that maximize your ability to recharge - Investment in your future with retirement savings plans - Professional development opportunities Teleport is an equal opportunity employer and does not discriminate against any employee or applicant on the basis of age, color, disability, gender, national origin, race, religion, sexual orientation, veteran status, or any classifications protected by federal, state, or local law. Candidate Privacy Notice: For information about our collection and processing of job applicant personal data for this position, please see our Job Applicant Privacy Policy and Notice of Collection at goteleport.com/legal/apply/job-applicant/

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Boston, MA vacancy
  •  ...change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability and performance... 
    Suggested
    Work at office
    Remote work
    Flexible hours

    Axon

    Boston, MA
    22 hours ago
  • $115.5k - $164.8k

     ...mission that matters at a company where you matter.Your ImpactAs an engineer on the APX SRE CloudOps team, you will spend a significant...  ...that replace what previously required human intervention with reliable, tested automation. You will also participate in on-call rotations... 
    Suggested
    Work experience placement
    Work at office
    Remote work

    Axon

    Boston, MA
    3 days ago
  • $160k - $200k

     ...data, ideally using promQLKey Responsibilities:Mentor and evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage & remediation processes as a player / coachPerform incident... 
    Suggested
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    Tulip Interface

    Somerville, MA
    1 day ago
  • $138.1k - $198.2k

     ...more intuitive with technology that simply works.  The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments...  ...Our customers are all engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter of our engineering... 
    Suggested
    Permanent employment
    Full time
    Temporary work
    Work experience placement
    Local area
    Remote work
    Flexible hours

    CISCO Systems

    Boston, MA
    5 days ago
  • $93.6k - $170.64k

    We currently have a career opportunity for a NOC AI-Ops Engineer to join our team located in Boston, MA. This is a hybrid role, 3 days...  ....Job Overview:We are seeking a Senior AIOps and Incident/Site Reliability Engineer to lead incident management, operational resilience,... 
    Suggested
    Work at office
    Local area
    Night shift
    3 days per week

    Perficient

    Boston, MA
    4 days ago
  • $130k - $150k

     ...systems and hybrid infrastructure, meaning experience with cloud technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are reliable, scalable, and performant across on-premises and cloud... 
    Work at office
    Work from home
    3 days per week

    CRA International

    Boston, MA
    5 days ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper).... 
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    3 days ago
  •  ...commuting distance of one of our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team, you will operate the...  ...ideal candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based... 
    Full time

    Federal Reserve Bank of Boston

    Boston, MA
    4 days ago
  • $95k - $171k

     .... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts... 
    Permanent employment
    Work experience placement
    Work at office
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Akamai

    Cambridge, MA
    4 days ago
  •  ...and AI agent a cryptographically secured identity, improving engineering velocity while maintaining security. We make trusted computing...  ...problems that allow our customers to trust us for secure and reliable access to their infrastructure. Excellent security is table stakes... 
    Work at office
    Local area
    Remote work
    Sleeping nights

    GrabJobs

    Boston, MA
    2 days ago
  • $140k - $205k

     ...Senior Technology Site Reliability Engineer Cooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operationsteam. Position summary: The Senior Technology Site Reliability Engineer("SRE") is responsible for ensuring the reliability... 
    Full time
    Temporary work
    Work at office
    Flexible hours
    Weekend work

    Cooley

    Boston, MA
    4 days ago
  • $121.4k - $218.6k

     ...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner with...  ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling robust... 
    Work experience placement
    Work at office

    Akamai

    Boston, MA
    4 days ago
  •  ...scale, we invite you to bring your talents to Zscaler to help shape the future of cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or Bellevue WA) to join our Zero Trust Exchange team. This is a remote role based in San... 
    Internship
    Work at office
    Local area
    Remote work
    Worldwide

    GrabJobs

    Boston, MA
    2 days ago
  •  ...Site Reliability Engineer Cambridge, MA About Watershed Our vision is to become the leading biocomputing platform. The future of biology is in big data analysis, and we are on a mission to accelerate digital drug discovery with the Watershed platform. Watershed... 

    Watershed Informatics

    Cambridge, MA
    3 days ago
  • $140k - $210.9k

     ...States. The position will be primarily on-site with residency commutable to one of our...  .../DevOps backgrounds or software engineering backgrounds (e.g., Java Python, Go) with...  ...strong interest in operating and improving reliability of distributed production systems. Responsibilities... 
    Full time
    Temporary work
    Part time
    Work at office
    Shift work

    Federal Reserve Bank

    Boston, MA
    2 days ago
  • $160k - $200k

     ...Senior Site Reliability Engineer This role is located in Somerville, MA - We are a hybrid work environment and are in the office 3+ days/per week. Tulip, the leader in AI-native frontline operations, is helping companies around the world equip their workforce with... 
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    Tulip Interfaces

    Somerville, MA
    12 hours ago
  •  ...itD is seeking a Site Reliability Engineer to develop and enhance automation solutions that improve the reliability, scalability, and operational efficiency of large-scale cloud infrastructure. The ideal candidate will bring hands-on experience in site reliability engineering... 
    Work experience placement
    Remote work

    GrabJobs

    Boston, MA
    2 days ago
  • $81.1k - $187k

     ...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection... 
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle

    Boston, MA
    2 days ago
  • $128k - $160k

     ...time. To those who see AI as a driver of progress, come build the future together. The Crown Is Yours As a Senior Site Reliability Engineer, you'll build and scale the critical infrastructure behind every product. In this role, you'll take on complex challenges... 
    Full time
    Immediate start

    DraftKings

    Boston, MA
    12 hours ago
  • $166k - $220k

     ...failure. As such, it is critical that Anduril services are reliable and maintainable. This means that all services &...  ...ground systems & Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build & operate Anduril... 
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Boston, MA
    1 day ago
  • $160k - $200k

    Role Overview We are looking for a Control System Engineer/Site Reliability Engineer (SRE) to integrate and maintain the hardware and software systems that enable QuEra’s quantum controls and software stack. You’ll work closely with software engineers, physicists, hardware... 
    Local area
    Remote work

    QuEra Computing

    Boston, MA
    1 day ago
  • $127k - $249k

    Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that...  ...maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering... 
    Local area
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    3 days ago
  • $130k - $180k

     ...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and...  ...an in-house AI R&D team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to... 
    Temporary work
    Work at office
    Immediate start
    Remote work
    Flexible hours

    GrabJobs

    Boston, MA
    3 days ago
  • $160k - $240k

     ...passionate about building unified IT solutions that simplify the way IT organizations work. We are currently looking for a Senior Site Reliability Engineer to join our SRE team in the Platform Engineering organization and help us scale our products to millions of end-users. We... 
    Permanent employment
    Full time
    Remote work
    Work from home
    Relocation
    Flexible hours

    GrabJobs

    Boston, MA
    4 days ago
  •  ...Information Technology group delivers secure, reliable technology solutions that enable DTCC to...  ...This RoleAs a Senior Application Support Engineer, you will help power DTCC's global...  ...trade processing and settlement.Leveraging Site Reliability Engineering (SRE) principles,... 
    Remote work
    Flexible hours

    DTCC- The Depository Trust & Clearing Corporation

    Boston, MA
    3 days ago
  • $105.79k - $141.05k

     ...shape the future of AI‑ready connectivity, join us today. The Role We are seeking a highly skilled and proactive Lead Site Reliability Engineer (SRE) to join our team, focusing on production support and performance optimization across our portal ecosystem. This role... 
    Temporary work
    Remote work

    Lumen Inc

    Boston, MA
    1 day ago
  • $160k - $225k

     ...Staff Site Reliability Engineer Manifold is the AI platform for life sciences, accelerating life-changing medicines to patients. Our products speed up workflows in areas from target identification and clinical development to market access and precision medicine in the... 

    Manifold

    Cambridge, MA
    5 days ago
  • $130k - $140k

     ...mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionSite Reliability EngineerLocation(s): Waltham, MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to ensuring systems are reliable, scalable, and efficient... 
    Ongoing contract
    Full time
    Temporary work
    Work experience placement

    SS&C Technologies

    Waltham, MA
    5 days ago
  • $166k - $220k

     ...through to the field. When something breaks in a deployed environment, we fix it. About the Role We're looking for a Site Reliability Engineer to join the Imaging team. This is not a product development role, and it isn't a traditional cloud-SRE role either. You... 
    Full time
    Work experience placement
    Immediate start
    Remote work
    Weekend work
    Day shift

    Anduril Industries

    Waltham, MA
    13 hours ago
  • $139k - $257.55k

     ...Individual Contributor The Challenge The Adobe Creative Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning, autonomous AI workflows, and cloud-native infrastructure. Adobe Stock... 
    Temporary work
    Local area
    Remote work
    Relocation

    Adobe

    Waltham, MA
    13 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!