Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

GrabJobs

About Teleport Don't wait for the future of infrastructure. Be part of it. Teleport is the AI Infrastructure Identity Company. We're solving one of the hardest problems in security: giving every human, machine, workload, and AI agent a cryptographically secured identity, improving engineering velocity while maintaining security. We make trusted computing simple. This gives you the freedom, power, and autonomy to build and innovate with confidence. Remote-first and globally distributed, we work with companies like Nasdaq, IBM, and Elastic to secure infrastructure for an AI world. About the Role Teleport Cloud takes our traditionally open-source and enterprise access plane and provides a SaaS option for our customers to adopt. As such, our team is building our production and software as a service infrastructure from scratch. We tackle the hard problems that allow our customers to trust us for secure and reliable access to their infrastructure. Excellent security is table stakes; a security breach can compromise our customers' infrastructure. We must also balance security with maintaining productivity and building a compelling product offering for our customers. Most of the code you will write will be written in Go. We strongly encourage you to explore our GitHub Repo to get a taste of what we are building. Important information We require to be in-person in our Oakland, CA, office for the onboarding week. We conduct background checks. What You'll do Re-engineer the core teleport product to scale globally and optimize routing latency for teams distributed around the world Re-write portions of the core Teleport product to enable our goals for the cloud product Build out our monitoring and observability stack to alert us to production issues and minimize false positives so we can all get a good sleep at night Work on automation to tackle and eliminate the highest toil activities Execute on traditional operation challenges, such as patching, scaling, backup and restore, disaster recovery, and more Investigate the outages and incidents our customers experience with our product Participate in the on-call rotation to ensure 24/7/365 system uptime. What We're Looking For Willingness to collaboratively work with Teleports’ engineers on coding challenge in Go as part of the interview process. Strong experience in Linux systems, networking, containers, and troubleshooting. Have solid Go and Kubernetes development experience. Strong experience developing scripts, automation, or lightweight programs, submitting patches to the product codebase, or building tooling that incorporates AI agents into operational workflows. AWS Cloud experience is preferred, GCP experience is acceptable. Systems Observability tools: Prometheus, Grafana, Loki etc. Operate and support the observability platform to maintain visibility and reliability. Experience operate in a team where sound security choices are critical, and where reasoning about correctness and system invariants (e.g. formal or property-based methods) is valued. Intellectual curiosity and a willingness to master new technologies. Transparency, honesty, and a no-ego mindset. Excellent communication skills. How We Hire Our process is designed to be straightforward and respectful of your time. We skip performative rounds and focus on what matters: understanding how you think and what you can do. For this role, we use a take-home challenge that mirrors real work at Teleport — on your time, your way. You'll have support from the team throughout. Zoom meeting with a Teleport recruiter. You’ll learn about the company, our products, compensation philosophy, interview process, and key requirements. Zoom meeting with the Hiring Manager or a Lead Engineer. They will walk you through the coding challenge and answer your questions. Coding challenge collaboration. The day after your meeting with the Hiring Manager or Lead Engineer, you join a Slack channel and complete a coding challenge in Go using GitHub. The challenge usually takes about 2 weeks and ends with a Zoom meeting with the Hiring Manager and a member of the interview team to review your solution. If your challenge solution meets our bar, you'll receive an offer to join Teleport. Why Teleport You're joining a company where the problem is real, the team is small, and your work shows up directly in the product. We're not a big company, you won't get lost in a crowd. You'll have the freedom and autonomy to do what you're great at, alongside teammates who care about doing it right and want to see you succeed. The work is collaborative, there's real room to grow, and the mission is one worth showing up for. Remote-first and globally distributed, we're genuinely passionate about what we're building. The Benefits At Teleport, we believe your career is more important than a list of perks. That's why we focus on making your day-to-day the best it can be — giving you the autonomy, access, and support to do the best work of your career. - Extensive health coverage - Annual expense budget - Rest and recovery policies that maximize your ability to recharge - Investment in your future with retirement savings plans - Professional development opportunities Teleport is an equal opportunity employer and does not discriminate against any employee or applicant on the basis of age, color, disability, gender, national origin, race, religion, sexual orientation, veteran status, or any classifications protected by federal, state, or local law. Candidate Privacy Notice: For information about our collection and processing of job applicant personal data for this position, please see our Job Applicant Privacy Policy and Notice of Collection at goteleport.com/legal/apply/job-applicant/

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Boston, MA vacancy
  • $160k - $200k

     ...data, ideally using promQLKey Responsibilities:Mentor and evangelize on observability best practices, SLIs/SLOs, and reliability culture across engineering teams. Contributing to and maintaining Tulip's triage & remediation processes as a player / coachPerform incident... 
    Senior
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    Tulip Interface

    Somerville, MA
    1 day ago
  • $127k - $249k

    The TeamPlatform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions...  ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper).... 
    Senior
    Work at office
    Local area
    Remote work
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    3 days ago
  • $93.6k - $170.64k

     ...currently have a career opportunity for a NOC AI-Ops Engineer to join our team located in Boston, MA. This is a hybrid...  ...3 days a week in office.Job Overview:We are seeking a Senior AIOps and Incident/Site Reliability Engineer to lead incident management, operational... 
    Senior
    Work at office
    Local area
    Night shift
    3 days per week

    Perficient

    Boston, MA
    4 days ago
  • $134.25k - $214.8k

     ...matters at a company where you matter.Your ImpactAre you an engineer who gets excited about the challenge of making complex distributed...  ...it.You will be part of the Observability team within Axon's Site Reliability organization — a focused team responsible for Axon's metrics,... 
    Senior
    Work experience placement
    Work at office
    Remote work

    Axon

    Boston, MA
    4 days ago
  • $127k - $249k

    Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that...  ...maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering... 
    Senior
    Local area
    Worldwide
    Flexible hours

    MongoDB

    Boston, MA
    3 days ago
  • $166k - $220k

     ...failure. As such, it is critical that Anduril services are reliable and maintainable. This means that all services &...  ...ground systems & Kubernetes infrastructure.ABOUT THE JOBAs a Site Reliability Engineer on the Observability team, you will build & operate Anduril... 
    Senior
    Full time
    Work experience placement
    Immediate start

    Anduril Industries

    Boston, MA
    1 day ago
  •  ...commuting distance of one of our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team, you will operate the...  ...candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based... 
    Senior
    Full time

    Federal Reserve Bank of Boston

    Boston, MA
    4 days ago
  • $128k - $160k

     ...a time. To those who see AI as a driver of progress, come build the future together. The Crown Is Yours As a Senior Site Reliability Engineer, you'll build and scale the critical infrastructure behind every product. In this role, you'll take on complex challenges... 
    Senior
    Full time
    Immediate start

    DraftKings

    Boston, MA
    13 hours ago
  • $121.4k - $218.6k

     ...will be responsible for ensuring best-in-class uptime and reliability of our AI hardware infrastructure offerings. Partner...  ...and defend them when they are breached. As a Senior Site Reliability Engineer, you will be responsible for: Developing and scaling... 
    Senior
    Work experience placement
    Work at office

    Akamai

    Boston, MA
    4 days ago
  • $140k - $210.9k

     ...position will be primarily on-site with residency commutable to...  ...DevOps backgrounds or software engineering backgrounds (e.g., Java...  ...interest in operating and improving reliability of distributed production...  ...Responsibilities As a Senior Engineer of the SRE / Production... 
    Senior
    Full time
    Temporary work
    Part time
    Work at office
    Shift work

    Federal Reserve Bank

    Boston, MA
    2 days ago
  • $140k - $205k

     ...Senior Technology Site Reliability Engineer Cooley is seeking a Senior Site Reliability Engineer to join the Infrastructure & Development Operationsteam. Position summary: The Senior Technology Site Reliability Engineer("SRE") is responsible for ensuring the reliability... 
    Senior
    Full time
    Temporary work
    Work at office
    Flexible hours
    Weekend work

    Cooley

    Boston, MA
    4 days ago
  • $160k - $200k

     ...Senior Site Reliability Engineer This role is located in Somerville, MA - We are a hybrid work environment and are in the office 3+ days/per week. Tulip, the leader in AI-native frontline operations, is helping companies around the world equip their workforce with... 
    Senior
    Temporary work
    Work at office
    Local area
    Flexible hours
    3 days per week

    Tulip Interfaces

    Somerville, MA
    13 hours ago
  •  ...Information Technology group delivers secure, reliable technology solutions that enable...  ...You Will Have in This RoleAs a Senior Application Support Engineer, you will help power DTCC's global...  ...processing and settlement.Leveraging Site Reliability Engineering (SRE) principles... 
    Senior
    Remote work
    Flexible hours

    DTCC- The Depository Trust & Clearing Corporation

    Boston, MA
    3 days ago
  • $160k - $240k

     ...passionate about building unified IT solutions that simplify the way IT organizations work. We are currently looking for a Senior Site Reliability Engineer to join our SRE team in the Platform Engineering organization and help us scale our products to millions of end-users.... 
    Senior
    Permanent employment
    Full time
    Remote work
    Work from home
    Relocation
    Flexible hours

    GrabJobs

    Boston, MA
    4 days ago
  • $81.1k - $187k

     ...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection... 
    Senior
    Temporary work
    Immediate start
    Flexible hours
    Shift work

    Oracle

    Boston, MA
    2 days ago
  •  ...real change. Constantly grow as you work hard for a mission that matters at a company where you matter.Your ImpactAs a Senior Site Reliability Engineer within the APX SRE organization, you’ll focus on delivering practical, scalable solutions to support the reliability... 
    Senior
    Work at office
    Remote work
    Flexible hours

    Axon

    Boston, MA
    22 hours ago
  • $139k - $257.55k

     ...Individual Contributor The Challenge The Adobe Creative Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning, autonomous AI workflows, and cloud-native infrastructure. Adobe... 
    Senior
    Temporary work
    Local area
    Remote work
    Relocation

    Adobe

    Waltham, MA
    14 hours ago
  • $138.1k - $198.2k

     ...more intuitive with technology that simply works.  The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments...  ...Our customers are all engineers at Cisco. Your Impact As a Site Reliability Engineer, you will be at the epicenter of our engineering... 
    Permanent employment
    Full time
    Temporary work
    Work experience placement
    Local area
    Remote work
    Flexible hours

    CISCO Systems

    Boston, MA
    5 days ago
  • $130k - $150k

     ...technologies is essential for this role. Position OverviewThe Site Reliability Engineer (SRE) helps ensure CRA’s critical business services are...  ...career mentoring and performance coaching from an assigned senior colleague. Additional leadership and collaboration opportunities... 
    Work at office
    Work from home
    3 days per week

    CRA International

    Boston, MA
    5 days ago
  • $160k - $200k

    Role Overview We are looking for a Control System Engineer/Site Reliability Engineer (SRE) to integrate and maintain the hardware and software systems that enable QuEra’s quantum controls and software stack. You’ll work closely with software engineers, physicists, hardware... 
    Local area
    Remote work

    QuEra Computing

    Boston, MA
    1 day ago
  •  ...Site Reliability Engineer Cambridge, MA About Watershed Our vision is to become the leading biocomputing platform. The future of biology is in big data analysis, and we are on a mission to accelerate digital drug discovery with the Watershed platform. Watershed... 

    Watershed Informatics

    Cambridge, MA
    3 days ago
  • $130k - $180k

     ...alongside some of the most experienced and innovative leaders and engineers in the field. Where we work Headquartered in Amsterdam and...  ...an in-house AI R&D team. The role Nebius is looking for a Site Reliability Engineer in Hardware Infrastructure team. You’re welcome to... 
    Temporary work
    Work at office
    Immediate start
    Remote work
    Flexible hours

    GrabJobs

    Boston, MA
    3 days ago
  • $95k - $171k

     .... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts... 
    Permanent employment
    Work experience placement
    Work at office
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Akamai

    Cambridge, MA
    4 days ago
  • $108k - $209k

     ...seeking an experienced, creative, and talented Principal / Senior Software Engineer. The ideal candidate will have a strong background in software...  .... Leverage AWS cloud infrastructure to build scalable, reliable, and efficient applications and AI-powered services. Uphold... 
    Senior

    Seres Therapeutics

    Cambridge, MA
    4 days ago
  •  ...itD is seeking a Site Reliability Engineer to develop and enhance automation solutions that improve the reliability, scalability, and operational efficiency of large-scale cloud infrastructure. The ideal candidate will bring hands-on experience in site reliability engineering... 
    Work experience placement
    Remote work

    GrabJobs

    Boston, MA
    2 days ago
  •  ...scale, we invite you to bring your talents to Zscaler to help shape the future of cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or Bellevue WA) to join our Zero Trust Exchange team. This is a remote role based in San... 
    Internship
    Work at office
    Local area
    Remote work
    Worldwide

    GrabJobs

    Boston, MA
    1 day ago
  • $125k - $165k

     ...Permanent Build a brilliant future with Hiscox Job Title: Senior Software Engineer - Integrations Location: Boston, MA- Hybrid Hiscox is...  ...when it comes to talented people. Hiscox is full of smart, reliable human beings that look out for customers and each other. We... 
    Senior
    Permanent employment
    Full time
    Temporary work
    Work at office

    Hiscox

    Boston, MA
    2 days ago
  •  ...mission-critical industries, helping partners move more quickly and reliably from algorithm to silicon. Our platform accelerates deployment...  .... The Roles We are looking for an experienced software engineer to help us build a new generation of transpilation tools... 
    Senior
    Full time
    Remote work
    Relocation package
    Flexible hours

    Code Metal

    Boston, MA
    1 day ago
  • $130k - $140k

     ...mid-market firms, rely on SS&C for expertise, scale, and technology.Job DescriptionSite Reliability EngineerLocation(s): Waltham, MA | HybridAbout the RoleSr Site Reliability Engineer- Guardian of the products to ensuring systems are reliable, scalable, and efficient... 
    Senior
    Ongoing contract
    Full time
    Temporary work
    Work experience placement

    SS&C Technologies

    Waltham, MA
    5 days ago
  • $150k - $195k

     ...empowers members to perform at a higher level through a deeper understanding of their bodies and daily lives.WHOOP is seeking a Senior Reliability Engineer to lead the charge in ensuring our hardware products deliver a consistent, high-reliability experience for members. In... 
    Senior
    Full time
    Work at office
    Relocation

    WHOOP

    Boston, MA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!