Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Site Reliability Engineer

GrabJobs

About Teleport Don't wait for the future of infrastructure. Be part of it. Teleport is the AI Infrastructure Identity Company. We're solving one of the hardest problems in security: giving every human, machine, workload, and AI agent a cryptographically secured identity, improving engineering velocity while maintaining security. We make trusted computing simple. This gives you the freedom, power, and autonomy to build and innovate with confidence. Remote-first and globally distributed, we work with companies like Nasdaq, IBM, and Elastic to secure infrastructure for an AI world. About the Role Teleport Cloud takes our traditionally open-source and enterprise access plane and provides a SaaS option for our customers to adopt. As such, our team is building our production and software as a service infrastructure from scratch. We tackle the hard problems that allow our customers to trust us for secure and reliable access to their infrastructure. Excellent security is table stakes; a security breach can compromise our customers' infrastructure. We must also balance security with maintaining productivity and building a compelling product offering for our customers. Most of the code you will write will be written in Go. We strongly encourage you to explore our GitHub Repo to get a taste of what we are building. Important information We require to be in-person in our Oakland, CA, office for the onboarding week. We conduct background checks. What You'll do Re-engineer the core teleport product to scale globally and optimize routing latency for teams distributed around the world Re-write portions of the core Teleport product to enable our goals for the cloud product Build out our monitoring and observability stack to alert us to production issues and minimize false positives so we can all get a good sleep at night Work on automation to tackle and eliminate the highest toil activities Execute on traditional operation challenges, such as patching, scaling, backup and restore, disaster recovery, and more Investigate the outages and incidents our customers experience with our product Participate in the on-call rotation to ensure 24/7/365 system uptime. What We're Looking For Willingness to collaboratively work with Teleports’ engineers on coding challenge in Go as part of the interview process. Strong experience in Linux systems, networking, containers, and troubleshooting. Have solid Go and Kubernetes development experience. Strong experience developing scripts, automation, or lightweight programs, submitting patches to the product codebase, or building tooling that incorporates AI agents into operational workflows. AWS Cloud experience is preferred, GCP experience is acceptable. Systems Observability tools: Prometheus, Grafana, Loki etc. Operate and support the observability platform to maintain visibility and reliability. Experience operate in a team where sound security choices are critical, and where reasoning about correctness and system invariants (e.g. formal or property-based methods) is valued. Intellectual curiosity and a willingness to master new technologies. Transparency, honesty, and a no-ego mindset. Excellent communication skills. How We Hire Our process is designed to be straightforward and respectful of your time. We skip performative rounds and focus on what matters: understanding how you think and what you can do. For this role, we use a take-home challenge that mirrors real work at Teleport — on your time, your way. You'll have support from the team throughout. Zoom meeting with a Teleport recruiter. You’ll learn about the company, our products, compensation philosophy, interview process, and key requirements. Zoom meeting with the Hiring Manager or a Lead Engineer. They will walk you through the coding challenge and answer your questions. Coding challenge collaboration. The day after your meeting with the Hiring Manager or Lead Engineer, you join a Slack channel and complete a coding challenge in Go using GitHub. The challenge usually takes about 2 weeks and ends with a Zoom meeting with the Hiring Manager and a member of the interview team to review your solution. If your challenge solution meets our bar, you'll receive an offer to join Teleport. Why Teleport You're joining a company where the problem is real, the team is small, and your work shows up directly in the product. We're not a big company, you won't get lost in a crowd. You'll have the freedom and autonomy to do what you're great at, alongside teammates who care about doing it right and want to see you succeed. The work is collaborative, there's real room to grow, and the mission is one worth showing up for. Remote-first and globally distributed, we're genuinely passionate about what we're building. The Benefits At Teleport, we believe your career is more important than a list of perks. That's why we focus on making your day-to-day the best it can be — giving you the autonomy, access, and support to do the best work of your career. - Extensive health coverage - Annual expense budget - Rest and recovery policies that maximize your ability to recharge - Investment in your future with retirement savings plans - Professional development opportunities Teleport is an equal opportunity employer and does not discriminate against any employee or applicant on the basis of age, color, disability, gender, national origin, race, religion, sexual orientation, veteran status, or any classifications protected by federal, state, or local law. Candidate Privacy Notice: For information about our collection and processing of job applicant personal data for this position, please see our Job Applicant Privacy Policy and Notice of Collection at goteleport.com/legal/apply/job-applicant/

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Site Reliability Engineer in Houston, TX vacancy
  •  ...and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.As a Site Reliability Engineer III at JPMorgan Chase within the Corporate Technology, Risk Technology team, you will solve complex and broad business problems... 
    Suggested

    JP Morgan Chase

    Houston, TX
    5 days ago
  • As a Site Reliability Engineer, you will be responsible for: Operational Excellence & Incident Management- Maintain and monitor production systems for availability, latency, and performance.- Lead incident response efforts, including communication, resolution, and postmortem... 
    Suggested
    Permanent employment

    National Oilwell Varco

    Houston, TX
    2 days ago
  • The NexTier Technology team is looking for a Site Reliability Engineer (SRE) to help build, scale, and maintain highly reliable systems on Google Cloud Platform (GCP). This role blends software engineering with infrastructure expertise to ensure our services are performant... 
    Suggested

    Patterson-UTI

    Houston, TX
    2 days ago
  •  ...Hynes & Khater is seeking a Cloud Support Engineer Lead to own the reliability, observability, and operational health of Azure-based applications. This is a hands-on leadership role; you will diagnose hard problems and define processes that keep systems running reliably... 
    Suggested

    Energy Jobline ZR

    Houston, TX
    2 days ago
  •  ...and AI agent a cryptographically secured identity, improving engineering velocity while maintaining security. We make trusted computing...  ...problems that allow our customers to trust us for secure and reliable access to their infrastructure. Excellent security is table stakes... 
    Suggested
    Work at office
    Local area
    Remote work
    Sleeping nights

    GrabJobs

    Houston, TX
    3 days ago
  • $100k - $170k

     ...Site Reliability Engineer Houston; San Francisco; Seattle About Nscale Nscale is the GPU cloud built for AI. We run high-performance, cost-efficient infrastructure for AI-native startups and global enterprises, from bare metal up through the platform services... 
    Flexible hours
    Shift work

    Nscale

    Houston, TX
    1 day ago
  • $1,500 per month

     ...game worlds they inhabit. Our approach is centered around World Engine, our state-of-the-art onchain game server framework. World...  ...architecture to keep our platform secure. Own delivery, scalability, and reliability of our backend infrastructure. Advise and collaborate with the... 
    Full time
    Flexible hours

    GrabJobs

    Houston, TX
    3 days ago
  •  ...Site Reliability Engineer III There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability... 

    Chase

    Houston, TX
    5 days ago
  • $120k - $175k

     ...of sports fandom. Ready to reimagine the DFS industry together? We are seeking a highly skilled and experienced Senior Site Reliability Engineer to join our team. We are passionate about delivering cutting-edge solutions and pushing the boundaries of what's possible.... 
    Full time
    Remote work
    Work visa
    Flexible hours

    GrabJobs

    Houston, TX
    1 day ago
  • $104.9k - $174.7k

     ...Management. You can learn more about LexisNexis Risk at the link below, About the Role: We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and improve the reliability of our production systems. This is not a purely advisory... 
    Work at office
    Local area
    Remote work
    Flexible hours

    RELX

    Houston, TX
    1 day ago
  • $74.1k - $148.3k

     ...Site Reliability Engineer Solve complex problems related to infrastructure cloud services and build automation to prevent problem recurrence. Design, write, and deploy software to improve the availability, scalability, and efficiency of Oracle products and services.... 
    Temporary work
    Immediate start
    Flexible hours

    Oracle

    Houston, TX
    1 day ago
  •  ...Senior Site Reliability Engineer The Senior Site Reliability Engineer is responsible for improving the reliability, availability, scalability, and operational excellence of our critical infrastructure platforms and services. This role partners closely with Engineering... 
    Work at office
    Local area

    Castleton Commodities International

    Houston, TX
    3 days ago
  • As a Lead Site Reliability Engineer at JPMorgan Chase within the Corporate Know Your Customer (KYC) Technology group, you hold a leadership role in your team, demonstrate strong knowledge across multiple technical domains, and advise others on the technical and business... 

    JP Morgan Chase

    Houston, TX
    2 days ago
  • As an Entry-Level DevOps Site Reliability Engineer, you will join a team responsible for continuous improvement and support of customer facing products. Responsibilities will include collecting system requirements; improving existing tools and processes through scripting... 
    Work from home
    2 days per week

    Reynolds & Reynolds

    Houston, TX
    2 days ago
  •  ...ENGINEERLocation: HOUSTON, TXFLSA Class: EXEMPTResponsible to: Directo of Software EngineeringPosition Summary: DevOps / Site Reliability Engineer to implement and evolve the infrastructure, deployment pipelines, and reliability posture of our systems. You'll work closely... 
    Full time
    Local area

    VoltaGrid

    Houston, TX
    4 days ago
  • $141.8k - $195k

     ...their best work, grow fast, and bring their full selves to the herd. Why You’ll Love This Role Cribl Inc is seeking a Senior Site Reliability Engineer to join our mission where you will unlock the value of all observability data, as we expand our team in the U.S. Cribl... 
    Temporary work
    Remote work

    GrabJobs

    Houston, TX
    1 day ago
  •  ...JOB DESCRIPTION As a Site Reliability Engineer, you will be responsible for: Operational Excellence & Incident Management - Maintain and monitor production systems for availability, latency, and performance. - Lead incident response efforts, including communication... 
    Permanent employment
    Full time

    NOV

    Houston, TX
    15 days ago
  •  ...Please extend your support for this role. Local candidate will get 1st preference. Job Title: SRE Engineer Location: Houston, TX and Jersey City, NJ - 3 Days Onsite Role FTE role with Mphasis Client: Mphasis H1B transfer will work... 
    Work experience placement
    H1b
    Local area

    Texas State Library and Archives Commision

    Houston, TX
    3 days ago
  • $102.97k - $131.69k

     ...compliance with regulatory requirements and ATS policies and procedures. Partners with internal/external customer for engineered solutions to improve reliability and throughput. Identifies opportunities for Capital Expenditures for equipment replacement with supervision (... 
    Full time
    Work at office

    Advanced Technology Services

    Houston, TX
    4 days ago
  • $76k - $155.7k

     ...RegularPercentage of Travel Required: Up to 10%Type of Travel: Continental US* * *The Opportunity:CACI is seeking Software Systems Engineers to support the Artemis Next Generation Space Suit program at NASA Johnson Space Center. This position contributes to systems... 
    Permanent employment
    Contract work
    For contractors
    Work experience placement
    Immediate start
    Flexible hours

    CACI International

    Houston, TX
    2 days ago
  • $113k - $141.53k

     ...leader in global energy. Senior Solutions Engineer - Systems Integration serves as a...  ...functionally in the field, ensuring safe, reliable, and performant operation across diverse...  ...Willingness to travel to factories and project sites (25%).Preferred QualificationsMaster’s degree... 
    Full time
    For contractors
    Local area
    Worldwide
    Flexible hours

    AES

    Houston, TX
    6 days ago
  • Reliability EngineerHouston, TXThe actual location of this job is in Houston, TX, US. Relocation...  ...families, if needed.This is a fully site‑based role. Working together in person supports...  ...environmentOpportunities to grow your engineering career in a global... 
    Full time
    Relocation package

    Lonza

    Houston, TX
    6 days ago
  •  ...infrastructure challenges. Job DetailsViridien is seeking a Platform Engineer - Infrastructure & Cloud Systems to design, build, and improve...  ...observability tooling. This role focuses on building scalable, reliable systems and ensuring strong integration between infrastructure... 
    Full time
    Relocation
    Flexible hours

    CGG

    Houston, TX
    2 days ago
  •  ...data platforms that support high-volume, data-intensive workflows.The team works across backend engineering, infrastructure, and data systems, collaborating to deliver reliable, high-performance services in a modern cloud-native environment.Key Responsibilities-Backend... 
    Full time
    Flexible hours

    CGG

    Houston, TX
    6 days ago
  •  ...accelerate autonomy development. We are seeking a software engineer with strong C++ expertise and a passion for building scalable simulation...  ...in architecture and technical design discussions Build reliable, maintainable, and well-tested systems Contribute to code... 
    Full time

    Bot Auto

    Houston, TX
    1 day ago
  •  ...Role: Release Engineer Type: Contract Location: Houston, TX(5 days onsite) Release Engineer with CI/CD pipelines...  ...systems architects infrastructure and security teams to deliver reliable and scalable cloud solutions. Key Responsibilities... 
    Contract work
    Shift work

    VBeyond

    Houston, TX
    1 day ago
  • Position Title: Senior Software Engineer - Platform Location: Houston, TX onsiteFLSA Class: ExemptReports To: Manager of Software EngineeringPosition Summary:VoltaGrid is seeking a Senior Technical Solutions Engineer to join our Platform Team, responsible for designing... 
    Full time
    Local area

    VoltaGrid

    Houston, TX
    1 day ago
  •  ...mission-driven team is seeking a bold and dynamic Software Safety Engineer who is fueled by high ownership, execution horsepower, growth...  ...that will interface with design engineers, safety engineers, reliability engineers and representatives from other subsystems to plan,... 
    Permanent employment
    Work at office
    Weekend work
    Afternoon shift

    Axiom Space

    Houston, TX
    16 days ago
  •  ...experience. Build software that matters alongside experienced engineers across software, firmware, and hardware domains. Grow into broader...  ...scenarios to uncover edge cases, performance limits, and reliability opportunitiesDebug and resolve challenging issues that span multiple... 
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start
    2 days per week

    Hewlett Packard Enterprise

    Houston, TX
    5 days ago
  •  ...General Description We are seeking a Senior Software Engineer with strong expertise in Functional Programming and F# to join our...  ...solutions, implement high-quality services, and drive performance and reliability improvements across our platform. This role requires both... 

    Crane Worldwide Logistics

    Houston, TX
    16 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!