Senior Site Reliability Engineer
$175k - $250kThe Recruiting Guy
1 day ago Be among the first 25 applicants This range is provided by The Recruiting Guy. Your actual pay will be based on your skills and experience — talk with your recruiter to learn more. Base pay range $175,000.00/yr - $250,000.00/yr Job Title: Senior Cloud Infrastructure Engineer Location: San Francisco, CA. Remote unavailable. Modality: On-Site only. Must live within commuting distance of San Francisco or be willing to relocate. Relocation Assistance: No Employment Type: Salaried W2 Full-Time. Salary Range: $175,000 - $250,000 About The Company We represent a pioneering open source technology company in San Francisco that is transforming the way creators interact with generative AI. They are the team behind a powerful, node based visual interface that gives artists, developers, and innovators the ability to design, control, and customize AI workflows with complete flexibility. Their platform allows users to connect modular components, build complex pipelines, and run everything locally with impressive speed and precision. Their mission is to make generative AI open, transparent, and accessible to everyone. Built around community collaboration and creative empowerment, their tools help users experiment freely and bring their ideas to life. Whether it is visual storytelling, image generation, or advanced machine learning, their technology gives creators the freedom to explore without limitations. About The Role In this role, you will take the lead on designing, deploying, and maintaining large-scale distributed systems that power AI workloads. The ideal candidate is deeply technical, self-sufficient, and motivated by solving complex infrastructure challenges. You will work closely with core engineers to shape the company’s long-term infrastructure vision while ensuring scalability, performance, and reliability across environments. What You’ll Do Design, build, and maintain the core infrastructure that powers AI workloads at scale Manage and automate GPU compute clusters using tools such as Python, Kubernetes, Terraform, and Ansible Architect and operate systems for orchestration, observability, distributed storage, and networking Ensure reliability, scalability, and performance across production environments Collaborate closely with core engineers to design infrastructure for new features and systems Contribute to technical strategy and long-term infrastructure vision Drive best practices for infrastructure automation, deployment, and monitoring Requirements 5+ years experience as an Infrastructure Engineer or Site Reliability Engineer building and operating large-scale distributed systems Skilled in Python and comfortable working with infrastructure-as-code tools such as Terraform and Ansible Familiar with container orchestration systems such as Kubernetes and related tooling like FluxCD, Prometheus, and Grafana Capable of managing high-performance GPU environments across cloud and bare metal setups Highly adaptable, resourceful, and motivated by building things from the ground up Excited to work in a small, fast-growing team where autonomy and accountability are key Comfortable working on-site in a startup setting where collaboration and speed matter most Bonus Points Experience contributing to or maintaining open-source projects Background working with AI infrastructure, ML pipelines, or GPU orchestration Strong computer science fundamentals and ability to work across different programming languages or frameworks Skills: prometheus,fluxcd,kubernetes,python,ansible,terraform,infrastructure,grafana #J-18808-Ljbffr
- ...and actionable to everyone, everywhere. That everyone now includes AI agents. The Role: You'll be the infrastructure and reliability engineer on the Data Replication team - a full-stack product team running over 3 million sync jobs a week powering thousands of data...SeniorFull timeWork at officeLocal areaFlexible hours
- ...founders with PhDs in AI, Math, and Computer Science - is poised to redefine computing. About the Role We're seeking a Site Reliability Engineer to ensure Hyperbolic's GPU marketplace and AI infrastructure operate with exceptional reliability, performance, and...Senior
- ...About the job Senior Site Reliability Engineer About the Company Stellar is a decentralized, public blockchain that gives developers the tools to create experiences that are more like cash than crypto. The network is faster, cheaper, and far more energy-efficient...Senior
$210k - $240k
...Join to apply for the Senior Site Reliability Engineer role at Alembic Technologies This range is provided by Alembic Technologies. Your actual pay will be based on your skills and experience — talk with your recruiter to learn more. Base pay range $210,000.00/yr - $2...SeniorFull time$160k - $250k
...public clouds when the right fit. As we continue to commercialize our machine learning models, we also need to grow our DevOps and Site Reliability team to maintain the reliability of our enterprise SaaS offering for our customers. Our ideal candidate is someone who is able...Senior- ...Engineering Hiring Sprint We're growing our engineering team and are accelerating hiring through a focused Engineering Hiring Sprint... ...: Platform Engineers Database Engineers Site Reliability Engineers Extensibility API Engineers AI Agents Engineers...SeniorWork at officeLocal areaFlexible hours
- ...Udaip Cloud-Based Data And Ai Platform Engineer At U.S. Bank, we're on a journey to do our best. Helping the customers and businesses we serve to make better and smarter financial decisions and enabling the communities we support to grow and succeed. We believe it...SeniorTemporary workWork experience placement
$210.8k - $272.8k
About Thumbtack Thumbtack helps millions of people confidently care for their homes. About the Site Reliability Engineering Team The Site Reliability Engineering team focuses on creating and maintaining a reliable, secure, and scalable platform vital for a seamless user...SeniorLocal area$232k - $319k
...to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and... ...with self-service Accelerate the velocity of SRE and product engineering by developing robust platforms, powerful tooling, and...SeniorPermanent employmentLocal areaWorldwideFlexible hours$174.92k - $209.91k
...same: to make access to data as simple and reliable as electricity. With Fivetran, customer... ..., canonical and ready to query, with no engineering or maintenance required. We're proud... ...integrate our teams, systems, and career sites. About the Role Fivetran is building...SeniorFull timeWork at officeRemote work$300k
...thousands of H100s, H200s, and B200s, ready for experimentation, full-scale model training, or inference. As a Platform Engineer/Senior Site Reliability Engineer, you’ll own the reliability, performance, and automation of this GPU-powered infrastructure, ensuring...SeniorPermanent employment- ...Job Description Job Description Senior Site Reliability Engineer (Payments Infrastructure) Kody is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, scalability, and operational excellence of our global payment platform. You will...Senior
- ...and 7Wire Ventures. About the Role We are looking for a Senior Software Engineer — an individual contributor who owns a product surface end-... ...governed by HIPAA, PCI-DSS, and SOC 2. Responsibilities Deliver Reliable Craft on Your Surface Deliver reliable, high-quality craft...SeniorFull time
$250k
...across Europe, while now significantly expanding its footprint in the United States. The company is looking for a Senior / Staff Site Reliability Engineer to support and scale large-scale HPC and cloud environments powering GPU-intensive workloads. The role involves...SeniorPermanent employmentRemote work- ...Flexible health coverage and 401K matching. The Role This is a senior individual contributor role for someone who can take a... ...a single feature, and help raise the bar for the rest of the engineering team. We’re looking for someone who can see how the pieces fit...SeniorFull timeFlexible hours
$130k - $196.5k
...from the ground up and migrate existing complex use cases into that system. Work with a team of supportive and passionate software engineers. Architect and implement systems that materialize our platform vision. Provide operational support for our production systems...SeniorFull timeWork from homeFlexible hoursNight shift$200k - $250k
Senior Software Engineer Engineering Prolific Prolific is not just another player in the AI space - we are the architects of the human data... ...decisions that balance scrappy startup execution with scalable, reliable engineering, as Prolific revolutionizes research for the AI...SeniorFull timeWork at officeRemote work2 days per week1 day per week$165k - $247k
...identity resolution, data import and export connectors, privacy and compliance, and the operational reliability of everything in between. As a Senior Software Engineer, you'll take on complex infrastructure challenges: designing for extreme throughput, optimizing for...SeniorFull timeHome officeFlexible hours$124k - $186k
...employees - and aim to leave a positive mark on culture.Job Title:Senior Software Engineer, CMS Platform - Paramount Streaming Department: Technology &... ...Paramount’s most dynamic teams. Opportunities for both on-site and virtual engagement events. Unique opportunities to make...SeniorFull timeLocal area- ...the orchestration layer that turns natural-language intent into reliable execution. Build systems for entity resolution, context... ...product areas. Collaborate closely with Product, Design, Sales Engineering, and Customer Success to turn ambitious ideas into production...SeniorFull timeWork at officeLocal areaFlexible hours
- ...complex, distributed, cloud-native systems. As a Staff Platform Engineer, you will play a critical role in ensuring these systems... ...hands-on engineering and technical leadership role. You will own reliability for major platform domains, design scalable solutions on Kubernetes...Senior
$257k - $302k
...Forbes Cloud 100 2025 List [ and is a Y Combinator 2024 Breakthrough Company [ Checkr is looking for an experienced Senior Staff Software Engineer to facilitate the long-term design of Checkr’s core systems and to lead critical cross-organizational initiatives,...SeniorFull timeWork at officeLocal areaRemote workRelocationFlexible hours3 days per week- ...mission-critical industries, helping partners move more quickly and reliably from algorithm to silicon. Our platform accelerates deployment... .... The Roles We are looking for an experienced software engineer to help us build a new generation of transpilation tools...SeniorFull timeRemote workRelocation packageFlexible hours
- ...Site Reliability Engineer We are looking for a dynamic engineer to join our rapidly growing SRE team. As an SRE, you will report to our VP of Technical Operations and be responsible for operating an extremely high performance and scalable, low latency platform built...Relocation package
$150k - $250k
...Site Reliability Engineer role USC or GC only are considered at this time. San Francisco - Local to Bay area only but role... ..., Rust, Kubernetes, FastAPI, Redis, Postgres, Prisma Seniority 4-8+ years of experience in production or reliability...Work experience placementCasual workLocal areaImmediate startRemote work$152.5k - $219.2k
...global cloud platform. As a team of six engineers distributed across the US, Canada, and the... ...with a strong focus on automation, reliability, and operational excellence. We are one... ...Qualifications ~2+ years of experience in Site Reliability Engineering, DevOps,...Permanent employmentFull timeTemporary workLocal areaWorldwideFlexible hours- ...company valued at $10 billion. We work in‑person five days a week in our new SanFrancisco headquarters. About the Role As a Site Reliability Engineer (SRE) at Mercor, you’ll own production reliability across our most critical systems, partnering directly with...
$100k - $170k
...Site Reliability Engineer Houston; San Francisco; Seattle About Nscale Nscale is the GPU cloud built for AI. We run high-performance, cost-efficient infrastructure for AI-native startups and global enterprises, from bare metal up through the platform services...Flexible hoursShift work$150k
...Site Reliability Engineer San Francisco, CA About The Role We are seeking an experienced Site Reliability Engineer (SRE) with a strong focus on DevSecOps to join our growing engineering team. In this role, you will oversee and maintain the reliability, security...$163.71k - $306k
...their own infrastructure, behind their own controls, with the reliability and operational clarity they would expect from any critical system... ..., Support, and TAMs to trust. Partner with product engineers on infrastructure requirements for new Retool products, especially...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer San Francisco, CA
- site reliability engineer sre San Francisco, CA
- sr hr business partner San Francisco, CA
- senior lighting artist San Francisco, CA
- senior planner San Francisco, CA
- senior hvac project manager San Francisco, CA
- home instead senior care San Francisco, CA
- research associate senior research associate San Francisco, CA
- senior technical product manager San Francisco, CA
- senior cloud infrastructure engineer San Francisco, CA


