Senior Site Reliability Engineer
Blitzy
Senior Site Reliability Engineer
As a Senior Site Reliability Engineer at Blitzy's Cambridge headquarters, you will be the backbone of our platform's reliability, scalability, and operational excellence. You'll work at the intersection of software engineering and infrastructure, ensuring our AI-powered development platform remains highly available and performant as we scale rapidly. This is a high-impact, hands-on role for an engineer who thrives in a fast-moving environment and takes deep ownership of the systems they build.
What success looks like:
- In 30 days: You have a deep understanding of Blitzy's infrastructure architecture, have identified key reliability risks, and are actively contributing to on-call rotations.
- In 90 days: You have shipped meaningful improvements to observability, incident response workflows, and deployment pipelines that measurably reduce MTTR and increase system uptime.
- In 6 months: You have driven at least one major reliability initiative from inception to production, established SLO/SLA frameworks for critical services, and are a trusted technical voice shaping our infrastructure roadmap.
Areas of ownership:
- Design, build, and operate scalable, fault-tolerant infrastructure across cloud environments (AWS, GCP, or Azure).
- Define and enforce SLOs, SLAs, and error budgets; lead blameless postmortems and drive systemic improvements.
- Build and maintain robust CI/CD pipelines, release automation, and deployment infrastructure.
- Own observability: design and maintain logging, metrics, tracing, and alerting stacks (e.g., Prometheus, Grafana, Datadog, OpenTelemetry).
- Partner closely with software engineering teams to embed reliability practices into the development lifecycle.
- Drive capacity planning, performance benchmarking, and cost optimization across our infrastructure.
- Champion security best practices within the infrastructure and deployment layers.
Required experience:
- 5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering roles.
- Strong proficiency in at least one major cloud platform (AWS preferred); experience with Kubernetes and container orchestration at scale.
- Hands-on experience with infrastructure-as-code tools (Terraform, Pulumi, or equivalent).
- Proven track record designing and maintaining high-availability, distributed systems.
- Deep expertise in observability tooling, incident management, and on-call practices.
- Strong scripting and automation skills (Python, Go, Bash, or similar).
- Excellent communication skills with the ability to collaborate across engineering teams and present technical findings to leadership.
What makes you stand out:
- Experience supporting AI/ML workloads or GPU-accelerated infrastructure.
- Prior experience in a high-growth startup environment where you wore multiple hats.
- Familiarity with eBPF, service mesh technologies (Istio, Linkerd), or advanced networking.
- Contributions to open-source SRE/DevOps tooling or communities.
- Experience building global, multi-region infrastructure with strict latency and availability requirements.
What makes this role different:
You won't be maintaining legacy systems or fighting fires in a sprawling monolith. At Blitzy, you're building reliability into a greenfield AI platform that is redefining how the world creates software. You'll have direct influence over architectural decisions, work side-by-side with world-class engineers, and see the tangible impact of your work as we scale to serve Fortune 500 customers. As a founding member of the Pune SRE team, you'll help shape the culture and technical standards of a team that will grow with the company.
Our culture:
Who we are:
Led by two pioneering co-founders we are one of the fastest growing companies in the U.S., creating our own category of enterprise autonomous software development. We automate thousands of hours of software development for our customers, which includes strong representation within the Fortune 500.
How we work:
We move Blitzy fast: Time is both our company's and our clients' most precious asset. We move quickly and decisively to innovate internally and deliver exceptional software externally.
Championship mindset: We operate like a professional sports team. We win as a team by holding ourselves and each other to high standards, collaborating in-person, and remaining focused on the mission.
Passion for invention: We're pushing the frontier of what's possible, requiring constant innovation and iteration.
We work for the customer: We focus on delivering outsized value to the customers we work with and expanding those relationships into deep, meaningful partnerships.
We believe in being 'everyday athletes'—taking care of ourselves so we can bring our best minds to work. We promote great sleep, movement, and restorative activities for optimal mental performance. It makes for a happier and more productive team.
Blitzy is an equal opportunity employer committed to building a diverse and inclusive team. We believe different perspectives make us stronger.
$146.4k - $263.6k
...Do you want to shape reliability practices for a new AI inference platform? Are you a senior technical leader who drives solutions... ...architecture decisions with product engineering teams, and shape SRE... ...at scale. As a Senior II Site Reliability Engineer, you will...SeniorWork experience placementWork at office$150k - $190k
...Senior Site Reliability Engineer (SRE) This role is located in Somerville, MA - We are a hybrid work environment and are in the office 3+ days per week. Tulip, the leader in AI-native frontline operations, is helping companies around the world equip their workforce...SeniorTemporary workWork at officeFlexible hours3 days per week$81.1k - $187k
...Job Description We are looking for a Site Reliability Engineer 3 to support mission-critical cloud services and production operations. The role focuses on improving service reliability, reducing operational risk, automating repetitive tasks, and driving faster detection...SeniorTemporary workImmediate startFlexible hoursShift work- ...commuting distance of one of our 12 Reserve Bank locations As a Senior Engineer of the SRE / Production Operations team, you will operate the... ...candidate is someone who loves building and maintaining reliable and scalable systems, CI/CD tooling, and automating cloud-based...SeniorFull time
- ...Site Reliability Engineer Cambridge, MA About Watershed Our vision is to become the leading biocomputing platform. The future of biology is in big data analysis, and we are on a mission to accelerate digital drug discovery with the Watershed platform. Watershed...Suggested
$95k - $171k
.... Opportunities exist to focus on GPU infrastructure, Kubernetes, and ensuring reliability for AI workloads within Akamai's serverless inference platform. As an Site Reliability Engineer II, you will be responsible for: Building and maintaining dashboards, alerts...Permanent employmentWork experience placementWork at officeRemote workWork from homeWorldwideFlexible hours$51.9 per hour
...Company: Allegheny Health Network Job Title: Site Reliability Engineering – Clinical & Facility Services General Overview This role ensures the reliability, availability, and performance of critical healthcare IT systems in the Environment of Care (EOC), supporting patients...Local area- ...collaborating with Verisk to connect them with exceptional professionals for this role. Description We are hiring a Senior Software Engineer with deep expertise in AI/ML engineering and data-intensive systems to join our Catastrophic and Risk Solutions team. You...SeniorWork at officeFlexible hours
$160k - $225k
...Staff Site Reliability Engineer Manifold is the AI platform for life sciences, accelerating life-changing medicines to patients. Our products speed up workflows in areas from target identification and clinical development to market access and precision medicine in the...$51.9 per hour
...OVERVIEW: This job is responsible for the reliability, availability, and performance of... ...operational efficiency. This role blends software engineering, clinical engineering, and security... .... Works cross-functionally with AHN site leaders and teams to navigate and to monitor...For contractorsLocal area- ...Evolvesquads is looking for a Senior Software Engineer to join our Product Engineering team in Boston. This hybrid role will empower you to lead the design and delivery of complex features, driving architectural decisions while mentoring other engineers. You will partner...Senior
- ...A leading open-source software company is looking for a Senior Software Engineer to support their Observability Service team. This role involves developing high-quality software features, collaborating with engineers, and optimizing user experience in large-scale distributed...Senior
- Good working experience in at least some of the below technologies: Middleware (i.e. Tomcat, WebSphere, WebLogic) Automation (i.e. UiPath RPA, Microsoft Power, Airflow) ETL (i.e. Glide) Data Management (i.e. Snowflake, Data Bricks) Architectural & Operational Knowledge...SeniorWork experience placement
- ...We’re seeking a future team member for the role of Senior Vice President, Third Party Application Engineering to join our Performance & Compliance Engineering team. This role is located in Boston, MA . This role requires strong experience in Eagle PACE Performance...SeniorWork experience placementWorldwide
$286.2k - $326.7k
A leading financial services firm is seeking a Sr. Distinguished Machine Learning Engineer who will drive technical strategies for personalizing product experiences. This role involves building robust ML systems and collaborating with cross-functional teams on advanced...SeniorRemote work- ...A leading technology firm in Boston seeks a Senior Industry Principal to advise C-suite stakeholders on supply chain transformation. This remote position requires 10-15 years of experience in consulting or industry leadership. The ideal candidate will possess deep expertise...SeniorRemote work
$165k - $185k
...This is a Senior Platform Engineer role focused on designing, building, and operating scalable, high-performance cloud infrastructure in AWS within a fast-paced investment banking environment. Title: Senior Platform Engineer Location: Boston, MA - Hybrid (3 days onsite...Senior- ...As a Senior Platform Engineer , you'll be at the heart of our infrastructure - designing, building, and scaling the core platform that powers... ...product teams to ensure our systems are performant, stable, reliable, and secure. You'll own key pieces of our cloud and...Senior
$150k - $195k
...Senior Reliability Engineer At WHOOP, we're on a mission to unlock human performance and healthspan. WHOOP empowers members to perform at a higher level through a deeper understanding of their bodies and daily lives. WHOOP is seeking a Senior Reliability Engineer...SeniorFull timeWork at officeRelocation- ...You Are You have genuine technical literacy - you can read a code snippet, understand what an API does, and engage credibly with engineering stakeholders without faking it. Proven experience in a Solutions Engineering, Forward Deployed Engineering, Sales Engineering, Technical...Senior
- ...Responsibilities Mentor junior staff and define engineering best practices Leverage past experiences to help lead the team through all phases of software development lifecycle Work effectively with developers, stakeholders and cross-functional teams Hands‑on technical...Senior
- ...team is part of CORE. We work at the intersection of software engineering, machine learning, sensors, and hardware compute platforms to... ...vehicles, we would love to talk with you. What You’ll Be Doing As a senior engineer in the Next-Gen Technologies team, you will help us...SeniorWork at office
- ...The Role As a Senior DevOps Engineer, you will lead the design, automation, and operation of secure, scalable AWS cloud infrastructure... ...evolve our environments, and own CI/CD, observability, and reliability across the full development lifecycle. Working closely...Senior
- ...unable to sponsor at this time. 12+ Month Contract with Possible Extension (DeWinter will hire full-time if desired) The Engineer will help build reliable, scalable systems that automate business processes, integrate data, and deliver timely information to business...SeniorFull timeContract work
$90k - $150k
...‑term strategic work. What skills do I need? 4-6+ years in IT engineering, system administration, and network management. Proficiency in... ...policies, device trust, and user lifecycle automation. Maintain reliable onboarding and offboarding automation for accurate, timely...SeniorTemporary workWork at officeLocal areaFlexible hours3 days per week- ...A leading nonprofit R&D company in Cambridge, MA is seeking a Senior Software Engineer to develop high-performance solutions for resource-constrained environments. The ideal candidate has expertise in embedded software development, a strong background in C/C++, and real...Senior
- ...Senior Director, Principal Gifts About the Company Philanthropic organization supporting Indigenous culture & individuals Industry Non-Profit Organization Management Type Non Profit Founded 2017 Employees 11-50 Categories...Senior
$314.8k - $359.3k
Capital One is seeking a Senior Distinguished Engineer, AI Compute to lead the architectural design and implementation of a scalable machine learning platform. The role involves working on distributed systems, developing solutions using Ray and Spark, and collaborating...SeniorRemote work- ...About the job Senior AI Solutions Engineer Location: Fully Remote (Global) Are you a seasoned AI Engineer with a passion for building robust, client... ...), to create sophisticated front‑end solutions that are reliable, scalable, and perfectly tailored to client needs. You...SeniorRemote workFlexible hours
- ...a class of unmet needs critical to streamline the functioning of a large, increasingly digital economy. Job Description Join our Engineering team and drive innovation that matters! We solve Identity Management problem with Decentralized Ledger Technology (DLT) or “Blockchain...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!
- site reliability engineer Cambridge, MA
- senior device engineer Cambridge, MA
- sr operations manager Cambridge, MA
- senior clinical data manager Cambridge, MA
- senior director clinical operations Cambridge, MA
- senior implementation project manager Cambridge, MA
- senior manager customer operations Cambridge, MA
- senior director epidemiology Cambridge, MA
- senior commercial counsel Cambridge, MA
- senior director diversity & inclusion Cambridge, MA



