Site Reliability Engineer
Backblaze
About Backblaze: Backblaze is the pioneering object storage leader in the open cloud movement, fueling customer success with cloud storage solutions built purposefully to unlock IT budgets, unburden system administrators, and unleash technology innovators. Founded in 2007 and scaling with under $3 million in outside funding until our traditional Nasdaq IPO in 2021, Backblaze generates over $100M in annual revenue as the leading specialized storage cloud. Today, we manage over three billion gigabytes of production data storage for more than 500,000 customers in 175+ countries, supporting businesses, developers, and IT professionals worldwide.
Position Overview
We are seeking a highly analytical, systems-focused Site Reliability Engineer II (SRE II) to play an essential role in ensuring the ongoing stability, scalability, and extreme reliability of our distributed services and open storage infrastructure. In this high-trust operational seat, your primary focus centers on engineering robust system automation, maintaining deep telemetry observability, and anchoring incident response paths to keep customer-facing environments performing flawlessly. Operating within an agile infrastructure framework, you will partner cross-functionally alongside Engineering, Product, and Operations divisions to embed proactive reliability practices, eliminate operational toil, and optimize resource performance.
Key Responsibilities
- Service Reliability & Operations: Support and safeguard the high availability, durability, and resilience of critical cloud storage services across multi-region production environments.
- Observability & SLO Tracking: Monitor cloud service health metrics utilizing Service Level Indicators (SLIs), Service Level Objectives (SLOs), and systemic error budgets, executing proactive escalations before thresholds are breached.
- Incident Response & On-Call: Participate responsibly in distributed team on-call rotations, live incident response triage, and post-incident blameless root-cause reviews to continually harden system limits.
- Infrastructure Automation: Code and deploy robust automation tools for common operational tasks, systematically eliminating manual engineering intervention and engineering toil.
- Telemetry Framework Engineering: Contribute directly to core monitoring, logging, and distributed alerting frameworks utilizing Prometheus, Grafana, Catchpoint, and ELK stacks .
- Infrastructure as Code & CI/CD: Configure, maintain, and scale infrastructure deployments using Infrastructure as Code (IaC) and configuration management tools including Terraform, Ansible, and Jenkins .
- Capacity & Disaster Recovery: Assist with proactive capacity planning modeling, hardware lifecycle tracking, and routine disaster recovery simulation exercises.
Required Skills & Qualifications
- 2–4 years of verified professional history operating as a Site Reliability Engineer, Systems Engineer, or Cloud Operations Administrator within high-scale software architectures.
- A **Bachelor’s degree in Computer Science, Engineering**, or an equivalent highly technical quantitative field (or equivalent professional experience).
- Solid, foundational **Linux systems administration** experience, including deep low-level system troubleshooting and network diagnostic skills.
- Demonstrated script automation proficiency authored natively in at least one backend language, preferably including Python, Go, or Bash .
- Familiarity with cloud-native runtime environments, container engines, and microservices topologies leveraging **Kubernetes or Docker**.
- Proven understanding of core service reliability concepts, structured root-cause analysis, and ITIL/OSS change and capacity management practices.
- Location Context: 100% remote-first full-time operational framework open to qualified infrastructure engineers permanently based within Bangalore, India .
Preferred Strategic Indicators (Nice to Have)
- Prior systems engineering experience operating within a distributed systems, high-volume SaaS, or specialized cloud service provider environment.
- Familiarity with foundational multi-account administration or resource management across hyper-scaler cloud environments (such as AWS, GCP, or Azure).
- Outstanding written communication mechanics, with a track record of authoring clear operational playbooks, runbooks, and architectural diagrams.
What We Offer
- The exceptional technical canvas to directly optimize, scale, and secure the foundational data pipelines handling billions of gigabytes of storage worldwide.
- Highly competitive compensation metrics structured transparently to match your verified Linux systems depth and infrastructure automation capability.
- Stable remote-first full-time parameters providing profound work-from-home schedule flexibility across Bangalore.
- A corporate workspace that champions diversity, equity, and inclusion at its core, fostering a high-trust engineering culture where individuals can deliver their best work.
- ...businesses, developers, and IT professionals worldwide. Position Overview We are seeking a highly analytical, systems-focused Site Reliability Engineer II (SRE II) to play an essential role in ensuring the ongoing stability, scalability, and extreme reliability of our...SuggestedFull timeRemote workWork from homeWorldwide
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that benefits...SuggestedFull timeWorldwide
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Summary: To support our continued growth and success, we are seeking a Sr. Database Administrator to join our team...SuggestedFull timeWorldwideShift work
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Overview The CTMC (Biz Ops React) team is seeking a Senior BizOps Engineer . As Business Operations Engineers, we...SuggestedFull timeWorldwideShift work
- ...part of an on-call rotation of 8 hours by 7 days a week. Requirements: ~6 -9 years of experience in Software Engineering, focusing on Site Reliability Engineering (SRE) or DevOps principles. ~ Experience provisioning large cloud environments using IAC tools such...SuggestedFull timeLocal areaWorldwideFlexible hours
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Site Reliability Engineer II Overview The Enterprise Monitoring team is looking for a Reliability Engineer II to drive our Application Performance...Full timeWorldwide
- Step into the world of Mrsool where convenience meets innovation! As one of the largest delivery platforms in the Middle East and North Africa (MENA) region, Mrsool has captivated users with its unique and seamless experience, earning it the highest ratings among all major...Full time
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Senior Site Reliability Engineer (SRE) Overview The Open Banking SRE team is responsible for building, operating,...Full timeWorldwide
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Site Reliability Engineer Overview The Mastercard Authentication Program owns how consumer authentication works across both in-store and e-...Full timeWorldwide
- ...vast fleet of dedicated on-demand couriers, ensures fast and reliable delivery no matter how far or remote the location may be.... ...at the tap of a button. We are seeking an experienced Site Reliability Engineer (SRE) to join our team. As an SRE, you will play a...Full timeRemote workNight shift
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Lead Site Reliability Engineer Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital economy that...Full timeWorldwide
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Lead Site Reliability Engineer The SRE team at Mastercard is looking for a Site Reliability Engineer (SRE) who thrives on solving complex problems,...Full timeWorldwideEarly shift
- ...and services that help people, businesses and governments realize their greatest potential. Title and Summary Manager, Site Reliability Engineering Overview: Who is Mastercard? At Mastercard technology, we work to connect and power an inclusive, digital...Full timeWorldwideFlexible hours
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Platform Engineer Who is Mastercard? Mastercard is a global technology company in the payments industry. Our mission is to connect and power...Full timeWorldwideWeekend work
- ...their greatest potential. Title and Summary Senior Platform Engineer ABOUT MASTERCARD Mastercard is a global technology company... ..., and evolve systems by pushing for changes that improve reliability and velocity. • Design and Develop scalable platform solutions...Full timeWorldwide
- Summary The Wikimedia Foundation is seeking a Senior Software Engineer to join the team supporting the Wikidata Platform — the structured... ...-scale, production-grade services while ensuring performance, reliability, and maintainability. Working closely with the technical and...Full time
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Platform Engineer I (DevOps) Who is Mastercard? Mastercard is a global technology company in the payments industry. Our mission is to connect and...Full timeWorldwide
- ...their greatest potential. Title and Summary Senior Platform Engineer Overview We are seeking a Senior Database Administrator... ...platforms. This role plays a critical part in ensuring database reliability, security, automation, and scalability across a complex,...Full timeWorldwide
- The Protocol Integrator position is part of a collaborative IT team that is responsible for publishing JoVE's latest science articles to the web. Main responsibilities include converting MS Word manuscripts into HTML and CSS marked-up pages ready for our website, ...Full timeWork at office
- ...phase of rapid growth at Harrison.ai and as we continue to grow, we have identified the need to find technically astute deployment engineers to join the Asia team to help the planned expansion of our company and our deployment strategies globally as new products are...Full timeRemote workWork from homeWorldwideFlexible hours
- The Protocol Integrator position is part of a collaborative IT team that is responsible for publishing JoVE's latest science articles to the web. Main responsibilities include converting MS Word manuscripts into HTML and CSS marked-up pages ready for our website, performing...Full time
- ...experiences, with the integrations, testing, monitoring, and reliability necessary to deploy voice and chat agents at scale. ElevenCreative... ...doing the best work of their lives. We are researchers, engineers, and operators. IOI medalists and ex-founders. If you want to...Remote jobFull timeImmediate start
- What we’re about At Harrison.ai, we’re redefining what’s possible in healthcare. Through our diagnostic AI solutions, we’re building tools that support clinicians to deliver earlier, more accurate diagnoses and raise the standard of care for millions of patients worldwide...Full timeWorldwide
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Software Engineer Overview Mastercard’s Network & Digital Payments group creates meaningful experiences for consumers while enabling merchants...Full timeWorldwide
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Software Engineer Experience Minimum 8+ years of hands-on experience in UI / Frontend engineering, building scalable, high-performance,...Full timeWorldwide
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Software Engineer-1 Overview • Responsible for the analysis, design, development and delivery of software solutions • Defines requirements for...Full timeContract workWorldwide
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Senior Software Engineer Overview Mastercard is the global technology company behind the worlds fastest payments processing network. We are a vehicle...Full timeWork experience placementWorldwide
- ...products and services that help people, businesses and governments realize their greatest potential. Title and Summary Software Engineer II Job Description Summary Who is Mastercard? Mastercard is a global technology company in the payments industry. Our mission...Full timeContract workWorldwide
- ...experiences, and identity and access management. We're looking for a skilled Front-End Developer to join Ping Identity's Protect Engineering team. You'll work at the intersection of design and engineering, helping our product teams implement polished, scalable...Full timeLocal areaWorldwideFlexible hours
- ...realize their greatest potential. Title and Summary Senior Software Engineer-2 Overview The Network Automation organization builds the software and automation capabilities that improve reliability, resiliency, and operational efficiency across Mastercard’s global...Full timeWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- site reliability engineer India
- on-site clinical research associate (traveling/remote) India
- junior website developer India
- site reliability engineer remote
- site reliability engineer sre
- site reliability engineering manager
- site reliability engineer
- lead site reliability engineer
- junior site reliability engineer
- site activation specialist





