Site Reliability Engineer
$139.73k - $160.37kID.me
Company Overview
ID.me is the next-generation digital identity wallet that simplifies how individuals securely prove their identity online. Consumers can verify their identity with ID.me once and seamlessly login across websites without having to create a new login and verify their identity again. Over 152 million users experience streamlined login and identity verification with ID.me at 20 federal agencies, 45 state government agencies, and 70+ healthcare organizations. More than 600+ consumer brands use ID.me to verify communities and user segments to honor service and build more authentic relationships. ID.me’s technology meets the federal standards for consumer authentication set by the Commerce Department and is approved as a NIST 800-63-3 IAL2 / AAL2 credential service provider by the Kantara Initiative. ID.me is committed to “No Identity Left Behind” to enable all people to have a secure digital identity. To learn more, visit
ID.me is a full-time, in-office culture. Unless a specific job description explicitly states otherwise, all roles are on-site five days per week at one of our offices in McLean, VA; Mountain View, CA; New York City, NY; or Tampa, FL. Certain roles — such as field-based sales or other remote-by-design positions — may have different work arrangements as noted in their individual postings.
At ID.me, we embrace the thoughtful use of AI tools in our daily work and there are even occasions where we leverage AI in our hiring process. However, during the interview process, we want to understand your individual skills and experiences. Therefore, we have guidelines on how AI can be appropriately used during your application and interviews which can be found here.
Role Overview We are seeking a Site Reliability Engineer to join our Core Platform Engineering organization. The SRE team builds the automation, observability, and operational foundations that ensure ID.me’s services are reliable, scalable, and secure.
As an SRE, you will play a pivotal role in building the platform and governance processes required to safely scale, deploy, and operate a high volume of machine-generated applications and features. You will design and implement the automated guardrails that maintain our high standards for resilience and security in an AI-accelerated development environment. You’ll focus on infrastructure automation, observability, performance optimization, and incident response, partnering closely with Software Engineering teams to foster a culture of reliability and operational excellence.
This role is based out of our Mountain View, CA or McLean, VA offices and requires full-time in-office attendance, 5 days per week .
Responsibilities
- Build and maintain automated reliability tooling , infrastructure as code, and observability systems that enhance uptime and service performance.
- Develop monitoring, logging, and alerting frameworks (e.g., Prometheus, Grafana, OpenTelemetry) to detect and remediate issues proactively.
- Implement automated architectural reviews and reliability guardrails for agent-developed applications to ensure machine-generated code meets long-term maintainability and performance standards.
- Partner with engineering teams to design and implement scalable, fault-tolerant systems that meet defined SLIs and SLOs.
- Automate repetitive operational tasks and develop self-healing and auto-remediation mechanisms to minimize human intervention.
- Participate in on-call rotations and lead incident response efforts, performing post-incident reviews and driving systemic improvements.
- Improve the deployment and release process using CI/CD pipelines and progressive delivery techniques to ensure stability and safety.
- Champion observability, reliability, and operational readiness reviews as part of the development process.
- Collaborate with Security and Compliance teams to ensure production systems meet FedRAMP, NIST, and internal policy requirements .
- Contribute to documentation, runbooks, and internal tooling to enhance knowledge sharing and operational maturity across teams.
Minimum Qualifications
- Bachelor’s degree in Computer Science, Software Engineering, or a related technical field.
- 3-5 years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering.
- 2+ years of hands-on experience managing and scaling services in cloud environments such as AWS, GCP, or Azure.
- 1+ years proficiency in at least one modern programming language (e.g., Java, Go, Python, Ruby, JavaScript).
Preferred Qualifications
- Strong understanding of containerization and orchestration technologies (Docker, Kubernetes).
- Experience implementing and maintaining CI/CD pipelines and automation frameworks.
- Working knowledge of observability systems —metrics, tracing, logging, and alerting.
- Experience building automated recovery, failover, or chaos-engineering systems to validate reliability.
- Familiarity with event-driven architecture and asynchronous processing systems.
- Knowledge of distributed systems design, load balancing, and performance optimization .
- Exposure to infrastructure-as-code tools (Terraform, Pulumi, Ansible) and GitOps practices.
- Understanding of security and compliance frameworks (FedRAMP, SOC2, or NIST 800-53).
- Strong analytical and troubleshooting skills across the stack—from network to application layer.
- Excellent communication and documentation skills, with a focus on cross-team collaboration and continuous improvement.
- Experience using AI agentic coding assistants and deploying custom AI agents or automated workflows into production environments.
The annual base salary listed does not include a company bonus, incentive for sales roles, equity and benefits which will be determined based on experience, skills, education, relevant training, geographic location and role.
ID.me offers comprehensive medical, dental, vision, health savings account, flexible spending accounts (medical, limited purpose, dependent care, commuter benefit accounts), basic and voluntary life and AD&D insurance, 401(k) with company match, parental leave, ability to participate in unlimited paid time off subject to the terms and conditions of the PTO policy, including 8 company wide holidays, short and long-term disability insurance, accident and critical illness insurance, referral bonus policy, employee assistance program, pet insurance, travel assistant program, wellbeing and childcare discounts, benefit advocates, and a learning and development benefit.
Final offers may vary from the amount listed based on qualifications, professional experiences, skills, education, relevant training, geographic location, and other job related factors.
U.S. Pay Range
$139,725—$160,367 USD
Mountain View, CA Pay Range
$172,528—$201,375 USD
ID.me maintains a work environment free from discrimination, where employees are treated with dignity and respect. All ID.me employees share in the responsibility for fulfilling our commitment to equal employment opportunity. ID.me does not discriminate against any employee or applicant on the basis of age, ancestry, color, family or medical care leave, gender identity or expression, genetic information, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran status, race, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable laws, regulations and ordinances. ID.me adheres to these principles in all aspects of employment, including recruitment, hiring, training, compensation, promotion, benefits, social and recreational programs, and discipline. In addition, ID.me's policy is to provide reasonable accommodation to qualified employees who have protected disabilities to the extent required by applicable laws, regulations and ordinances where a particular employee works. Upon request we will provide you with more information about such accommodations.
Please review our Privacy Policy, including our CCPA policy, at id.me/privacy. If you provide ID.me with any personally identifiable information you confirm that you have read and agree to be bound by the terms and conditions set out in our Privacy Policy.
ID.me participates in E-Verify.
$165k - $280k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARLINK)At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy Starlink, the world’s most...SuggestedPermanent employmentTemporary workWorldwideWeekend work$160k - $240k
...one another millions of times a day - quickly, reliably, and securely. Any time you swipe your credit... ...come make a difference at Fiserv.Job TitleSenior Site Reliability EngineerWhat does a successful Site Reliability Engineer do at Fiserv?You will join our global team in...SuggestedFull time$170k - $200k
We are seeking a talented and motivated Site Reliability Engineer to join our engineering team. You will be responsible for building, maintaining, and troubleshooting cloud service/cluster, infrastructure, and monitoring systems to ensure high availability, performance,...SuggestedFull timeWorldwide$100k - $200k
...OPPO US Research Center is seeking a skilled and proactive Site Reliability Engineer (SRE) to join our team. In this role, you will be responsible for ensuring the stability, scalability, and performance of our application systems. The ideal candidate is passionate about...SuggestedFull time$104.4k - $171k
...The mission of the Cloud Intelligence Group SRE (Site Reliability Engineering) Team is to ensure the stability of production environments, enterprise-grade cloud data reliability, and service continuity for the Cloud Intelligence Group. Our greatest challenge lies in...Suggested$145k - $165k
...: Selflessly collaborate towards our shared purpose. About the role Bolt Graphics is seeking a highly experienced Site Reliability Engineer (SRE) to design, build, and operate highly reliable developer and production systems. This role is mission-critical to maintaining...Work at officeImmediate start- ...keep the world running. Location: 5 on-site days a week in Sunnyvale, CA Headquarters. Our Team's Vision: Our Engineering team is shaping the future of cybersecurity... ...looking for an experienced Senior Site Reliability Engineer (SRE) with a strong background in...Work experience placement
- ...Site Reliability Engineer Onsite- Bay Area, CA Skills Relevant Skills and Experience What You’ll Do (Day-to-Day) Own and manage our cloud infrastructure (GCP or AWS, on-prem). Build, maintain, and optimize Kubernetes clusters (including GPU-backed clusters...
- ...Investigate and resolve performance and reliability issues across application, infrastructure, database, Kubernetes, and Linux layers... ..., plan capacity, improve observability, and collaborate with engineering teams and business stakeholders. Requirements: Requires hands...
- ...hands‑on in designing, building, and operating our cloud platform—and driving the reliability, performance, and security that empower our engineering organization. As a Site Reliability Engineer at DrumWave, you will: Infrastructure as Code & CI/CD: Automate...
$140k - $165k
...'s most complex electronics. We capture digital exhaust and engineering context from assembly lines - images, test logs, BOM data, performance... ...and the best access to that technology to win. As a Site Reliability Engineer, you'll operate, improve, and scale our AWS-based...- ...technologies. Our mission is to double America’s compute capacity without building new data centers. We are seeking a skilled Site Reliability Engineer to join our growing team. The ideal candidate will help ensure the reliability, scalability, and performance of our hybrid...Work at officeWeekend work
$129k - $158k
...Site Reliability Engineer Join Fortinet, a cybersecurity pioneer with over two decades of excellence, as we continue to shape the future of cybersecurity and redefine the intersection of networking and security. At Fortinet, our mission is to safeguard people, devices...Full timeWorldwide- ...Site Reliability Engineer III There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems. As a Site Reliability...Work at office
- ...that keep the world running. Location: 5 On-Site Days a Week in Sunnyvale, CA Headquarters Our Engineering team is driven by a culture that thrives on visionary... ...to-day basis, you will work on enhancing system reliability and scalability of Illumio SaaS products, and...Work experience placementImmediate start
$170k - $230k
...Site Reliability Engineer (SRE) Palo Alto / San Francisco Bay Area About Mithril Mithril is an AI infrastructure platform built to make GPU compute more accessible and affordable for the world's leading enterprises, AI startups, and the AI research community,...Work at officeLocal area1 day per week$150k - $195k
...customers worldwide. Our team is growing, and we are looking for engineers with passion for automation. You will help support the... ...alongside engineering/operations teams to improve the scalability and reliability of internal processes. Participate in an on‑call rotation....Full timeWorldwide$81.5k - $141.3k
...Site Reliability Engineer II Abbott is a global healthcare leader that helps people live more fully at all stages of life. Our portfolio of life-changing technologies spans the spectrum of healthcare, with leading businesses and products in diagnostics, medical devices...Remote work$255.7k - $300k
...designs from peers, providing feedback to ensure best practices in reliability, security, and efficiency.Triage and resolve complex system... ...execution of software development initiatives.Mentor other engineers and contribute to the engineering community through documentation...Full time$174k - $252k
...systems by pushing for changes that improve reliability and velocity.Practice sustainable... ...:Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical... ...degree in Computer Science or Engineering.Site Reliability Engineering (SRE) is what you...$222k - $300.5k
...possible.Job OverviewAbout the TeamIntuit's Infrastructure and Site Reliability organization owns the operational backbone that keeps... ...hundreds of millions of customers. The Fintech Platform Systems Engineering team builds and operates the AWS-based infrastructure, resiliency...WorldwideShift work$165k - $265k
...is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.SR. SITE RELIABILITY ENGINEER (STARSHIELD) At SpaceX we’re leveraging our experience in building rockets and spacecraft to deploy the Starshield constellation...Permanent employmentTemporary workImmediate startWeekend work- ...AI, IBM and Accern. Position Summary We are hiring for a highly experienced Senior Staff SRE Engineer to act as a senior technical authority within our reliability function. This is a deeply hands-on individual contributor role, to build and operate SRE practices...Shift work
$200k - $260k
...Site Reliability Engineering Lead Glean is seeking a Site Reliability Engineering Lead to foster a culture of engineering excellence, drive technical strategy, and develop a high-performing, collaborative team. Your role is pivotal in ensuring our services meet stringent...Work at officeHome office$150k - $180k
..., environmental, and innovation outcomes. Role Verrus is looking for candidates to serve as software-focused Senior Site Reliability Engineer at Verrus. This is a full‑time position based out of the Mountain View, CA office. Verrus takes a very technology‑forward...Full timeWork at officeLocal areaFlexible hours$169k - $338k
...advanced agentic AI systems that can autonomously handle complex reliability engineering workflows, predictive failure analysis, and self-... ...associates, or business operations across any Walmart system.Site Reliability Engineering Technical Excellence:Design, write and...Full timeTemporary workPart time- ...function to support one of the world’s fastest-growing AI inference services, powered by the Wafer-Scale Engine (WSE). This team will help deliver world-class, ultra-reliable inference infrastructure for leading model builders such as OpenAI and other frontier labs.As a...Shift work
- ...prioritize a diverse F5 community where each individual can thrive.Role SummaryWe are seeking a proactive and detail-oriented Site Reliability Engineer II (SRE II) to join our 24/7 Operations team in a hybrid capacity. In this role, you will provide round-the-clock, eyes-on...Full timeLocal areaImmediate startShift workNight shiftAfternoon shiftWeekday work
$148k - $235.75k
...see how you can make a lasting impact on the world.Join our team of innovative engineers who are building an AI Data Center AIOps platform that turns raw, high-volume telemetry into reliable, job-centric insights and automation for GPU fleets. We’re hiring a DevOps Engineer...Full time$152k - $241.5k
...infrastructure platforms for automated host lifecycle management, fleet reliability/auto-healing, E2E observability or data-driven operations (... ...languages such as Python, Go, Perl, or Ruby.Mentored other engineers and influenced technical direction through design reviews,...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- official site Mountain View, CA
- construction site safety Mountain View, CA
- site safety Mountain View, CA
- junior website developer Mountain View, CA
- on-site clinical research associate (traveling/remote) Mountain View, CA
- site reliability engineering manager
- site reliability engineer sre
- site reliability engineer
- junior site reliability engineer
- site reliability engineer remote


