Site Reliability Engineer
$180.5k - $236.91kGrabJobs
Hi, we're Oscar. We're hiring a Senior Software Engineer, Cloud Infrastructure / SRE to join our Engineering team. Oscar is the first health insurance company built around a full stack technology platform and a relentless focus on serving our members. We started Oscar in 2012 to create the kind of health insurance company we would want for ourselves—one that behaves like a doctor in the family. About the role: Our Core Technology teams build and maintain the foundational platform upon which all Oscar engineering is built. We are responsible for architecting a world-class, resilient ecosystem using a modern stack centered on AWS/GCP, Terraform, and Kubernetes with developer focused tooling and CI/CD. Our mission is to provide an automated, self-service infrastructure that empowers our engineering organization to move fast without sacrificing security or stability. You will report into a Staff/Senior Staff Engineer. Work Location: This is a remote position, open to candidates who reside in: San Francisco, CA. You will be fully remote; however, our approach to work may adapt over time. Future models could potentially involve a hybrid presence at the hub office associated with your metro area. #LI-Remote Pay Transparency: The base pay for this role is: $180,504 - $236,911 per year. You are also eligible for employee benefits, participation in Oscar's unlimited vacation program, company equity grants, and annual performance bonuses. Responsibilities: Become the expert on your team's business and technical domains such as DevOps, site reliability, and cloud best practices Lead the planning, execution and release of complex technical projects across multiple teams outside of Core Technology Work with partners, product managers, and designers to solve challenging problems Lead and mentor engineers on the team to improve technology and apply best practices Independently responsible for large or complex technology capabilities (set of components or services) within their team's domain or spanning multiple domains Facilitates, encourages, and enhances cross-team execution and collaboration; knows when cross-team projects are at risk and actively mitigates risk to deliver on time Prolific contributor to the objectives of their functional group, as well as organization-wide projects Drives prioritization of technical roadmap and influences prioritization of product roadmap and process enhancements within their team Actively identifies and reduces failure domains, designs and builds resilient systems, and strives to reduce adverse effects of an outage. Builds software to minimize effort and business impact during maintenance and failures Guides the development of Service-Level Objectives (SLOs) for systems they are responsible for Own medium to large features or infrastructure projects from technical design through completion Compliance with all applicable laws and regulations Other duties as assigned Requirements: 6+ years of professional software engineering experience, working with a variety of technologies, and have increasingly impactful accomplishments Experience as a major contributor cross-pod or cross-company deliverables Experience leading technical contributions, improving the quality of what your teams create, and are excited to build fault-tolerant, and scalable software systems. Demonstrates expertise of the practical application of CS concepts within their team. Sets and enforces the standard for writing stable, correct, and maintainable code Experience mentoring and training more junior engineers Bonus points: Cloud Proficiency: Deep expertise in managing production environments within AWS or GCP at scale. Infrastructure as Code: Advanced experience with Terraform or similar IaC tools to manage complex, multi-account structures. Orchestration & Delivery: Proven track record with Kubernetes and workflows using ArgoCD. SRE Discipline: Strong background in Site Reliability Engineering, including Service Level Objectives (SLOs), error budgets, and incident management. CI/CD & Automation: Experience building robust deployment pipelines via GitHub Actions. Observability: Proficiency with monitoring using tools like Prometheus, Grafana, or similar. Security & Networking: Knowledge of cloud-native security (IAM, VPC peering) and service mesh technologies like Istio. Programming: Understanding of at least one coding language that you are able to use to develop scripts and software Education: B.S. in Computer Science, a related technical field, or equivalent high-level industry experience. This is an authentic Oscar Health job opportunity. Learn more about how you can safeguard yourself from recruitment fraud here . At Oscar, being an Equal Opportunity Employer means more than upholding discrimination-free hiring practices. It means that we cultivate an environment where people can be their most authentic selves and find both belonging and support. We're on a mission to change health care -- an experience made whole by our unique backgrounds and perspectives. Pay Transparency: Final offer amounts, within the base pay set forth above, are determined by factors including your relevant skills, education, and experience. Full-time employees are eligible for benefits including: medical, dental, and vision benefits, 11 paid holidays, paid sick time, paid parental leave, 401(k) plan participation, life and disability insurance, and paid wellness time and reimbursements. Artificial Intelligence (AI): Our AI Guidelines outline the acceptable use of artificial intelligence for candidates and detail how we use AI to support our recruiting efforts. Reasonable Accommodation: Oscar applicants are considered solely based on their qualifications, without regard to applicant’s disability or need for accommodation. Any Oscar applicant who requires reasonable accommodations during the application process should contact the Oscar Benefits Team (View email address on click.appcast.io) to make the need for an accommodation known. California Residents: For information about our collection, use, and disclosure of applicants’ personal information as well as applicants’ rights over their personal information, please see our .
- ...your big ideas, and your desire to team up with some of the best and brightest in technology and entertainment. The RoleThe Site Reliability Engineer (SRE) II is responsible for designing, implementing, and maintaining scalable and reliable systems and applications. Focus...SuggestedFull timeLocal areaWorldwideFlexible hours
- ...A leading livestream shopping platform is seeking a Senior Software Engineer for the Logistics Platform team. This role focuses on improving logistical data systems, enhancing buyer and seller experience, and fostering collaboration across departments. Ideal candidates...SuggestedRemote work
- ...The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational... ...fleet, alongside the critical components that ensure cluster reliability and security (e.g., CoreDNS, cert-manager, and Gatekeeper). As...SuggestedWork at officeLocal areaRemote workWorldwide
- ...What you will do: Partner with a team of high-performing engineers and developers who are focused on delivering best in class software... ...our shift to a SecDevOps culture, solving for security, reliability, cost-effectiveness, and observability Building Zero trust...SuggestedFull timeContract workLocal areaFlexible hoursShift work
$32 - $35 per hour
...and assignment.) Key Responsibilities: In this role, you will help ensure the reliability, performance, and stability of key restaurant-facing platforms by working closely with engineering and infrastructure teams. You will use observability tools such as DataDog, Grafana...SuggestedContract workLocal areaImmediate start- ...foundation of success and bringing it to the digital space - ready to join us? What’s the position? We are looking for a Senior Site Reliability Engineer who combines deep infrastructure expertise with a forward-thinking approach to AI-driven operations. In this role you will...Remote workFlexible hoursNight shift
$150k - $180k
...Senior Cloud Reliability EngineerIrvine, California, United States; Los Angeles, California, United StatesWhat You'll DoThe Senior Cloud Reliability Engineer will be responsible for writing and integrating various open source and closed sources tools. The ideal candidate...Work experience placementLocal area- ...SRE Support Engineer While this position is not currently open, we are interviewing strong candidates for upcoming opportunities on this team. Location: Remote | Time Zone: (iNDIA)(8AM–5PM IST) Domain: Compute(Linux Fundamentals, Linux Networking, Kubernetes, Docker)...Remote work
- ...Senior Site Reliability Engineer (SRE) Our client is a global technology consulting and digital solutions company that enables enterprises across industries to reimagine business models, accelerate innovation, and maximize growth by harnessing digital technologies....Local area
$140k - $180k
...fundamentally different class of spacecraft. Engineered to survive the harshest radiation... ...create highly available, deployable, and reliable products Reduce operational toil through... ...experience in Software Engineering, Site Reliability Engineering or DevOps ~ Deep...Permanent employmentShift work$164k - $270k
...for the 21st century and beyond.The Role What You’ll DoOwn the reliability of our robotics systems, from PLCs through ROS2/middleware to... ...remediation.Partner with controls, robotics, and platform engineering teams to bake reliability in early. Review designs, develop SLOs...Permanent employmentFull timeLocal areaFlexible hours$164k - $270k
Hadrian - Manufacturing the FutureHadrian is building autonomous factories that help aerospace and defense companies manufacture rockets, satellites, jets, and ships up to 10x faster and up to 2x cheaper. By combining advanced software, robotics, and full-stack manufacturing...Permanent employmentFull timeLocal areaRemote workFlexible hours- ...Role: Site Reliability Engineering (SRE) Location: Los Angeles, CA Remote position Fulltime position JD Site Reliability Engineer Experience in Cloud platforms (AWS, Azure, Google Cloud) and hybrid environments. Proficiency...Full timeRemote work
$150k - $200k
...our CEO's funding announcement: . The Reliability team owns the availability, performance,... ...enforcing reliability standards across engineering Designing incident response processes and... ...strong ownership of production systems. As a Site Reliability Engineer on the Reliability...Remote workVisa sponsorshipWork visaFlexible hours- ...scale, we invite you to bring your talents to Zscaler to help shape the future of cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or Bellevue WA) to join our Zero Trust Exchange team. This is a remote role based in San...InternshipWork at officeLocal areaRemote workWorldwide
$100k - $200k
...DevOps / Site Reliability Engineer Los Angeles, CA General Matter is enriching uranium in America. Our mission is to restore our country's ability to make nuclear fuel. Our fuel will help power AI, manufacturing, and other critical industries. It will power our next...Full timeWeekend work- ...worldwide, Disney Cruise Line, Aulani, a Disney Resort & Spa, and Disney Vacation Club. This role sits in the Commerce Site Reliability Engineering (SRE) specifically supporting Ecommerce, Consumer Products and Licensing and Publishing organization within Technology &...Work experience placementWorldwide
$197k - $291k
...troubleshooting distributed systems. Preferred qualifications Master's degree in Computer Science or Engineering. 1 year of people management experience. About The Job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale,...Full time- ..., and thrive! KēSTA I.T. is actively seeking a Principal Engineer for an immediate full-time opportunity with our industry creating... ...An innovative technology company is seeking experienced Site Reliability Engineers to take ownership of building reliable, scalable platforms...Permanent employmentFull timeTemporary workImmediate start
$120k - $180k
...organizations that test and validate complex systems—think drones, rocket engines, satellites, and nuclear reactors. Supported by leading... ...to roll : Frequently traveling to spend time with end-users on-site (e.g. rocket test stands, spacecraft clean rooms, automated...Full timeTemporary workWork experience placement- ...Job Description Job Description Forhyre is looking for engineers who can bring unique perspectives and innovative ideas to all areas... ...evangelize cloud best practices while building a culture of reliability and observability Engage in and improve the end to end lifecycle...
- ...defense programs. Our platform gives hardware engineering teams a single place to ingest data,... ..., frequent travel to end-user sites, and requires U.S. TS Clearance eligibility... ...environments. Reduce complexity and improve reliability as we grow.Drive priorities: Identify the...Permanent employmentWork at office
$141.9k - $190.3k
...and transcends generations. We’re looking for passionate engineers who love learning new technologies at a rapid pace. You should... ..., and clear observability ~ Maintain and improve the reliability of services and infrastructure ~ Troubleshoot and resolve...Work experience placement- A leading technology company is seeking an Engineering Manager to lead a team focused on Site Reliability Engineering. The role demands a strong background in software development, data structures, and team management. Responsible for the uptime and performance of critical...
$120k - $150k
...love for you to join us on our mission of providing humankind access to the galaxy beyond our planet. About the RoleAs a Software Engineer, Business Systems you will have the opportunity to architect and manage the Apex “Operating System” platform from the ground up. Reporting...Full timeWork at office- ...next-generation defense programs. Our platform gives hardware engineering teams a single place to ingest data, analyze performance,... ...CI/CD pipeline performance, release automation, and release reliability.Own developer infrastructure that enables engineering velocity...Permanent employmentWork at officeLocal area
$128.52k - $204.09k
...panels, 3D printed components, body hardware, etc. Develop engineering concepts and designs, drive 3D packaging, generate manufacturing... ...NX and Teamcenter. Work Environment This position is on-site at our offices in Torrance, CA. Some work will be done in our...Temporary work$229.2k - $319.5k
...Principal Software Engineer - Developer Connections, Game Release Job Id: REQ-0010111 Riot engineers bring deep knowledge of specific... ...across Game Studios, R&D, and Central Tech to help teams reliably and efficiently ship and operate games. Our mission is to provide...Temporary workLocal areaImmediate startFlexible hours$155.9k - $233.9k
...excellence and creativity.SIE Studio IT is seeking a Lead Systems Engineer to lead the design, operation, and evolution of our production... ...by Systems Administrators to ensure quality, consistency, and reliability.Act as the technical lead during infrastructure incidents,...Shift workAfternoon shift- ...tier investors and has over $13M in government funding. About the Role Antares is seeking a Research & Development Software Engineer to build the software systems that enable fast, rigorous experimental engineering across our R&D organization. This role supports...Permanent employmentFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Site Reliability Engineer. Be the first to apply!
- on-site clinical research associate (traveling/remote) Glendale, CA
- junior website developer Glendale, CA
- site leader Glendale, CA
- historic site Glendale, CA
- construction site safety Glendale, CA
- official site Glendale, CA
- site services specialist Glendale, CA
- site safety Glendale, CA
- IT site lead Glendale, CA
- site reliability engineer remote




