Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid)
$140k - $215kCrowdStrike
As a global leader in cybersecurity, CrowdStrike protects the people, processes and technologies that drive modern organizations. Since 2011, our mission hasn’t changed — we’re here to stop breaches, and we’ve redefined modern security with the world’s most advanced AI-native platform. We work on large scale distributed systems, processing almost 3 trillion events per day and this traffic is growing daily. Our customers span all industries, and they count on CrowdStrike to keep their businesses running, their communities safe and their lives moving forward. We're proud to work for a mission-driven company leveraging AI to transform the way we work. CrowdStrikers drive their careers through flexibility and autonomy while also being expected to contribute to a culture of responsible AI adoption, experimentation, and innovation. We use an AI-first mindset as a force multiplier to proactively and continuously accelerate execution, build expertise, uncover insights, and solve complex problems. We’re always looking to add talented CrowdStrikers to the team who have limitless passion, a relentless focus on innovation and a fanatical commitment to our customers, our community and each other. Ready to join a mission that matters? The future of cybersecurity starts with you.About the Role: CrowdStrike Falcon is the industry standard in cloud‑native cybersecurity and threat hunting, processing trillions of events per day. As a Principal SRE, you will operate at the intersection of our Core Platform and Embedded Reliability charters: building the foundational libraries, services, and tooling that every product group depends on, while embedding directly with product engineering teams and their leadership to drive reliability outcomes at scale.While we embrace the SRE moniker, at CrowdStrike it means something far more service‑oriented and engineering‑heavy than traditional operations. This is hands‑on systems engineering — writing production code, re‑architecting critical systems, and eliminating entire classes of failure — not ticket management. It is far and away our most self‑driven and autonomous backend engineering role, with the freedom to move up, down, and laterally across the stack as needed.You'll join a group with bottom‑up visibility and ownership of a fast‑expanding codebase (including AI‑based Detection & Response) and a far‑reaching mandate for service resiliency. Several key feature additions and critical re‑architecture initiatives are underway: deploying core services to additional cloud providers, modularizing into reusable components, improving core libraries and frameworks, maturing observability tooling (tracing, profiling, alerting, SLOs), and automating away manual toil across all of the above. Recent examples of the team's work include introducing adaptive concurrency into our core Kafka library to implement flow control that protects database performance, resolving critical issues in leader election libraries, and building infrastructure‑as‑code tooling that eliminated manual deployment processes.At the Senior Engineer level, your influence is organizational. Product engineers and engineering leaders will come to you for guidance on architectural decisions because you've earned credibility through hands‑on work and delivered results. You will shape architectural choices that affect every feature development team and provide shared architectural components leveraged throughout the Falcon Platform — including Unified Search, Protobuf libraries, and other shared‑tier infrastructure.Why This Role Matters: Our customers depend on us to protect their businesses from sophisticated threats, and reliability isn't optional — it's fundamental to our mission. Your work directly impacts whether organizations around the world can defend themselves against cyberattacks. You'll work on problems that matter, at a scale few companies can match, with the autonomy to make real architectural decisions.Location: This position requires candidates to be based in Midtown Manhattan, NY; Redmond, WA; Sunnyvale, CA; or Austin, TX. It's a hybrid role, with employees typically working 2 to 3 days per week in the office.What You'll Do:Partner with engineering leadership across multiple product groups to define and drive multi‑year reliability roadmaps.Design and implement architectural improvements to services, libraries, and platforms that impact teams across all of CrowdStrike.Develop and maintain services that meet aggressive reliability and scalability demands.Extend and build new libraries for cross‑cutting concerns spanning CrowdStrike's cloud platform, which comprises hundreds of libraries and services.Lead initiatives around reliability, scalability, performance, and cost efficiency in large‑scale distributed systems.Establish foundational observability practices: ensure teams instrument services properly, react to signals effectively, and leverage observability to drive automation such as continuous delivery.Define and implement service‑level objectives and error budgets that drive real decision‑making and prioritization.Lead performance and cost optimization efforts: profiling, bottleneck analysis, capacity planning, and cloud efficiency improvements.Conduct resilience engineering: chaos experiments, failure injection, failure modeling, and designing for graceful degradation.Design and implement automation and infrastructure‑as‑code to improve infrastructure reliability and eliminate manual toil.Provide technical leadership during complex incidents and ensure follow‑through on retrospectives with concrete improvements that eliminate entire classes of failures.Identify opportunities to extract common patterns into shared libraries and tools, or partner with platform teams on improvements benefiting multiple product groups.Continuously re‑evaluate our products to improve architecture, knowledge models, developer and user experience, performance, and stability.Drive strategic technical decisions and influence infrastructure and operational improvements across the organization.Mentor and coach engineers, raising the technical IQ of the team and driving architectural standards across the org.Use and give back to the open source community; evangelize software engineering best practices, especially as they pertain to Go.Brainstorm, define, and build collaboratively with members across multiple teams as an energetic self‑starter who owns and is accountable for deliverables.What You'll Need: 10+ years of experience building and operating distributed systems and service‑oriented backends at scale.5+ years developing microservices for a SaaS product in a modern backend language (Go, Java, Scala, Kotlin, Python, Node.js).Expert‑level proficiency in at least one programming language, with expert‑level Go or demonstrated ability and willingness to reach expert level in Go.Deep understanding of distributed systems: consensus algorithms, replication, consistency models, failure modes, and scalability patterns.Proven experience scaling backend systems — sharding, partitioning, horizontal scaling, capacity planning, and performance optimization are second nature.Deep understanding of multi‑threading, concurrency, and parallel processing.Track record of making impactful architectural decisions at organizational scope and seeing them through to production.Strong systems thinking and the ability to influence without direct authority across organizational boundaries.Thorough command of engineering best practices: appropriate testing paradigms, effective peer code review, and resilient architecture.Ability to thrive in a fast‑paced, test‑driven, collaborative, and iterative environment; strong team‑player orientation.A desire to ship code and a love of seeing your bits run in production.Degree in Computer Science, or commensurate experience in data structures, algorithms, and distributed systems.Proven experience utilizing AI technologies to enhance decision-making, streamline workflows and processes, improve efficiency and drive business outcomes.Bonus Points:Experience driving reliability improvements in organizations with hundreds or thousands of microservices. Deep knowledge of Kubernetes or other large‑scale orchestration systems. Hands‑on experience with AWS, Cassandra, Kafka, Elasticsearch/OpenSearch, or similar large‑scale distributed technologies. Experience with Google Cloud Platform (GCP).Experience with Oracle Cloud Infrastructure (OCI).Experience delivering or operating services across multiple cloud providers, including multi‑cloud abstraction layers, portability, and cloud‑agnostic tooling. Track record of building internal platforms, developer platforms, or tools that other engineers depend on. Experience with infrastructure cost optimization at scale. Background in performance engineering: profiling, optimization, and identifying system bottlenecks. Experience with chaos engineering or resilience testing practices. History of establishing SLO/SLI frameworks and error budgets in production environments.Contributions to the open source community (GitHub, Stack Overflow, technical blogging). Prior experience in the cybersecurity or intelligence fields.#LI-SC1 Benefits of Working at CrowdStrike:Market leader in compensation and equity awardsComprehensive physical and mental wellness programs Competitive vacation and holidays for recharge Paid parental and adoption leavesProfessional development opportunities for all employees regardless of level or roleEmployee Networks, geographic neighborhood groups, and volunteer opportunities to build connectionsVibrant office culture with world class amenitiesGreat Place to Work Certified across the globeCrowdStrike is proud to be an equal opportunity employer. We are committed to fostering a culture of belonging where everyone is valued for who they are and empowered to succeed. We support veterans and individuals with disabilities through our affirmative action program.CrowdStrike is committed to providing equal employment opportunity for all employees and applicants for employment. The Company does not discriminate in employment opportunities or practices on the basis of race, color, creed, ethnicity, religion, sex (including pregnancy or pregnancy-related medical conditions), sexual orientation, gender identity, marital or family status, veteran status, age, national origin, ancestry, physical disability (including HIV and AIDS), mental disability, medical condition, genetic information, membership or activity in a local human rights commission, status with regard to public assistance, or any other characteristic protected by law. We base all employment decisions--including recruitment, selection, training, compensation, benefits, discipline, promotions, transfers, lay-offs, return from lay-off, terminations and social/recreational programs--on valid job requirements.If you need assistance accessing or reviewing the information on this website or need help submitting an application for employment or requesting an accommodation, please contact us at View email address on click.appcast.io for further assistance.Find out more about your rights as an applicant.CrowdStrike participates in the E-Verify program.Notice of E-Verify ParticipationRight to WorkCrowdStrike, Inc. is committed to fair and equitable compensation practices. Placement within the pay range is dependent on a variety of factors including, but not limited to, relevant work experience, skills, certifications, job level, supervisory status, and location. The base salary range for this position for all U.S. candidates is $140,000 - $215,000 per year, with eligibility for bonuses, equity grants and a comprehensive benefits package that includes health insurance, 401k and paid time off.For detailed information about the U.S. benefits package, please click here. SummaryLocation: USA - New York, NY; USA - Austin, TX; USA - Sunnyvale, CA; USA - Redmond, WAType: Full time
$186.9k - $267.7k
...are received.This is a Hybrid position requiring... ...approximately 2 days per week on-site at Cisco offices in... ...intended, improving reliability and reducing risks.... ...Staff Site Reliability Engineer (SRE), you will... ...Agent Observability's platform. You will define the long...SuggestedFull timeTemporary workLocal areaFlexible hours2 days per week$158.5k - $172k
...technology, easy-to-use platforms, and an improved... ...OpportunityAs a Senior Engineer on the Runtime Automation... ...automate, and secure our core infrastructure,... ...position driving continuous reliability, deep system... ...(e.g., EKS, GKE).Our hybrid model requires 3 days...SeniorFull timeWork at office3 days per week$141k - $216.6k
...Axon. Together, we are creating the only platform that combines modern 911... ...connected world.Position OverviewAs a Site Reliability Engineer, you'll own the reliability, observability... ...company's most important products.This is a hybrid role, with an expectation of four days...SeniorWork experience placementWork at office$165k - $241.4k
...This role follows a hybrid work model, with... ...operates our US GovCloud platform. This team is... ...Federal region’s core infrastructure... ...looking for talented engineers with a software or... ...teams to ensure the reliability, performance and... ...the Cisco careers site to discover more benefits...SeniorFull timeTemporary workWork at officeLocal areaFlexible hours1 day per week- ...are seeking an experienced Site Reliability Engineer (SRE) – Microsoft Hyper-V &... ...Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V.... .... Microsoft Hyper-V (Core Expertise) Deep... ...Replica. Exposure to hybrid cloud and private cloud platforms...SeniorLocal area
$104.9k - $174.7k
...member of myGwork – the largest global platform for the LGBTQ+ business community.... ...We are hiring a hands-on Senior Site Reliability Engineer (SRE) to actively build, operate, and... ...near one of our offices, you may work a hybrid schedule. If not, this role is fully remote...SeniorFull timeWork at officeLocal areaRemote workFlexible hours$185k - $227k
.... ROLE AND RESPONSIBILITIES: A Senior Site Reliability Engineer (SRE) is expected to own the operational stability and performance ofJuul’s hybrid cloud infrastructure (Nutanix, AWS/GCP... ...theplatform is scalable and efficient. Nutanix Platform Management Design, deploy, and...SeniorRemote work$153k - $210k
...Senior Software Engineer, Site Reliability Engineering Reno, NV; San Ramon, CA; NYC - Hybrid Are you passionate about building... ...resilient, highly available cloud platforms that enable engineering... ...of experts bound by a set of core values and motivated to revolutionize...Senior$220k - $235k
...Staff/Senior Staff Site Reliability Engineer Ironclad is the leading AI contracting platform that transforms agreements into assets. Contracts move faster, insights... ...Sequoia, BOND, and Franklin Templeton. This is a hybrid role. Office attendance is required at least...SeniorFull timeContract workWork at office$138.1k - $198.2k
...simply works. The SRE Engineering Enablement Team supports our CI Platforms, Developer Environments... ...lifecycle (SDLC), including core developer tools and... .... Your Impact As a Site Reliability Engineer, you will be at... ...Qualifications Experience in a hybrid cloud environment (bare...Permanent employmentFull timeTemporary workWork experience placementLocal areaRemote workFlexible hours$110k - $120k
...technology.Job DescriptionJob Title: Site Reliability Engineer (SRE) / L3 Support EngineerGetting to... ...financeWhy You Will Love It Here! Flexibility: Hybrid Work Model & a Business Casual Dress... ...Engineer (SRE) to join our Platform Engineering team and take ownership of...Ongoing contractFull timeCasual workRemote workFlexible hours$190k - $210k
...imaging products, Butterfly Embedded™ is the Company's... ...semiconductor chip and software platform. Butterfly's innovations have... ...We're looking for a Staff Site Reliability Engineer to raise the bar for how we... ...Location Butterfly offers a hybrid work model for most positions...H1bWork at officeImmediate start2 days per week3 days per week$194k - $267k
...too, let's talk.The TeamThe Site Reliability team is dedicated to architecting... ...tooling and CI/CD platforms that support Okta’s SRE ecosystem... ...maximize platform reliability and engineering velocity.The ideal candidate... ...(FAR) 2.101#LI-SM1#LI-Hybrid P17611_3494641The annual base...Local areaWorldwideFlexible hours$131k - $164k
...OverviewWe are seeking a highly skilled Staff Site Reliability Engineer with deep technical expertise across... ...and performance of our global platform and drive continuous improvement through... ...and service account management across hybrid environments. Collaborate with network...Work at officeLocal areaVisa sponsorshipFlexible hours- ...team of researchers, engineers, designers, and more,... ...performance, scalable and reliable machine learning... ...next generation of AI platforms powering advanced NLP... ...We are looking for a Site Reliability Engineer to... ...multi-cloud on-prem / hybrid servingExperience in designing...Full timeWork experience placementWork at officeLocal areaRemote workHome office
$194k - $267k
...concepts and tools.Position Overview:The Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications... ...during the first week of employment.#LI-Hybrid#LI-LSS1requisition ID- (P16373_3396241)The...Permanent employmentWork at officeLocal areaWorldwideFlexible hours$150k - $160k
Front-End & AdTech Site Reliability Engineer (SRE)Haymarket Media, Inc. is seeking a Front-End & AdTech... ...monetization via Prebid.js against Core Web Vitals. In this hands-on role you... ...fieldNice To Have SkillsExperience with hybrid bidding strategies or Prebid Server migration...Work at officeLocal area- ...globally, the Zscaler Zero Trust Exchange platform combined with advanced AI combats... ...cybersecurity. Role We are looking for a Site Reliability Engineer-SkillBridge Intern (San JosA Ca or... ...about Zscaler’s Future of Work strategy, hybrid working model, and benefits here . By...InternshipWork at officeLocal areaRemote workWorldwide
$115k - $125k
...Site Reliability Engineer New York City, NY Pico fuels the global capital markets community by providing exceptional market data services... ...any given situation. Working Arrangements This is a hybrid position with weekly time in the office with the flexibility...Work experience placementWork at officeWork from homeMonday to FridayFlexible hoursShift workWeekend workAfternoon shiftEarly shift$111k - $160k
...Join Mizuho as a Site Reliability Engineer! In this role you will play a crucial role in maintaining... ...Use Grafana and other monitoring platforms to track system reliability and performance... ...Mizuho has in place a hybrid working program, with varying opportunities...Work at officeLocal areaRemote workWorldwide$86k - $105k
...Francisco, Chicago, or New York follow a hybrid work model to allow for a more... ...infrastructure and to be responsible for reliability, automation and scalability using and the... ...Minimum of 2 years prior DevOps, software engineering or related experience. Must be able...Hourly payWork at officeImmediate startVisa sponsorshipWork visaFlexible hours$139k - $257.55k
...Community CCM organization is seeking an outstanding Senior Site Reliability Engineer (SRE) to support innovation through machine learning,... ...advanced services. Our infrastructure spans multiple cloud platforms and is orchestrated by cloud-native, containerized systems...SeniorFull timeTemporary workLocal areaRemote workWorldwide$113.1k - $232.3k
Position Summary Lead Applied AI Site Reliability Engineer II Role Overview: As a Lead Applied... ...of high-visibility products and platforms and the environments they run in. Your... ...segregation-of-duties line as a dedicated, embedded function while partnering with...Work at officeLocal areaVisa sponsorshipFlexible hours3 days per week$120k - $142k
...Responsibilities: At The New York Times, our Site Reliability Engineering (SRE) team is central to how we... ...for reliability programs and platforms that help teams ship resilient systems... ...experts in every SRE domain.This is hybrid role based in New York City, NY.Responsibilities...Local areaFlexible hours- Guide and shape the future of technology at a globally recognized firm, driven by pride in ownership.As a Senior Manager of Site Reliability Engineering at JPMorgan Chase within the Corporate Investment Bank, Markets team, you are the non-functional requirement owner and...SeniorBank staffShift work
$160k - $180k
...Socure is seeking a Site Reliability Engineer in New York to enhance our identity trust infrastructure. In this role, you will take full ownership of AWS and Kubernetes platforms, ensuring high reliability and operability. The ideal candidate will possess extensive experience...Senior- ...Karsun Solutions, LLC is seeking a Site Reliability Manager to lead a multi-disciplinary team responsible for reliability, security, and platform lifecycle across AWS-based services. The role emphasizes collaboration, observability, and continuous improvement in a client...Senior
- ...Karsun Solutions in the DMV area is seeking a Site Reliability Manager to ensure reliability, scalability, and performance of our systems.... ...lead a team focusing on Application Reliability, DevSecOps, and Platform Lifecycle Management. The ideal candidate has 10+ years in...Senior
- ...adaptive training and intervention. For higher-risk users, our platform integrates seamlessly with the broader security stack to... ..., more resilient organizations. The Role: As a Senior Site Reliability Engineer (SRE) at Dune Security, you will play a critical role in ensuring...SeniorFull timeWork at office
$225k - $325k
...-day Ensure the scalability, reliability, and observability of our systems to maintain and improve the firm's core infrastructure environment. Lead a range of engineering projects, from developing proprietary platforms for configuration management and monitoring...SeniorHourly pay
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid). Be the first to apply!
- site reliability engineering manager New York, NY
- site reliability engineer New York, NY
- site reliability engineer sre New York, NY
- site reliability engineer remote New York, NY
- data platform engineer New York, NY
- platform engineer New York, NY
- platform engineering manager New York, NY
- client platform engineer New York, NY
- senior platform engineer New York, NY
- platform developer New York, NY

