Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Senior Site Reliability Engineer

XRC Ventures

The Role

We're looking for a Senior Site Reliability Engineer to own the reliability, scalability, and operational excellence of the production systems that power Nectar's platform. We run high-volume data ingestion pipelines and real-time AI agents on top of a fast-growing customer base, and we need a seasoned SRE to help us scale these systems safely and keep them running flawlessly.

As one of our first dedicated SREs, you'll have outsized impact and ownership. You'll define how we measure, operate, and harden our infrastructure -- establishing the reliability foundations that the rest of the engineering team builds on as we scale.

What You'll Be Doing
  • Own the reliability and scalability of our production systems as they handle rapidly growing volumes of social data and real-time AI workloads

  • Define and drive SLOs, SLIs, and error budgets, and build the observability, alerting, and on-call practices to support them

  • Lead incident response and blameless postmortems, then turn what we learn into systemic improvements that prevent recurrence

  • Improve performance, cost efficiency, and capacity planning across our cloud infrastructure as the platform scales

  • Harden our infrastructure-as-code, deployment, and CI/CD pipelines for resilience and repeatability

  • Partner with engineering teams to embed reliability into system design and raise the operational bar across the org

What We're Looking For
  • 5+ years of experience operating production systems as an SRE, infrastructure, or platform engineer

  • Experience scaling databases, data infrastructure, or complex production platforms under significant load

  • Hands-on expertise with cloud infrastructure (AWS or similar) and infrastructure-as-code tooling

  • Solid programming skills for building automation, tooling, and operational services

  • Comfortable operating in fast-moving startup environments with high ownership and autonomy

  • A reliability-first mindset balanced with pragmatism about velocity and cost

Bonus Points
  • Experience standing up or maturing an SRE practice at an early-stage or rapidly scaling company

  • Familiarity with our tech stack: AWS, Pulumi, Postgres, ClickHouse, Turbopuffer, or Temporal

  • Background in capacity planning, performance engineering, or cost optimization at scale

What We Offer
  • Competitive compensation and early equity

  • Health, vision, and dental benefits + 401(k) match

  • Clear career growth opportunities as the company scales

  • Free lunch in the heart of University Ave. in Palo Alto

  • Deep exposure to cutting-edge AI tooling and the opportunity to shape how brands use it

  • A collaborative, ambitious team defining a new category of AI-native marketing infrastructure

Vacancy posted 5 days ago
Similar jobs that could be interesting for youBased on the Senior Site Reliability Engineer in Palo Alto, CA vacancy
  • $158k - $225k

     ...Senior Site Reliability Engineer (SRE) Manufacturing advanced electronics requires understanding millions of signals generated across complex assembly processes. Instrumental builds systems that capture and analyze those signals — images, test results, and process... 
    Senior

    Instrumental Inc

    Palo Alto, CA
    2 days ago
  • $150k - $175k

     ...Site Reliability Engineer At ASAPP, our mission is simple: deliver the best AI-powered customer experience—faster than anyone else. To achieve that, we're guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed... 
    Senior
    Remote work

    ASAPP

    Mountain View, CA
    4 days ago
  •  ...Site Reliability Engineer There are NO limits to your career: come shape the future and be part of a truly unique global culture at OutSystems! Hybrid Onsite in Menlo Park, CA Site Reliability Engineering (SRE) is a discipline that incorporates aspects of software... 
    Senior
    Immediate start
    Remote work
    Worldwide

    OutSystems

    Menlo Park, CA
    3 days ago
  • $137.77k - $194.59k

     ...distributed team of roughly 80 scientists and engineers building and operating Rubin's petascale...  .... Your role: You will own the reliability and robustness of Rubin Observatory's...  ...of this position, SLAC is open to on-site, hybrid, and remote work options. Work... 
    Senior
    Remote work
    Flexible hours
    Night shift

    Stanford University

    Menlo Park, CA
    3 days ago
  •  ...Senior Site Reliability Engineer Latitude AI develops automated driving technologies, including L3, for Ford vehicles at scale. We're driven by the opportunity to reimagine what it's like to drive and make travel safer, less stressful, and more enjoyable for everyone... 
    Senior
    Work at office
    Immediate start

    Latitude AI

    Palo Alto, CA
    3 days ago
  •  ...join our small team focused on growth and productivity. The role involves scaling our platform and infrastructure while enhancing reliability and the overall developer experience. Ideal candidates will have strong expertise in distributed systems, cloud-native... 
    Senior
    Remote job

    BuildBuddy

    Palo Alto, CA
    1 day ago
  • $180k - $260k

     ...effortless integration into customers' logistics operations. About the role We are seeking an experienced Senior/Staff Site Reliability Engineer to support the operation, monitoring, and scaling of our growing fleet of autonomous vehicles. In this role, you will... 
    Senior
    Odd job
    Work at office
    Remote work

    Gatik AI

    Mountain View, CA
    3 days ago
  • $140k - $220k

    About the Job You’ll own reliability and operational excellence for Pylon’s production systems. This means designing and implementing...  ...scale as we grow. You’ll build tooling that makes the entire engineering team more effective, establish on‑call rotations and runbooks... 
    Senior

    Pylon

    Palo Alto, CA
    3 days ago
  • $181k - $197k

    Senior SRE Palo Alto, CA • Engineering • Hybrid • Full-time Founded by a team of ex-Apple engineers, Instrumental provides a collection of software...  ...on, and measuring KPIs to ensure ongoing performance, reliability and efficiency. Network/application security and... 
    Senior
    Full time

    Clutch Canada

    Palo Alto, CA
    5 days ago
  • JPMorgan Chase & Co. is seeking a Director of Site Reliability Engineering to partner with the Infrastructure Platforms and Foundational Services team in Palo Alto. This role involves guiding stakeholders through complex projects, leading the application of AI capabilities... 
    Senior

    JPMorgan Chase & Co.

    Palo Alto, CA
    2 days ago
  • A leading tech recruiting firm is seeking a Site Reliability Engineer to manage and optimize cloud infrastructure primarily using GCP or AWS. The role involves maintaining high availability through Kubernetes clusters and improving CI/CD pipelines with Terraform. Ideal... 
    Senior

    Amiri Recruiting

    Mountain View, CA
    1 day ago
  • Cerebras is looking for a Senior Site Reliability Engineer to join their Infrastructure team in Palo Alto, California. This role involves designing and optimizing infrastructure for distributed AI applications, contributing to the open-source Ray project, and ensuring... 
    Senior

    Cerebras

    Palo Alto, CA
    2 days ago
  • $181.69k - $213.75k

     ...Senior Site Reliability Engineer San Francisco, California; Santa Clara, California; Seattle, WA The Company You'll Join Carta connects founders, investors, and limited partners through world-class software, purpose-built for everyone in venture capital, private... 
    Senior
    Full time
    Work at office

    Carta

    Santa Clara, CA
    3 days ago
  •  ...Senior Site Reliability Engineer LeanData helps the world's fastest-growing companies automate, simplify, and accelerate revenue. We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud infrastructure. Reporting directly... 
    Senior
    Full time
    Work at office
    Flexible hours
    2 days per week

    LeanData

    Santa Clara, CA
    3 days ago
  • $148k - $235.75k

     ...and Processes organization where you will be working as a Senior SRE Engineer. The position will be part of a fast-paced crew that develops...  ...Manage NVIDIA's on-prem infrastructure. Maintain uptime, reliability and readiness of on-prem engineering cloud spread across multiple... 
    Senior
    Remote work

    NVIDIA

    Santa Clara, CA
    4 days ago
  • $150k - $180k

    A technology-focused data center developer in Mountain View, CA is looking for a Senior Site Reliability Engineer to manage software infrastructure. This full-time position requires experience in Software Engineering or DevOps, with strong proficiency in Golang. The role... 
    Senior
    Full time

    Verrus, LLC

    Mountain View, CA
    3 days ago
  • $126k - $204.5k

     ...As part of this role, you will collaborate closely with our engineering teams to develop innovative solutions that provide clear and...  ...team to influence the operability of the product and ensure the reliability and availability of our services. Qualifications... 
    Senior
    Full time
    Work at office

    Palo Alto Networks

    Santa Clara, CA
    2 days ago
  •  ...building an AI Data Center AIOps platform that turns raw, high‑volume telemetry into reliable, job‑centric insights and automation for GPU fleets. Join our team of innovative engineers who are building this platform and operating it (not the compute cluster): uptime, performance... 
    Senior

    NVIDIA Gruppe

    Santa Clara, CA
    1 day ago
  • $174k - $252k

    Senior Software Engineer, Site Reliability Engineering X Applicants in San Francisco: Qualified applications with arrest or conviction records will be considered for employment in accordance with the San Francisco Fair Chance Ordinance for Employers and the California... 
    Senior
    Full time

    Google Inc.

    Sunnyvale, CA
    1 day ago
  • $152k - $241.5k

    Overview NVIDIA is looking for a Senior Site Reliability Engineer (SRE) to join our Compute Farm team and help build the next generation of our global services platform. The role focuses on keeping critical systems operational while leveraging AI technologies to deliver... 
    Senior

    NVIDIA Corporation

    Santa Clara, CA
    4 days ago
  • $152k - $241.5k

     ...intelligence. Job Overview We’re looking for a Senior SRE to join our Compute Farm team and...  ...host lifecycle management, fleet reliability/auto‑healing, E2E observability or data‑...  ...Python, Go, Perl, or Ruby. Mentored other engineers and influenced technical direction through... 
    Senior

    NVIDIA Gruppe

    Santa Clara, CA
    4 days ago
  •  ...capital, our team is looking to grow its world class team of engineers. We are transforming an age old industry with intelligent robotics...  .... Engineers write application code; you make sure it deploys reliably, scales correctly, stays secure, and is observable in... 
    Senior

    Robot RX

    Newark, CA
    1 day ago
  • $200k - $322k

    Senior Manager, Site Reliability Engineering page is loaded## Senior Manager, Site Reliability Engineeringlocations: US, CA, Santa Claratime type: Full timeposted on: Posted Yesterdayjob requisition id: JR2016119For over 25 years, NVIDIA has been at the forefront of transforming... 
    Senior

    NVIDIA Corporation

    Santa Clara, CA
    3 days ago
  • $145k - $165k

    A technology solutions firm in Sunnyvale, CA is looking for a highly experienced Site Reliability Engineer (SRE). This role involves maintaining uptime and performance across systems. Exceptional Linux expertise and automation skills in Bash and Python are crucial. Key... 
    Senior

    Bolt Graphics, Inc.

    Sunnyvale, CA
    3 days ago
  • $175k - $225k

     ...vector database company for enterprise-grade AI. Founded by the engineers behind Milvus, the world’s most popular open-source vector...  ...you will do: Work at the intersection of development and site reliability. Creating SRE tools and systems, as well as supporting existing... 
    Senior

    Zilliz

    Redwood City, CA
    4 days ago
  •  ...Senior Site Reliability Engineer (SRE) / DevOps Engineer Location: Onsite - Mountain View, CA Experience Required: 5+ years Infrastructure Footprint: Global production infrastructure across AWS, South America, and Europe Role Type: Hands-on engineering... 
    Senior
    Full time

    Prophet Town

    Mountain View, CA
    a month ago
  • $175.8k - $264.2k

    Senior Site Reliability Engineer - Apple Services Engineering (ASE) / iCloud Cupertino, CA People at Apple don't just build products - they craft experiences our customers love and depend on. Apple Services Engineering (ASE) builds and supports the systems that make many... 
    Senior

    Hong Kong Study Skills Research Institute

    Cupertino, CA
    4 days ago
  • $232.9k - $335.81k

     ...About the Role: We're looking for a Principal Site Reliability Engineer to join our Platform Engineering team - someone equally at...  ...infrastructure rather than daily firefighting. This is a senior individual-contributor role. You will not have direct reports... 
    Permanent employment

    Uniphore

    Palo Alto, CA
    22 hours ago
  • $170k - $230k

     ...Site Reliability Engineer (SRE) Palo Alto / San Francisco Bay Area About Mithril Mithril is an AI infrastructure platform built to make GPU compute more accessible and affordable for the world's leading enterprises, AI startups, and the AI research community,... 
    Work at office
    Local area
    1 day per week

    Mithril

    Palo Alto, CA
    3 days ago
  • $98.58k - $138.02k

     ...Site Reliability Engineer II Restaurant365 is a SaaS company disrupting the restaurant industry! Our cloud-based platform provides a unique, centralized solution for accounting and back-office operations for restaurants. Restaurant365's culture is focused on empowering... 
    Work at office

    Restaurant365

    Palo Alto, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Senior Site Reliability Engineer. Be the first to apply!