Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Infra Staff Systems Engineer Onsite or Remote San Francisco, California

$200k - $250k

Jobleads-US

At Evidently, we are raising the quality of care for every patient by empowering clinicians with the right information at the speed of thought. Our Clinical Data Intelligence platform extracts meaning from an ocean of clinical notes, outside records, and scanned documents — and delivers it to clinicians in the moment it matters. We're a small, mission-driven team building something genuinely hard and genuinely impactful. Learn more at evidently.com.

Who We're Looking For

We're looking for a Senior Systems Engineer who owns reliability the way a founding engineer owns their product: with a sense of personal accountability, eyes everywhere, and no tolerance for fires that were preventable.

This is a clear home-base role. Infrastructure and reliability are your domain and your primary responsibility. You are the person who notices before anyone else that something is trending wrong. You work alongside leadership as a partner in operational decision-making, not getting paged when things break, but instead making sure they don't break in the first place, and standing shoulder-to-shoulder with the team when they do. You bring a calm, systematic approach to incident response, help divide up the work intelligently under pressure, and run thorough postmortems that actually change behavior.

At the same time, you have the seniority and breadth to patrol adjacent territory. Integration work, security and compliance posture, database performance, AI inference infrastructure: you can cover these areas when they touch systems, and you're not waiting to be asked. You see the whole field.

We are an AI-first company, and that applies to how you work as much as what you build. We expect you to use AI tools aggressively, as an operator who directs them, not a passive user who accepts their output uncritically.

What You'll Do

Own Reliability

  • Hold end-to-end accountability for the reliability, availability, and performance of Evidently's production systems: our SMART on FHIR platform, AI inference pipelines, data systems, and the integrations that connect them all.
  • Define and own SLOs, SLIs, and error budgets for critical services; build the dashboards and alerting that make the health of the system visible to everyone, not just you.
  • Lead incident response: triage fast, communicate clearly, coordinate the team, contain blast radius. Run blameless postmortems that actually result in durable fixes and systemic improvements, not just tickets.
  • Proactively identify reliability risks before they become incidents, through capacity planning, load testing, chaos engineering, and systematic review of production signals.
  • Build runbooks and on-call practices that make rotating on-call humane and effective, not a rotation people dread.

Build and Operate Infrastructure

  • Design, implement, and maintain highly available, scalable, and secure cloud infrastructure on Google Cloud Platform: GKE, Cloud Run, AlloyDB, Cloud Storage, Pub/Sub, and adjacent services.
  • Own Infrastructure as Code across all environments with a cloud-agnostic, open-minded, code-centric approach, applying standard software engineering discipline (version control, code review, testing, promotion gates).
  • Build and maintain CI/CD pipelines (GitHub Actions) that make deployments fast, safe, and boring, with rollback, canary, and feature toggling patterns baked in.
  • Manage observability infrastructure end-to-end: instrumentation, tracing, log pipelines, alerting, and dashboards structured around what actually matters operationally.
  • Own data infrastructure health across GCP data systems, including AlloyDB, Cloud Storage, BigQuery, and Cloud Logging: performance tuning, query analysis, storage lifecycle management, connection management, backup/restore, and disaster recovery.
  • Own AI and LLM inference infrastructure end-to-end: Vertex AI, Model Garden, and model serving platforms (vLLM, Cloud Run, GPU provisioning), latency SLOs, cost optimization, auto-scaling characteristics, and capacity planning for production AI workloads.

Set the Standard on Security and Compliance

  • Maintain and improve our security posture across all infrastructure: network segmentation, IAM, secrets management, vulnerability management, audit logging.
  • Own infrastructure-level compliance for HIPAA and SOC 2, not just checkbox compliance, but genuinely secure and auditable systems.
  • Be a thought partner to engineering on security architecture decisions, particularly where AI workloads introduce new data handling surface area.

Patrol Adjacent Territory

  • Cover integration infrastructure where it intersects with systems: EHR integrations, third-party APIs, data pipelines, and AI inference serving. You don't need to own these, but you understand them well enough to step in.
  • Support backend engineers on runtime performance: Python async, FastAPI/Django serving characteristics, concurrency patterns, and where infrastructure choices constrain application-level performance.

What You'll Bring

  • 8+ years of experience in SRE, infrastructure engineering, or DevOps, with a clear track record of owning reliability, not just contributing to it.
  • You think in SLOs. You have defined them, defended them, and used error budgets to make real decisions about when to slow down feature work.
  • Deep experience with incident management: you've led responses to serious production incidents, written the postmortems, and driven the follow-through.
  • Expert-level command of GCP (GKE, Cloud Run, Cloud SQL/AlloyDB, Pub/Sub, IAM, VPC, Cloud Armor) or directly comparable experience on AWS or Azure with a credible GCP migration story.
  • Very strong shell scripting skills (Bash) and fluency in Python for automation; comfortable reading application code well enough to understand its operational behavior.

Depth That Matters

  • Observability done right: you've built meaningful monitoring from the ground up, not just enabled a tool. You know the difference between dashboards that look good and dashboards that catch problems.
  • Deep knowledge of GCP data infrastructure and database operational characteristics: query planning and tuning in PostgreSQL/AlloyDB/BigQuery, log management and retention in Cloud Logging, data lifecycle management in Cloud Storage, analytical workloads in BigQuery, and how to diagnose issues across these services as they scale.
  • Understanding of container and Kubernetes internals well beyond YAML: scheduling, resource management, network policies, pod disruption budgets, autoscaling behavior.
  • Security and compliance grounding: you've operated in regulated environments (HIPAA, SOC 2, or similar) and understand the difference between compliance theater and systems that are actually secure.
  • Deep operational mastery of AI and LLM inference serving, including Vertex AI and Model Garden: GPU provisioning and orchestration, model deployment strategies, optimizing inference latency and throughput, GPU cost controls, and managing the unique reliability patterns of production LLM workloads at scale.

Mindset

  • You have a strong operational instinct. You notice anomalies, you ask why things are trending the way they are, and you're not satisfied until you understand the root cause.
  • You treat reliability as a product you own, not a set of tasks assigned to you.
  • You're calm under pressure, structured in how you communicate during incidents, and rigorous in follow-through after them.
  • You're senior enough to have strong opinions about how things should be done and humble enough to change them when presented with better evidence.
  • You use AI tools fluently as a practitioner, writing infrastructure code, debugging incidents, drafting runbooks, with the judgment to know when the output is right and when it needs correction.

Bonus

  • Familiarity with healthcare data standards: FHIR, HL7, SMART on FHIR application architecture.
  • Experience operating infrastructure for LLM or AI inference workloads at production scale.
  • Background in building internal developer platforms or golden-path tooling that makes other engineers more effective.
  • Strong hands-on experience or familiarity with Infrastructure as Code tooling such as Kubernetes, Terraform, or Pulumi.
  • Experience with OpenTelemetry for instrumentation and distributed tracing across complex environments.

Who You'll Work With

You'll join a small, highly effective, self-driven engineering team where resourcefulness is prized and engineers operate with a high degree of autonomy. You will serve as a peer to the CTO and technical leadership—not as a support function, but as a co-owner of how the platform runs. You'll work closely with backend, data, and AI/ML engineers on the systems they build and depend on, and directly with leadership on operational priorities and incident response. The team is concentrated in the Bay Area with a weekly San Francisco office day; remote candidates are welcome, though you must be able to work core Pacific hours and must be based in the US.

Compensation: $200k – $250k + equity

We're a team on a mission to elevate the quality of care for every patient, by empowering clinicians with the right information at the speed of thought. Learn more about us.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Infra Staff Systems Engineer Onsite or Remote San Francisco, California in San Francisco, CA vacancy
  • Staff Engineering Program Manager, Core Tech Our mission at Oura is to...  ...initiatives across multiple system technologies leveraging...  ...This role will be located in San Francisco, California with in-office...  ...recruiters, especially for remote roles. Please note: Our jobs... 
    Remote work
    Work at office
    Local area
    Flexible hours

    Ouraring

    Brooklyn, NY
    2 days ago
  • $166k - $343k

    Pre-Sales Systems Engineer - Higher Education and State Government (Sacramento/San Francisco, CA)This role has been designated as ‘Remote/Teleworker’, which means you will primarily work from...  ...immediately.SummaryLocation: All, California, United States of America;... 
    Remote work
    Full time
    Work experience placement
    Local area
    Immediate start
    Work from home

    Hewlett Packard Enterprise

    San Francisco, CA
    3 days ago
  • $166k - $343k

     ...Pre-Sales Systems Engineer - Higher Education and State Government (Sacramento/San Francisco, CA)This role has been designated as ‘Remote/Teleworker’, which means you will primarily work from...  ...Salary USD 166,000 - 343,000 in California This range reflects the minimum... 
    Remote work
    Work experience placement
    Work from home

    Jobleads-US

    Sacramento, CA
    5 days ago
  • $225k - $350k

    # Staff AI EngineerFast-growing software company* LocationSan Francisco, California* TypeFull-time* Posted6 October 2026Apply...  ...and ship software systems layered on top of...  ...a founder or early engineer.## Benefits* Health,...  ...Location and work model* San Francisco,... 
    Suggested
    Full time
    Work at office
    Flexible hours

    Jobleads-US

    San Francisco, CA
    1 day ago
  • $130k - $140k

     ...residential experience for entrepreneurs in San Francisco. They are looking for a proactive, high-...  ..., offering more flexibility to work remotely and/or take time off 5+ years of experience...  ...periods San Francisco, CA (Mainly onsite - with flexibility to WFH depending on... 
    Remote work
    Work from home
    Flexible hours
    Shift work
    Afternoon shift

    Burke + Co.

    San Francisco, CA
    4 days ago
  •  ...edge AI startup in San Francisco to hire a Full Stack AI Engineer with strong zero-to...  ...experience.This is a fully onsite position....  ...if you are seeking remote or hybrid work.About...  ...infrastructure, and AI systems. You should enjoy...  ...Francisco, California - 0Physical Therapist... 
    Remote work
    Full time
    Live in
    Work at office
    Local area
    Relocation

    Affinity Executive Search

    San Francisco, CA
    2 days ago
  • $101.25k - $126.56k

     ...Technical Consultant I#### San Francisco, California, United States**Why join us...  ...four weeks per year of fully remote work!**Responsibilities***...  ...integration of customer systems with Brex, managing the end...  ...insights for our Product and Engineering partners.* Actively contribute... 
    Remote work
    Work at office
    Immediate start
    Work from home

    Brex

    San Francisco, CA
    1 day ago
  • $320k - $405k

     ...Staff / Senior Physical Security Systems Engineer Remote-Friendly (Travel Required) | San Francisco, CA About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and... 
    Remote job
    For contractors
    Work at office
    Visa sponsorship
    Flexible hours

    Jobleads-US

    San Francisco, CA
    2 days ago
  •  ...talent-first flexible/remote work approach, see...  ...direction for our evaluation systems; metric, judge and...  ...alignment across engineering, product and AI leaders...  ...accordance with the San Francisco Fair Chance Ordinance...  ...Chance Ordinance, the California Fair Chance Act, and... 
    Remote job
    Full time
    Local area
    Flexible hours

    Jobleads-US

    Kentucky
    2 days ago
  •  ...Manager - Analytical Instrumentation San Francisco Area (Remote) Swan Analytical USA Full-...  ...our presence throughout the Northern California region. This is an excellent...  ...For Bachelor's degree preferred (Engineering, Chemistry, Environmental Science, Business... 
    Remote work
    Base plus commission
    Full time
    Work at office
    Worldwide
    Monday to Friday

    Swan Analytical USA

    San Francisco County, CA
    3 days ago
  • $50 per hour

     ...may also elect to schedule in-person clinic sessions at our San Francisco location by selecting from available clinic blocks, at their...  ...ability to obtain additional state licenses as needed Valid California DEA Registration in good standing with all schedules to prescribe... 
    Remote work
    Hourly pay
    Live in
    Afternoon shift

    Circle Medical - a UCSF Health Affiliate

    San Francisco, CA
    1 day ago
  •  ...RoleWe are seeking a highly motivated Systems Engineer to help shape the architecture and integration...  ...Applicants for Jobs Located in NYC or Remote Jobs Associated With Office in NYC...  ...of non-discrimination.Pursuant to the San Francisco Fair Chance Ordinance, Los Angeles... 
    Remote work
    Hourly pay
    Work at office
    Local area
    Flexible hours

    Doordash

    Oakland, CA
    1 day ago
  • $166k - $343k

    Presales Systems Engineer - HPE Networking (Northern California)This role has been designated as ‘Remote/Teleworker’, which means you will primarily work from home.Who We Are:Hewlett...  ...States of America; Oakland, California; San Francisco, California, United States of America;... 
    Remote work
    Full time
    Work experience placement
    Local area
    Immediate start
    Work from home

    Hewlett Packard Enterprise

    San Francisco, CA
    2 days ago
  • $114.8k - $165.2k

     ...are looking for a Staff Reliability Engineer to join our team in...  ...position will be fully Onsite based in San Jose, CA Role...  ...Bloom Energy fuel cell systems. Work with others...  ...Field Service and Remote Monitoring to...  ...headquartered in San Jose, California, is growing quickly... 
    Remote work
    Work experience placement
    Worldwide
    Shift work

    101 Bloom Energy

    Concord, NH
    4 days ago
  •  ...all of their business systems through natural language...  ...Moveworks’ Reasoning Engine and natural language...  ...all started in sunny San Diego, California in 2004 when a visionary...  ...learning systems. The ML infra team covers a variety...  ...personas (flexible, remote, or required in office... 
    Remote work
    Permanent employment
    Work at office
    Flexible hours

    Moveworks

    Mountain View, CA
    3 days ago
  • $122.6k - $183.95k

    Job SummaryThe Sr. Staff OT Systems Engineer is a senior technical leadership role responsible for designing, implementing, integrating, and supporting...  ..., including segmentation, switching, routing, firewalls, remote access, servers, virtualization, identity management,... 
    Remote work
    Full time
    Work at office

    Insulet Corporation

    Acton, MA
    2 days ago
  • $160k - $200k

     ...autonomous logistics system, delivering critical supplies...  ...issues requires an engineer who can move across...  ...telemetry, fault codes, and remote-debug capabilities so...  ...-cause methods.At the Staff level, lead cross-...  ...role based in South San Francisco and requires domestic... 
    Remote work
    Permanent employment
    Local area

    Zipline

    South San Francisco, CA
    21 hours ago
  •  ...focused practice in the heart of San Francisco. This is a great opportunity...  ...re a licensed Optometrist in California with the required liability...  ...For roles that are remote (i.e., Work From Home (WFH)) or hybrid (i.e., partial onsite at a VSP location and WFH), must... 
    Remote work
    Full time
    Local area
    Work from home
    Flexible hours

    VSP Vision

    San Francisco, CA
    5 days ago
  •  ...from leading investors. We are based in San Francisco and work in person. The Role You...  ...OTA updates, logging, monitoring, and remote diagnostics for the fleet Profile and...  ...Experience with real-time or embedded Linux systems Familiarity with ROS 2 or similar... 
    Remote work

    Jobleads-US

    San Francisco, CA
    2 days ago
  • $114k - $182k

    Job TitleSenior/Staff Ultrasound System Engineer - Medical Device (San Diego, CA)Job DescriptionSenior/Staff Ultrasound System Engineer - Medical Device (San Diego...  ...means working in-person at least 3 days per week. Onsite roles require full-time presence in the company’s... 
    Full time
    Work at office
    Immediate start
    Work visa
    Relocation package
    3 days per week

    Philips

    San Diego, CA
    21 hours ago
  • $270k - $290k

     ...training, and deploying AI systems—designed to take ideas from...  ...with offices in New York City, San Francisco, Seattle, and London, and is...  ...NYC office, we are open to remote US candidates for this role....  ...Leadership with Product Management, Engineering, Revenue Teams, and within... 
    Remote work
    Work at office
    Local area
    Work from home
    Flexible hours

    Jobleads-US

    San Francisco, CA
    2 days ago
  • $297k - $350k

     ...quality. We are looking for a Senior Staff Software Engineer to set technical direction across...  ...spans multiple engineering teams across San Francisco, New York, Vancouver, and Warsaw....  ...Org that includes AI platform, Agent systems and Studio in partnership with engineering... 
    Work at office
    Local area
    Work from home
    Worldwide
    Flexible hours
    Shift work

    Jobleads-US

    San Francisco, CA
    1 day ago
  • $68.9k - $131.1k

     ...Date: 2026-10-06Location: US-CA-SAN DIEGO-SD1 ~ 8650 Balboa Ave ~...  ...BLDG Position Role Type:Onsite U.S. Citizen, U.S. Person, or...  ...at RTX.What You Will DoAs a Systems Engineer at Raytheon, you will integrate...  ...as on-site, hybrid or remote.The salary range for this role... 
    Remote work
    Full time
    Temporary work
    Work experience placement
    Internship
    Work at office
    Relocation
    Flexible hours

    Raytheon

    San Diego, CA
    12 hours ago
  • $234.84k

     ...global Site Reliability Engineering (SRE) function to...  ...uses code, data, and systems thinking to make product...  ...safe and fast. As a Staff Site Reliability...  ...TTC / OTE) Zone 1: San Francisco Bay Area, New York City...  ...Zone 2: Washington, California (excluding San Francisco... 
    Remote work
    Base plus commission
    Local area
    Worldwide
    Flexible hours
    Shift work

    Veeam Software

    United States
    2 days ago
  • $83k - $108k

     ...looking for a Key Account Manager for our San Francisco, California area. In this role, you will be...  ...classes and so much more! Location: Remote - US Actual compensation may vary depending...  ...have any difficulty using our online system. If you need such an accommodation,... 
    Remote work
    Full time
    Temporary work
    Work experience placement
    Work at office
    Flexible hours

    iRhythm Technologies

    San Francisco, CA
    2 days ago
  •  ...of all of us—from design and engineering to the manufacturing and...  ...RoleWe’re looking for a Sr. Staff Systems Engineer with a strong systems...  ...teams in Ireland, including onsite work to support system integration...  ...and work-life balance. Remote or field-based positions will... 
    Remote work
    Full time
    Work at office

    BD, Becton and Dickinson

    Tempe, AZ
    1 day ago
  • $125.7k - $220k

     ...DescriptionIt all started when engineer Fred Luddy wrote code that...  ...screening.About the Role:The systems we build run on a variety of cloud...  ...below:Cloud (Azure, AWS, or GCP)Infra as code (Terraform, cloud...  ...trust. Work personas (flexible, remote, or required in office) are categories... 
    Remote work
    Permanent employment
    Work at office
    Immediate start
    Flexible hours

    ServiceNow

    Toronto, OH
    21 hours ago
  • $7.86k - $10.54k

     ...problem-solving, system reliability, and helping...  ...up to two days of remote work and three days or more onsite, per week....  ...July 1, 2025, The California Department of Human...  ...Title: Systems Engineer Classification...  ...the Napa Valley, San Francisco, Lake Tahoe, and... 
    Remote work
    Permanent employment
    Full time
    Work at office
    Monday to Friday
    Flexible hours

    California Correctional Health Care Services

    Sacramento, CA
    1 day ago
  • $177k - $265.6k

     ...opportunities to work on revolutionary systems that impact people's lives...  ...and Verification - Staff Systems Engineer to join our team of diverse...  ...role can be located in either San Diego , El Segundo , or...  ...our facility. There is no remote / hybrid / telework available... 
    Remote work
    Full time
    Relocation package
    Shift work

    Northrop Grumman

    El Segundo, CA
    1 day ago
  • $104.9k - $194.7k

     ...DescriptionPerforms technical planning, system integration, verification and validation,...  ...perform IT responsibilities, use E146, Systems Engineer.Basic Qualifications• Active Secret...  ...software, data, and procedures.• Full-time onsite role in Orlando, FL, working the majority... 
    Full time
    Temporary work
    Work experience placement
    Casual work
    Flexible hours

    Lockheed Martin

    Orlando, FL
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Infra Staff Systems Engineer Onsite or Remote San Francisco, California. Be the first to apply!