Infra Staff Systems Engineer Onsite or Remote San Francisco, California
$200k - $250kJobleads-US
- Remote job
At Evidently, we are raising the quality of care for every patient by empowering clinicians with the right information at the speed of thought. Our Clinical Data Intelligence platform extracts meaning from an ocean of clinical notes, outside records, and scanned documents — and delivers it to clinicians in the moment it matters. We're a small, mission-driven team building something genuinely hard and genuinely impactful. Learn more at evidently.com.
Who We're Looking For
We're looking for a Senior Systems Engineer who owns reliability the way a founding engineer owns their product: with a sense of personal accountability, eyes everywhere, and no tolerance for fires that were preventable.
This is a clear home-base role. Infrastructure and reliability are your domain and your primary responsibility. You are the person who notices before anyone else that something is trending wrong. You work alongside leadership as a partner in operational decision-making, not getting paged when things break, but instead making sure they don't break in the first place, and standing shoulder-to-shoulder with the team when they do. You bring a calm, systematic approach to incident response, help divide up the work intelligently under pressure, and run thorough postmortems that actually change behavior.
At the same time, you have the seniority and breadth to patrol adjacent territory. Integration work, security and compliance posture, database performance, AI inference infrastructure: you can cover these areas when they touch systems, and you're not waiting to be asked. You see the whole field.
We are an AI-first company, and that applies to how you work as much as what you build. We expect you to use AI tools aggressively, as an operator who directs them, not a passive user who accepts their output uncritically.
What You'll Do
Own Reliability
- Hold end-to-end accountability for the reliability, availability, and performance of Evidently's production systems: our SMART on FHIR platform, AI inference pipelines, data systems, and the integrations that connect them all.
- Define and own SLOs, SLIs, and error budgets for critical services; build the dashboards and alerting that make the health of the system visible to everyone, not just you.
- Lead incident response: triage fast, communicate clearly, coordinate the team, contain blast radius. Run blameless postmortems that actually result in durable fixes and systemic improvements, not just tickets.
- Proactively identify reliability risks before they become incidents, through capacity planning, load testing, chaos engineering, and systematic review of production signals.
- Build runbooks and on-call practices that make rotating on-call humane and effective, not a rotation people dread.
Build and Operate Infrastructure
- Design, implement, and maintain highly available, scalable, and secure cloud infrastructure on Google Cloud Platform: GKE, Cloud Run, AlloyDB, Cloud Storage, Pub/Sub, and adjacent services.
- Own Infrastructure as Code across all environments with a cloud-agnostic, open-minded, code-centric approach, applying standard software engineering discipline (version control, code review, testing, promotion gates).
- Build and maintain CI/CD pipelines (GitHub Actions) that make deployments fast, safe, and boring, with rollback, canary, and feature toggling patterns baked in.
- Manage observability infrastructure end-to-end: instrumentation, tracing, log pipelines, alerting, and dashboards structured around what actually matters operationally.
- Own data infrastructure health across GCP data systems, including AlloyDB, Cloud Storage, BigQuery, and Cloud Logging: performance tuning, query analysis, storage lifecycle management, connection management, backup/restore, and disaster recovery.
- Own AI and LLM inference infrastructure end-to-end: Vertex AI, Model Garden, and model serving platforms (vLLM, Cloud Run, GPU provisioning), latency SLOs, cost optimization, auto-scaling characteristics, and capacity planning for production AI workloads.
Set the Standard on Security and Compliance
- Maintain and improve our security posture across all infrastructure: network segmentation, IAM, secrets management, vulnerability management, audit logging.
- Own infrastructure-level compliance for HIPAA and SOC 2, not just checkbox compliance, but genuinely secure and auditable systems.
- Be a thought partner to engineering on security architecture decisions, particularly where AI workloads introduce new data handling surface area.
Patrol Adjacent Territory
- Cover integration infrastructure where it intersects with systems: EHR integrations, third-party APIs, data pipelines, and AI inference serving. You don't need to own these, but you understand them well enough to step in.
- Support backend engineers on runtime performance: Python async, FastAPI/Django serving characteristics, concurrency patterns, and where infrastructure choices constrain application-level performance.
What You'll Bring
- 8+ years of experience in SRE, infrastructure engineering, or DevOps, with a clear track record of owning reliability, not just contributing to it.
- You think in SLOs. You have defined them, defended them, and used error budgets to make real decisions about when to slow down feature work.
- Deep experience with incident management: you've led responses to serious production incidents, written the postmortems, and driven the follow-through.
- Expert-level command of GCP (GKE, Cloud Run, Cloud SQL/AlloyDB, Pub/Sub, IAM, VPC, Cloud Armor) or directly comparable experience on AWS or Azure with a credible GCP migration story.
- Very strong shell scripting skills (Bash) and fluency in Python for automation; comfortable reading application code well enough to understand its operational behavior.
Depth That Matters
- Observability done right: you've built meaningful monitoring from the ground up, not just enabled a tool. You know the difference between dashboards that look good and dashboards that catch problems.
- Deep knowledge of GCP data infrastructure and database operational characteristics: query planning and tuning in PostgreSQL/AlloyDB/BigQuery, log management and retention in Cloud Logging, data lifecycle management in Cloud Storage, analytical workloads in BigQuery, and how to diagnose issues across these services as they scale.
- Understanding of container and Kubernetes internals well beyond YAML: scheduling, resource management, network policies, pod disruption budgets, autoscaling behavior.
- Security and compliance grounding: you've operated in regulated environments (HIPAA, SOC 2, or similar) and understand the difference between compliance theater and systems that are actually secure.
- Deep operational mastery of AI and LLM inference serving, including Vertex AI and Model Garden: GPU provisioning and orchestration, model deployment strategies, optimizing inference latency and throughput, GPU cost controls, and managing the unique reliability patterns of production LLM workloads at scale.
Mindset
- You have a strong operational instinct. You notice anomalies, you ask why things are trending the way they are, and you're not satisfied until you understand the root cause.
- You treat reliability as a product you own, not a set of tasks assigned to you.
- You're calm under pressure, structured in how you communicate during incidents, and rigorous in follow-through after them.
- You're senior enough to have strong opinions about how things should be done and humble enough to change them when presented with better evidence.
- You use AI tools fluently as a practitioner, writing infrastructure code, debugging incidents, drafting runbooks, with the judgment to know when the output is right and when it needs correction.
Bonus
- Familiarity with healthcare data standards: FHIR, HL7, SMART on FHIR application architecture.
- Experience operating infrastructure for LLM or AI inference workloads at production scale.
- Background in building internal developer platforms or golden-path tooling that makes other engineers more effective.
- Strong hands-on experience or familiarity with Infrastructure as Code tooling such as Kubernetes, Terraform, or Pulumi.
- Experience with OpenTelemetry for instrumentation and distributed tracing across complex environments.
Who You'll Work With
You'll join a small, highly effective, self-driven engineering team where resourcefulness is prized and engineers operate with a high degree of autonomy. You will serve as a peer to the CTO and technical leadership—not as a support function, but as a co-owner of how the platform runs. You'll work closely with backend, data, and AI/ML engineers on the systems they build and depend on, and directly with leadership on operational priorities and incident response. The team is concentrated in the Bay Area with a weekly San Francisco office day; remote candidates are welcome, though you must be able to work core Pacific hours and must be based in the US.
Compensation: $200k – $250k + equity
We're a team on a mission to elevate the quality of care for every patient, by empowering clinicians with the right information at the speed of thought. Learn more about us.
#J-18808-Ljbffr Jobleads-US- Staff Engineering Program Manager, Core Tech Our mission at Oura is to... ...initiatives across multiple system technologies leveraging... ...This role will be located in San Francisco, California with in-office... ...recruiters, especially for remote roles. Please note: Our jobs...Remote workWork at officeLocal areaFlexible hours
$166k - $343k
Pre-Sales Systems Engineer - Higher Education and State Government (Sacramento/San Francisco, CA)This role has been designated as ‘Remote/Teleworker’, which means you will primarily work from... ...immediately.SummaryLocation: All, California, United States of America;...Remote workFull timeWork experience placementLocal areaImmediate startWork from home$166k - $343k
...Pre-Sales Systems Engineer - Higher Education and State Government (Sacramento/San Francisco, CA)This role has been designated as ‘Remote/Teleworker’, which means you will primarily work from... ...Salary USD 166,000 - 343,000 in California This range reflects the minimum...Remote workWork experience placementWork from home$225k - $350k
# Staff AI EngineerFast-growing software company* LocationSan Francisco, California* TypeFull-time* Posted6 October 2026Apply... ...and ship software systems layered on top of... ...a founder or early engineer.## Benefits* Health,... ...Location and work model* San Francisco,...SuggestedFull timeWork at officeFlexible hours$130k - $140k
...residential experience for entrepreneurs in San Francisco. They are looking for a proactive, high-... ..., offering more flexibility to work remotely and/or take time off 5+ years of experience... ...periods San Francisco, CA (Mainly onsite - with flexibility to WFH depending on...Remote workWork from homeFlexible hoursShift workAfternoon shift- ...edge AI startup in San Francisco to hire a Full Stack AI Engineer with strong zero-to... ...experience.This is a fully onsite position.... ...if you are seeking remote or hybrid work.About... ...infrastructure, and AI systems. You should enjoy... ...Francisco, California - 0Physical Therapist...Remote workFull timeLive inWork at officeLocal areaRelocation
$101.25k - $126.56k
...Technical Consultant I#### San Francisco, California, United States**Why join us... ...four weeks per year of fully remote work!**Responsibilities***... ...integration of customer systems with Brex, managing the end... ...insights for our Product and Engineering partners.* Actively contribute...Remote workWork at officeImmediate startWork from home$320k - $405k
...Staff / Senior Physical Security Systems Engineer Remote-Friendly (Travel Required) | San Francisco, CA About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and...Remote jobFor contractorsWork at officeVisa sponsorshipFlexible hours- ...talent-first flexible/remote work approach, see... ...direction for our evaluation systems; metric, judge and... ...alignment across engineering, product and AI leaders... ...accordance with the San Francisco Fair Chance Ordinance... ...Chance Ordinance, the California Fair Chance Act, and...Remote jobFull timeLocal areaFlexible hours
- ...Manager - Analytical Instrumentation San Francisco Area (Remote) Swan Analytical USA Full-... ...our presence throughout the Northern California region. This is an excellent... ...For Bachelor's degree preferred (Engineering, Chemistry, Environmental Science, Business...Remote workBase plus commissionFull timeWork at officeWorldwideMonday to Friday
$50 per hour
...may also elect to schedule in-person clinic sessions at our San Francisco location by selecting from available clinic blocks, at their... ...ability to obtain additional state licenses as needed Valid California DEA Registration in good standing with all schedules to prescribe...Remote workHourly payLive inAfternoon shift- ...RoleWe are seeking a highly motivated Systems Engineer to help shape the architecture and integration... ...Applicants for Jobs Located in NYC or Remote Jobs Associated With Office in NYC... ...of non-discrimination.Pursuant to the San Francisco Fair Chance Ordinance, Los Angeles...Remote workHourly payWork at officeLocal areaFlexible hours
$166k - $343k
Presales Systems Engineer - HPE Networking (Northern California)This role has been designated as ‘Remote/Teleworker’, which means you will primarily work from home.Who We Are:Hewlett... ...States of America; Oakland, California; San Francisco, California, United States of America;...Remote workFull timeWork experience placementLocal areaImmediate startWork from home$114.8k - $165.2k
...are looking for a Staff Reliability Engineer to join our team in... ...position will be fully Onsite based in San Jose, CA Role... ...Bloom Energy fuel cell systems. Work with others... ...Field Service and Remote Monitoring to... ...headquartered in San Jose, California, is growing quickly...Remote workWork experience placementWorldwideShift work- ...all of their business systems through natural language... ...Moveworks’ Reasoning Engine and natural language... ...all started in sunny San Diego, California in 2004 when a visionary... ...learning systems. The ML infra team covers a variety... ...personas (flexible, remote, or required in office...Remote workPermanent employmentWork at officeFlexible hours
$122.6k - $183.95k
Job SummaryThe Sr. Staff OT Systems Engineer is a senior technical leadership role responsible for designing, implementing, integrating, and supporting... ..., including segmentation, switching, routing, firewalls, remote access, servers, virtualization, identity management,...Remote workFull timeWork at office$160k - $200k
...autonomous logistics system, delivering critical supplies... ...issues requires an engineer who can move across... ...telemetry, fault codes, and remote-debug capabilities so... ...-cause methods.At the Staff level, lead cross-... ...role based in South San Francisco and requires domestic...Remote workPermanent employmentLocal area- ...focused practice in the heart of San Francisco. This is a great opportunity... ...re a licensed Optometrist in California with the required liability... ...For roles that are remote (i.e., Work From Home (WFH)) or hybrid (i.e., partial onsite at a VSP location and WFH), must...Remote workFull timeLocal areaWork from homeFlexible hours
- ...from leading investors. We are based in San Francisco and work in person. The Role You... ...OTA updates, logging, monitoring, and remote diagnostics for the fleet Profile and... ...Experience with real-time or embedded Linux systems Familiarity with ROS 2 or similar...Remote work
$114k - $182k
Job TitleSenior/Staff Ultrasound System Engineer - Medical Device (San Diego, CA)Job DescriptionSenior/Staff Ultrasound System Engineer - Medical Device (San Diego... ...means working in-person at least 3 days per week. Onsite roles require full-time presence in the company’s...Full timeWork at officeImmediate startWork visaRelocation package3 days per week$270k - $290k
...training, and deploying AI systems—designed to take ideas from... ...with offices in New York City, San Francisco, Seattle, and London, and is... ...NYC office, we are open to remote US candidates for this role.... ...Leadership with Product Management, Engineering, Revenue Teams, and within...Remote workWork at officeLocal areaWork from homeFlexible hours$297k - $350k
...quality. We are looking for a Senior Staff Software Engineer to set technical direction across... ...spans multiple engineering teams across San Francisco, New York, Vancouver, and Warsaw.... ...Org that includes AI platform, Agent systems and Studio in partnership with engineering...Work at officeLocal areaWork from homeWorldwideFlexible hoursShift work$68.9k - $131.1k
...Date: 2026-10-06Location: US-CA-SAN DIEGO-SD1 ~ 8650 Balboa Ave ~... ...BLDG Position Role Type:Onsite U.S. Citizen, U.S. Person, or... ...at RTX.What You Will DoAs a Systems Engineer at Raytheon, you will integrate... ...as on-site, hybrid or remote.The salary range for this role...Remote workFull timeTemporary workWork experience placementInternshipWork at officeRelocationFlexible hours$234.84k
...global Site Reliability Engineering (SRE) function to... ...uses code, data, and systems thinking to make product... ...safe and fast. As a Staff Site Reliability... ...TTC / OTE) Zone 1: San Francisco Bay Area, New York City... ...Zone 2: Washington, California (excluding San Francisco...Remote workBase plus commissionLocal areaWorldwideFlexible hoursShift work$83k - $108k
...looking for a Key Account Manager for our San Francisco, California area. In this role, you will be... ...classes and so much more! Location: Remote - US Actual compensation may vary depending... ...have any difficulty using our online system. If you need such an accommodation,...Remote workFull timeTemporary workWork experience placementWork at officeFlexible hours- ...of all of us—from design and engineering to the manufacturing and... ...RoleWe’re looking for a Sr. Staff Systems Engineer with a strong systems... ...teams in Ireland, including onsite work to support system integration... ...and work-life balance. Remote or field-based positions will...Remote workFull timeWork at office
$125.7k - $220k
...DescriptionIt all started when engineer Fred Luddy wrote code that... ...screening.About the Role:The systems we build run on a variety of cloud... ...below:Cloud (Azure, AWS, or GCP)Infra as code (Terraform, cloud... ...trust. Work personas (flexible, remote, or required in office) are categories...Remote workPermanent employmentWork at officeImmediate startFlexible hours$7.86k - $10.54k
...problem-solving, system reliability, and helping... ...up to two days of remote work and three days or more onsite, per week.... ...July 1, 2025, The California Department of Human... ...Title: Systems Engineer Classification... ...the Napa Valley, San Francisco, Lake Tahoe, and...Remote workPermanent employmentFull timeWork at officeMonday to FridayFlexible hours$177k - $265.6k
...opportunities to work on revolutionary systems that impact people's lives... ...and Verification - Staff Systems Engineer to join our team of diverse... ...role can be located in either San Diego , El Segundo , or... ...our facility. There is no remote / hybrid / telework available...Remote workFull timeRelocation packageShift work$104.9k - $194.7k
...DescriptionPerforms technical planning, system integration, verification and validation,... ...perform IT responsibilities, use E146, Systems Engineer.Basic Qualifications• Active Secret... ...software, data, and procedures.• Full-time onsite role in Orlando, FL, working the majority...Full timeTemporary workWork experience placementCasual workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Infra Staff Systems Engineer Onsite or Remote San Francisco, California. Be the first to apply!
- assistant engineer San Francisco, CA
- senior staff systems engineer San Francisco, CA
- technology administrator San Francisco, CA
- engineering aide San Francisco, CA
- senior staff engineer San Francisco, CA
- staff design engineer San Francisco, CA
- assistant electrical engineer San Francisco, CA
- staff data engineer San Francisco, CA
- software engineer staff San Francisco, CA
- staff engineer San Francisco, CA



