Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Inference Systems Engineer — High-Throughput AI

Kindredventures

Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high-throughput systems, latency-optimized serving, and distributed orchestration across Kubernetes, Ray, and Slurm. Ideal candidates have a track record in building efficient inference stacks, GPU-aware optimization, and deep learning frameworks like PyTorch or JAX. Collaboration with researchers and a focus on reliability are essential. #J-18808-Ljbffr Kindredventures

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff Inference Systems Engineer — High-Throughput AI in San Francisco, CA vacancy
  • Causal Labs in San Francisco is seeking an experienced Infrastructure Engineer to build high-throughput inference systems for large-scale evaluation and backtesting against historical physical observations. You will design techniques to improve latency and throughput, optimize... 
    Suggested

    Causal Labs

    San Francisco, CA
    13 hours ago
  • Sail Research in San Francisco is seeking a talented engineer to design and implement robust systems that ensure fast and cost-efficient AI inference at global scale. You will be responsible for building high-performance schedulers and optimizing global routing while focusing... 
    Suggested

    Sail Research

    San Francisco, CA
    2 days ago
  • A tech startup focused on AI workloads is seeking a Member of Technical Staff to design and optimize inference systems. The role involves managing KV cache allocation and improving...  ...candidates should have strong software engineering skills and experience with ML inference... 
    Suggested

    Gimlet Labs

    San Francisco, CA
    13 hours ago
  • Crusoe in San Francisco is seeking a Staff Technical Program Manager to lead the Managed Inference platform team. You will ensure end-to-end program delivery for LLM workloads, driving innovation in AI infrastructure powered by clean energy. The ideal candidate has extensive... 
    Suggested

    Crusoe

    San Francisco, CA
    1 day ago
  •  ...the da Vinci surgical system and Ion—have transformed...  ...worldwide.We’re a team of engineers, clinicians, and...  ...Systems GPU Engineer - AI & Robotics, you will be...  ...improving and integrating high performance robotic AI...  ...models to real-time onboard inference—while serving as a core... 
    Suggested
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    San Francisco, CA
    23 hours ago
  •  ...Responsibilities As a senior Machine Learning Systems Engineer on the Search Platform team, you will...  .... Contribute to the architecture of high-throughput, low-latency search systems that meet...  ...workflows. Partner with Rovo and AI platform teams to evolve search infrastructure... 
    Work at office
    Local area

    Atlassian

    San Francisco, CA
    1 day ago
  • $227.2k - $284k

    Scale's Physical AI business unit is dedicated to solving the...  ...Physical AI.The RoleAs an ML Systems Engineer on the Physical AI team, you...  ...system design. You’ll work in a highly collaborative environment,...  ...performance tracking of model inference.Lead: Own projects end-to-end... 
    Full time

    Scale AI

    San Francisco, CA
    23 hours ago
  • OpenAI in San Francisco is seeking an experienced systems generalist to build an automated inference optimization platform across hardware, compiler, and runtime contexts. You will design the OpenAI-hosted control plane and partner-side software, focusing on reliable long... 

    Slope

    San Francisco, CA
    4 days ago
  •  ...powers mission-critical inference for the world's most dynamic AI companies, like...  ...build the platform engineers turn to to ship AI...  ...the global operating system for distributed,...  ...converging. The massive throughput of H100, B200, and...  ...experience with high-performance networking... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    23 hours ago
  • Senior ML Systems Engineer, Frameworks & Tooling at Cohere Our mission is...  ...enterprises who are building AI systems to power magical experiences...  ...). Improve training throughput and stability on multi-node clusters...  ...configurations support high-performance training. Investigate... 
    Full time
    Work at office
    Remote work
    Flexible hours

    Cohere

    San Francisco, CA
    13 hours ago
  • $200.8k - $251k

    A leading AI technology company in San Francisco seeks a team member to build and optimize a machine learning...  ...for large language models. Candidates should have system optimization experience and solid software engineering skills, particularly in tools like CUDA and Pytorch... 
    Full time

    Scale AI

    San Francisco, CA
    1 day ago
  •  ...building the observability layer for AI-assisted software development, measuring how much of an engineering org's code is actually written...  ...What You'll DoBuild the core systems that detect and attribute AI-...  ...to the userHigh-agency, high-responsibility individuals with... 
    Work at office
    Local area

    Atlassian

    San Francisco, CA
    4 days ago
  • $124k - $280k

     ...people in data and analytics engineering focus on leveraging advanced technologies...  ...and implementing advanced AI and ML solutions to drive...  ...algorithms, models, and systems to enable intelligent decision...  ...ability to develop and sustain high performing, diverse, and inclusive... 
    Full time
    H1b

    PwC

    San Francisco, CA
    4 days ago
  • $293k - $385k

     ...operating deeply technical systems that must work reliably...  ...is seeking a Security Engineer, Host Assurance to help...  ...excited by ambiguous, high-impact problems at the...  ...preserving deployment throughput and operational reliability...  ...OpenAIOpenAI is an AI research and deployment... 
    Work at office
    Local area
    Relocation package
    Flexible hours

    OpenAI

    San Francisco, CA
    4 days ago
  • $190k - $250k

     ...flagship product—an AI-driven, non-...  ...patients worldwide. As a Staff Application...  ...writing and tuning high-performance queries...  ...Own OLTP Database Systems: Take hands-on ownership...  ...— for high-throughput, low-latency web services...  ...; partner with engineering and SRE teams to... 
    Work experience placement
    Local area
    Worldwide
    Relocation

    HeartFlow

    San Francisco, CA
    1 day ago
  • $264.8k - $331k

     ...AI is becoming vitally important in every function...  ...As an ML Sys Research Engineer, you'll work on building...  ...to optimize our ML system. Your customer will be...  ...optimize our training and inference framework. Post-train...  ...products provide the high-quality data and full-stack... 
    Full time

    Scale AI

    San Francisco, CA
    4 days ago
  •  ...AI Systems EngineerTransluce is a fast-moving research lab building...  ...for an exceptional AI systems engineer to lead the design and development...  .... As an early member of a highly collaborative team, you will...  ...checkpointsInterpretability: Inference stacks that are as performant... 
    Flexible hours

    Transluce

    San Francisco, CA
    1 day ago
  • $272k - $336k

     ...Waymo Systems Engineering RoleWaymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since...  ...and hardware systems in groundbreaking new ways. We set the high performance standards that ensure our vehicles run smoothly and... 
    Odd job
    Full time
    Remote work

    Latent Logic

    San Francisco, CA
    1 day ago
  • $165k - $206k

     ...and we are hiring the world’s best engineers, scientists, designers, product managers...  ...AND RESPONSIBILITIES:The Enterprise Systems Engineer who leads with AI - not as a tool on the side, but as...  ...manual authoring time.Produce high-fidelity technical documentation using... 

    Juul

    San Francisco, CA
    2 days ago
  • $145k - $195k

    San Francisco, CAProduct Systems Engineering - Product Systems /Full time /On-siteWanna join the adventure...  ...across the Hub product scopeTranslate high-level concepts into actionable technical...  ...observation, IoT connectivity, on-orbit AI, national security missions, and more.... 
    Full time
    Temporary work

    Loft Orbital

    San Francisco, CA
    4 days ago
  •  ...superintelligence is an AI that uses quantum...  ...-qubit systems, turning raw cryogenic...  ...the world's software engineers. AI is already generating...  ...of Technical Staff you will shape Conductor...  ..., labelling, and inference. Integrate with external...  ...We Are Looking For High intelligence, high... 

    SwiftCruit

    San Francisco, CA
    1 day ago
  •  ...About the Team: The Database Systems team specializes in high-performance distributed...  ...: We are looking for engineers passionate about distributed...  ...database reliability and throughput as usage grows by orders of...  ...OpenAI OpenAI is an AI research and deployment company... 
    Full time

    OpenAI

    San Francisco, CA
    23 hours ago
  • $264.8k - $331k

    Machine Learning Systems Research Engineer, Agent Post-training - Enterprise GenAI AI is becoming vitally important in every function...  ...and optimize our training and inference framework. Post-train state of...  ...decisions. Our products provide the high-quality data and full-stack... 
    Full time
    Contract work
    For contractors
    For subcontractor
    Work at office

    Scale LLP

    San Francisco, CA
    1 day ago
  • $224.5k - $251.5k

     ...DialpadDialpad is the AI platform for...  ...advantage.Unlike legacy systems built to route and...  ...themselves to a high bar. Our ambition...  ...an AI-native engineering culture.This position...  ...agent frameworks, LLM inference optimization,...  ...leadership (as a Staff, Senior Staff, or... 
    Work at office

    Dialpad

    San Francisco, CA
    4 days ago
  •  ...Responsibilities As a Principal Machine Learning Systems Engineer on the Search Platform team, you set the...  ..., Confluence, and the broader Atlassian AI platform. You operate with broad...  ...alike.Leading Without AuthorityDrive high-impact initiatives across teams that do... 
    Work at office
    Local area
    Shift work

    Atlassian

    San Francisco, CA
    2 days ago
  • $102.3k - $161.76k

     ...matters and so do you. About the Role As a Staff AI FinOps Governance Lead, you will lead...  .... You will partner closely with Engineering, Product, Finance, Procurement, and Cloud...  ...drivers of Generative AI, including LLMs, inference services, vector databases, GPUs, and AI... 

    Ultimate Software

    San Francisco, CA
    4 days ago
  • $204k - $300k

     ...future. At Dolby, science meets art, and high tech means more than computer code. As...  ...to computer science and electrical engineering, such as AI/ML, algorithms, digital signal processing...  ...data science & analytics, distributed systems, cloud, edge & mobile computing,... 
    Full time
    Local area
    Worldwide
    Flexible hours

    Dolby

    San Francisco, CA
    2 days ago
  • $250k - $300k

     ...the only vertically integrated AI infrastructure company built...  ...strategies, and be part of a high-performing team that believes...  ...RoleCrusoe is looking for a Senior Staff Network Architect to help...  ...Network Architect and network engineering leadership to design and operate... 
    Temporary work
    Work at office

    Crusoe

    San Francisco, CA
    4 days ago
  •  ...capital that fragmented legacy systems can't provide.We are...  ...infrastructure underneath them. You're an engineer first: you write clean, fast...  ...'ve retooled ourselves around AI to clear out the busy work,...  ...feels off in a UI and hold a high bar, even if you spend more time... 
    Flexible hours
    Shift work

    Ambrook

    San Francisco, CA
    6 hours ago
  • $189.6k - $237k

     ...language model training and inference. The platform has been...  ...heart of the field of AI as an indispensable...  ...Strong excitement about system optimizationExperience...  ...systemsStrong software engineering skills, proficient in frameworks...  ...products provide the high-quality data and full-... 
    Full time

    Scale AI

    San Francisco, CA
    23 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Inference Systems Engineer — High-Throughput AI. Be the first to apply!