Member of Technical Staff — Inference Infrastructure
Uncover
Our mission is general causal intelligence; AI that is capable of (1) predicting the future and (2) identifying the actions to alter it. To achieve this breakthrough, we are building a Large Physics foundation Model (LPM) because physical systems, unlike text or images, are governed by verifiable cause and effect. We believe that scaling on physics will enable an understanding of causality required to predict and control physical systems, starting with weather. Our founding team has built and deployed AI against the physical world in robotics, drug discovery, and particle physics at institutions like DeepMind, Waymo, Cruise, Insitro, Nabla Bio, and CERN. Responsibilities Your mission is to make inference so fast and cheap that evaluation never gates research. Build high-throughput inference systems for large-scale evaluation, backtesting, and scoring against historical physical observations Design and implement techniques that improve latency, throughput, and efficiency for real-time inference Optimize the inference stack to fully utilize hardware FLOPs, bandwidth, and memory Extend orchestration frameworks (e.g. Kubernetes, Ray, Slurm) for distributed inference and large-batch evaluation sweeps Establish standards for reliability, observability, and reproducibility across the inference stack, so every evaluation is trustworthy and repeatable Collaborate with researchers to enable high-performance inference for novel architectures as they emerge What we're looking for We value a relentless approach to problem-solving, rapid execution, and the ability to quickly learn in unfamiliar domains. Experience building or optimizing inference and serving systems for throughput and latency (e.g. TensorRT) Understanding of distributed compute, GPU parallelism, and hardware-aware optimization Deep familiarity with deep learning frameworks (e.g. PyTorch, JAX) and their underlying system architectures Strong engineering skills: performant, maintainable code and the ability to debug complex codebases Bonus: contributions to open-source inference or systems infrastructure (e.g. vLLM, SGLang, Triton) #J-18808-Ljbffr Uncover
- ...designing, building, and scaling core infrastructure that powers a high-volume data platform... ...applications. We are looking for team members who love building enabling systems that... ...applications, including model integrations, inference workloads, evaluation pipelines,...SuggestedWork at office
$150k - $300k
...spanning cloud LLM serving, LLM inference optimization and RL systems.... ...areas are: Building the infrastructure to serve LLMs efficiently at... ...RL training stack. Core Technical Responsibilities LLM Serving... ...development and encourage team members to contribute to the broader...SuggestedWork at officeRemote workVisa sponsorshipRelocation packageFlexible hoursShift work$225k
Member of Technical Staff, Inference & RL Systems Magic’s mission is to build safe AGI that accelerates humanity’s progress on the world’s most important... ...the boundary between model execution and distributed infrastructure. You will work on systems that determine inference...SuggestedRelocationVisa sponsorship- ...recognize parts of inputs that are unimportant, reducing inference costs for scale-ups and enterprises that integrate... ...is 5 people with a research and product focus. As a Member of Technical Staff on our infrastructure team, you'll own the cloud systems that serve our...SuggestedVisa sponsorship
- ...AI Models, software, and infrastructure. We are recognized as the... ...investors, distils our deep technical research and knowledge into... .... Position Overview Member of Technical Staff will play a crucial role in developing training & inference benchmarks & system modelling...SuggestedFull timeWork at officeRemote workWorldwide
$250k
...building the next generation of agentic infrastructure for GPU-intensive workloads.... ...opportunity offers the chance to join as a Member of Technical Staff at a pivotal stage in the company's... ...on orchestration intelligence, inference gateways, and agentic operations tooling...Full time- Member of Technical Staff - ML Systems & Inference Bay Area, CA | Onsite Join a well-funded AI infrastructure startup building the orchestration layer for next-generation AI workloads This role sits at the intersection of ML systems, inference, distributed systems, and...
- # Founding Member of Technical Staff, AI Infrastructure**Location:** San Francisco / Bay Area preferred. Remote exceptional for the right person.We look... ...AI workloads cheaper and easier to own by turning inference behavior, traces, workload replay, GPU signals, and task...Full timeRemote work
$200k - $400k
...to an algorithm. We're building the infrastructure to understand human behavior at scale... ...simulating a society means running inference over populations of agents, not single... ...to run. About the Role As a Member of Technical Staff in Research Infrastructure, you will...Live inFlexible hours$200k
Member of Technical Staff, Supercomputing Platform & Infrastructure Magic’s mission is to build safe AGI that accelerates humanity’s progress on the world’s most important... ..., domain-specific RL, ultra-long context, and inference-time compute to achieve this goal. About the...RelocationVisa sponsorship- ...with today’s homogeneous, vertically integrated infrastructure. Gimlet addresses this by decoupling AI workloads... ...datacenters. Mission Gimlet Labs is seeking a Member of Technical Staff focused on ML systems and inference. In this role, you will design and build inference...
- ...About Us Gimlet is building the next generation of AI infrastructure: large-scale AI datacenters and the orchestration platform that coordinates... ...the cluster infrastructure behind Gimlet’s heterogeneous inference cloud. Unlike traditional cloud platforms built around a...
$150k - $300k
Building Open Superintelligence Infrastructure Prime Intellect is building the open superintelligence... ...for GPU Infrastructure, you'll be the technical expert who transforms customer... ...deployment strategies for LLM training, inference, and HPC workloads Present architectural...- Perplexity is looking for a technical program manager to be the connective tissue between... ..., and product teams, driving our core inference platform forward. Perplexity runs one of... ..., across performance engineering, infrastructure, and product teams Lead cross-functional...Shift work
$300k
...About the role This role is for strong infrastructure engineers who can build the systems layer... ...rollouts, training orchestration, inference, evals, data pipelines, observability,... ..., platforms, or services used by other technical users. Strong judgment around technical...Work at officeLocal area$275k - $315k
...amount of untrusted, freshly generated kernels on real silicon, quickly and safely, which is why we're hiring a Member of Technical Staff for Sandbox Infrastructure to build the layer that makes it possible: a serverless GPU container service across NVIDIA, AMD, TPU and...Full timeWork at officeRelocationRelocation package$200k - $400k
About The Role We're looking for an inference runtime engineer to push the boundaries of... ...Contributions to open-source ML or system infrastructure projects. Bonus points if you have:... ...LlamaFactory, etc). Written widely-shared technical blogs or side projects on vLLM or LLM...Remote workVisa sponsorshipShift work- Job Description - Member of Technical Staff (Inference) Location: San Francisco (on-site at our offices) About Artificial Analysis Artificial Analysis is the leading independent AI benchmarking company. We support labs, engineers and enterprises to understand AI capabilities...Shift work
$250k - $300k
...development? Join one of the most exciting AI infrastructure companies in the market, building a... ...next-generation AI training and inference at scale. This role offers the opportunity... ...Have: ~ A track record of impressive technical work you can speak to in depth, the...Full timeRemote work- About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure... ..., and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at...Work at office
- ...of AI. Our mission: make intelligence open and accessible to all. Role Overview Reflection.AI is looking for a Member of Technical Staff - Infrastructure Security to secure our geographically diverse multi‑cloud Kubernetes and cloud environments. In this role, you’ll...Work at officeVisa sponsorship
- Member of Technical Staff - Infrastructure Security We're partnering with a frontier AI research company that is building next-generation open-weight foundation models with the mission of making advanced AI broadly accessible. Their team includes researchers, engineers...
$180k - $250k
Job Title Member of Technical Staff, Backend Salary $180k-$250k + Equity Company Description Well-funded AI infrastructure startup Job Description Join a high-growth team building the core infrastructure powering how modern AI applications process and understand documents...- ...power our diffusion LLMs in production. Your work will make inference faster, more cost-effective, and more reliable. Key Responsibilities... ...decoding, continuous batching). Knowledge of ML-specific infrastructure challenges (checkpointing, resource scheduling, etc.). #J-18...
- ...Pixeltable Inc. Member of Technical Staff San Francisco, CA·Full time Apply for Member of Technical... ...to focus on innovation, not on infrastructure. We aim to simplify the AI development... ...transformation, training/fine-tuning, and inference? You will also: Find opportunities...Full timePart timeWork at officeWork from homeFlexible hours2 days per week
$150k - $350k
...Member of Technical Staff | Distributed Systems San Francisco - Onsite $150k-$350k base + equity I'm working with a small, deeply technical $80 million Series A AI infrastructure company , building an inference cloud for agentic workloads that can partition and orchestrate...$150k - $350k
...Member of Technical Staff - Distributed Systems San Francisco, CA- 5 days per week onsite $150,000–... ...Opportunity Join a rapidly growing AI infrastructure company building a multi-silicon cloud platform for fast, efficient inference. The future of AI inference will not...- ...people take ownership, grow together, and share both the challenges and the wins. What You'll Do Build the supercomputing infrastructure that runs our agents. Our agents tackle long-horizon, high-performance workloads, and you'll design the cloud compute,...Work at officeRemote workFlexible hours
- ...Job Title Member of Technical Staff: Infrastructure Salary Not Disclosed Company Description Observable Intuition is an early‑stage AI infrastructure... ...first infrastructure hire, you will build a production inference platform from the ground up. You’ll design portable, multi...
- ...AI Models, software, and infrastructure. We are recognized as the... ...investors, distils our deep technical research and knowledge into... ...for a highly motivated member of technical staff to join our engineering team... ...both frontier LLM training & inference models Implement modern...Full timeWork at officeRemote workWorldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Member of Technical Staff — Inference Infrastructure. Be the first to apply!
- senior technical analyst San Francisco, CA
- systems support technician San Francisco, CA
- senior help desk analyst San Francisco, CA
- help desk technical support San Francisco, CA
- trade support analyst San Francisco, CA
- support analyst San Francisco, CA
- technical analyst San Francisco, CA
- technical support assistant San Francisco, CA
- help desk assistant San Francisco, CA
- IT assistant San Francisco, CA


