Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff Engineer, AI Benchmarking & Inference Systems

S27a

SemiAnalysis is seeking a highly motivated Member of Technical Staff to join our growing engineering team. You will contribute to inference benchmarks, system modeling, and the preparation of technical newsletters with potential authorship recognition. We hire individuals at all experience levels and offer competitive compensation. As part of the process, you’ll complete a paid coding challenge reflecting daily tasks and collaborate with a global team across multiple AI and hardware domains. #J-18808-Ljbffr S27a

Vacancy posted 9 hours ago
Similar jobs that could be interesting for youBased on the Staff Engineer, AI Benchmarking & Inference Systems in San Francisco, CA vacancy
  • SemiAnalysis LLC is hiring a Member of Technical Staff to develop training and inference benchmarks and system modelling for AI hardware. You will contribute to writing our newsletter articles with potential authorship recognition and collaborate with a global team of... 
    Suggested

    S27a

    San Francisco, CA
    9 hours ago
  • Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high-throughput systems, latency-optimized serving, and distributed orchestration across Kubernetes, Ray, and Slurm... 
    Suggested

    Kindredventures

    San Francisco, CA
    3 days ago
  • Sail is hiring for an engineering role in San Francisco to design and implement high-performance...  ...fleet. You will also build LV routing systems to dispatch workloads with latency...  ...caching for memory/compute trade-offs in LLM inference stacks. You will contribute to deep... 
    Suggested

    Sail

    San Francisco, CA
    2 days ago
  • Sail Research in San Francisco is seeking a talented engineer to design and implement robust systems that ensure fast and cost-efficient AI inference at global scale. You will be responsible for building high-performance schedulers and optimizing global routing while focusing... 
    Suggested

    Sail Research

    San Francisco, CA
    3 days ago
  •  ...building the world's most efficient software for inference and agent hosting. In this role, you'll be one of the first engineers on Sailboxes, contributing across the stack—...  ...networking stack to building large-scale systems that maximize efficiency. You'll focus on distributed... 
    Suggested
    Work at office

    Sail Research Inc.

    San Francisco, CA
    19 hours ago
  • Together AI in San Francisco is seeking a Staff ML Systems Engineer to design and prototype algorithms, architectures, and scheduling for low-latency, high-throughput inference. You will implement changes in production-grade inference engines, including kernel backends... 

    Together

    San Francisco, CA
    2 days ago
  •  ...Francisco, CA, is seeking a Member of Technical Staff for distributed systems to design, build, and operate the platform that schedules AI workloads across thousands of nodes. This...  .... You will collaborate with founders and engineers from Nvidia, Google AI, Intel, and Pixie... 

    Acceler8 Talent

    San Francisco, CA
    2 days ago
  •  ...building the best way to talk to AI and humans together —...  .... Member of Technical Staff is the title we use for engineers who own hard problems end...  ...Representative projects Designing systems that give 3M+ AI agents...  ...fine-tuning, evaluation, inference, or RAG at scale High-... 

    Shapes, Inc

    San Francisco, CA
    1 day ago
  • B Capital is seeking a skilled engineer for GPU infrastructure in San Francisco. This...  ...designing and operating high-performance systems for model inference, synthetic data generation, and...  ...a passion for working in cutting-edge AI. Benefits include top-tier compensation... 

    B Capital

    San Francisco, CA
    4 days ago
  •  ...Francisco is hiring Members of Technical Staff to build systems that accelerate LLM inference and own customer workloads end to...  ...-performance kernels, inference engine internals, and production...  ...collaborating with a fast-growing AI inference company. #J-18808-Ljbffr... 

    Simplify

    San Francisco, CA
    1 day ago
  • Modal is building an infrastructure layer for AI workloads, covering training, deployment, observation, and inference. You will perform hands-on inference research, selecting...  ...work with customers alongside Forward Deployed Engineers to deploy and tune models, while expanding... 

    Mixpeek

    San Francisco, CA
    9 hours ago
  • $200k - $400k

    A leading AI technology company located in San Francisco is seeking an infrastructure engineer to build distributed systems for their AI inference engine. The role involves designing systems that ensure minimal latency and maximum reliability. Candidates should have a... 
    Visa sponsorship

    Inferact

    San Francisco, CA
    4 days ago
  • $190.9k - $232.8k

    A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine. The role requires...  ...in CUDA, GPU programming, and distributed systems design. Ideal candidates will have a strong... 

    Jobleads-US

    San Francisco, CA
    1 day ago
  • $225k

    Dormont Manufacturing Co is looking for a Software Engineer on the Inference & RL Systems team in San Francisco. The role involves designing distributed systems, optimizing performance, and ensuring high reliability for RL and post-training workflows. The ideal candidate... 

    Dormont Manufacturing Co

    San Francisco, CA
    2 days ago
  • $215k - $260k

     ...the only vertically integrated AI infrastructure company built...  .... That means owning the inference stack end to end: profiling where...  ...not good enough. This is core systems and performance work on some...  ...work directly with customer engineering teams to tailor deployments to... 
    Temporary work

    Crusoe

    San Francisco, CA
    4 days ago
  •  ...in San Francisco is seeking a Member of Technical Staff to lead distributed systems at the core of our AI infrastructure. You’ll design, build, and operate scheduling...  ..., GPUs, and accelerators. You’ll work closely with inference, runtimes, compilers, kernels, and hardware teams... 

    Acceler8 Talent

    San Francisco, CA
    2 days ago
  • $220k - $280k

     ...About the Role Together AI is building the best inference infrastructure for voice...  .... We're looking for a Staff ML Engineer to drive the model...  ...architect and implement systems targeting leading TTFB, throughput...  ...; establish the internal benchmark methodology that informs... 
    Full time

    Together Ai

    San Francisco, CA
    19 hours ago
  • $155k - $269k

     ...Description Job Description Waabi, founded by AI visionary Raquel Urtasun, is the leader...  ...team of Research Scientists and Engineers building the content backbone of a best-in...  ...scale, not just prototypes. You have shipped systems that other engineers and researchers... 
    Full time
    Work at office
    Work from home
    Flexible hours

    Waabi

    San Francisco, CA
    6 days ago
  • $150k - $225k

     ...with unprecedented speed and accuracy. Our AI-enabled platform turns siloed and...  ...team — which means you're not inheriting a system, you're designing one. You'll be the second...  ...identity, SaaS, and cloud platforms.The work is engineering-first: workflow automation, system... 
    Local area

    Peregrine Technologies

    San Francisco, CA
    2 days ago
  • $203.5k - $299.3k

     ...the next generation of causal decisioning systems for New Verticals: grocery, convenience,...  ...are hiring a Causal Machine Learning Engineer to help build the causal ML foundation behind...  ...…Deep practical experience with causal inference, econometrics, experimentation, or... 
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordash

    San Francisco, CA
    1 day ago
  •  ...the next generation of AI-driven game...  ...Senior Machine Learning Engineer for On-Device & Mobile...  ...significant parts of the inference stack — from a trained...  ...PIX, Instruments/Metal System Trace,Snapdragon Profiler...  ...and automated on-device benchmarking in CI.Research... 
    Full time
    Work at office
    Remote work
    Worldwide

    Unity Technologies

    San Francisco, CA
    2 days ago
  • Vals AI, Inc. is seeking exceptional engineers to own and scale our benchmarking platform. You’ll work across the stack—Python/Django backend, React/TypeScript frontend...  ...infrastructure—building reliable, high-performance systems for evaluating LLMs at scale. You will drive... 

    Vals AI, Inc.

    San Francisco, CA
    4 days ago
  • $220k

    We build and run the inference engine behind every Perplexity query and deploy dozens of model architectures at scale with tight latency and...  ...work (CUDA, Triton, CUTLASS, or similar). Any other deep systems programming experience is a plus. You understand modern LLM... 

    Perplexity

    San Francisco, CA
    3 days ago
  • $155k - $180k

     ...and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we...  ...us at Crusoe.About the Role:We are looking for a Staff Enterprise Technology Administrator, Engineering Systems to serve as the senior technology owner for Crusoe... 
    Temporary work

    Crusoe

    San Francisco, CA
    1 day ago
  • $300 per month

     ...only vertically integrated AI infrastructure company...  ...This Role:We are seeking a Staff Hardware Systems Engineer to strengthen Crusoe’s Hardware...  ...across training and inference - dense, MoE, long-context...  ....Experience with workload benchmarking, performance profiling, and... 
    Temporary work

    Crusoe

    San Francisco, CA
    1 day ago
  • $150k - $350k

     ...Inc. is seeking a Member of Technical Staff to focus on distributed systems in San Francisco, California. This...  ...and building the core platform for AI workloads, developing resource management...  ...should have strong software engineering fundamentals and experience with distributed... 

    Gimlet Labs, Inc.

    San Francisco, CA
    1 day ago
  • Lance is seeking a flexible, high-ownership engineer for our Members of Technical Staff role in San Francisco. You will own hard problems end-to-end, designing systems that span product, AI, infrastructure, and the physical world. You may work across frontend, backend,... 
    Flexible hours

    Lance

    San Francisco, CA
    1 day ago
  •  ...governing autonomous AI agents across industries...  ...We are seeking a Staff Research Engineer, AI/ML & Cybersecurity...  ...production-grade AI systems, secure model deployment...  ...components Improve inference performance, observability...  ...model evaluation, benchmarking, and stress-testing... 

    Ephapsys

    San Francisco, CA
    1 day ago
  •  ...in San Francisco is seeking a Member of Technical Staff to design and build distributed systems for AI workloads. The role involves developing scheduling...  ...APIs. Ideal candidates should have strong software engineering skills and experience with distributed systems. This... 

    Gimlet Labs

    San Francisco, CA
    1 day ago
  •  ...the next generation of causal decisioning systems for New Verticals: grocery, convenience,...  ...We are hiring a Causal Machine Learning Engineer to help build the causal ML foundation...  ...you Deep practical experience with causal inference, econometrics, experimentation, or... 
    Hourly pay
    Work at office
    Local area
    Flexible hours

    DoorDash, Inc.

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff Engineer, AI Benchmarking & Inference Systems. Be the first to apply!