Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff AI Inference Benchmarking Engineer

Artificial Analysis

Artificial Analysis, the leading AI benchmarking company, is hiring a Member of Technical Staff to own serverless inference coverage and drive performance benchmarks across provider ecosystems. You will extend methodologies, collaborate with neoclouds, and help shape how the market measures serving performance. You will work with engineers and leadership to direct the roadmap of the benchmarking platform, analyze throughput, time-to-first-token, and cost-per-token economics, and contribute to #J-18808-Ljbffr Artificial Analysis

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Staff AI Inference Benchmarking Engineer in San Francisco, CA vacancy
  • Artificial Analysis, Inc. is seeking a Member of Technical Staff (Hardware) to design and run AI hardware benchmarks for GPUs, TPUs and custom silicon in San Francisco...  .../data-analysis skills, and deep knowledge of inference economics. #J-18808-Ljbffr Artificial Analysis,... 
    Suggested

    Artificial Analysis, Inc.

    San Francisco, CA
    2 days ago
  • An innovative company is seeking a talented software engineer to join their dynamic Inference team. This role involves designing and implementing infrastructure...  ...researchers and product teams to push the boundaries of AI technology, ensuring reliable production services. If... 
    Suggested

    Jobleads-US

    San Francisco, CA
    3 days ago
  • Crusoe in San Francisco is seeking a Staff Technical Program Manager to lead the Managed Inference platform team. You will ensure end-to-end program delivery for LLM workloads, driving innovation in AI infrastructure powered by clean energy. The ideal candidate has extensive... 
    Suggested

    Crusoe

    San Francisco, CA
    2 days ago
  • Dynamo AI in San Francisco seeks an ML Engineer focused on evaluating LLMs, benchmarking, and ensuring safe, responsible AI in production. You will own end-to-end evaluation pipelines and contribute to synthetic data generation. You will collaborate with researchers to... 
    Suggested

    DynamoFL

    San Francisco, CA
    3 days ago
  • Anyscale is seeking a Distributed LLM Inference Engineer in San Francisco, California. This pivotal role involves pushing the boundaries of performance for ML inference at scale. You'll work closely with product teams to deliver end-to-end solutions while leveraging open... 
    Suggested

    Anyscale

    San Francisco, CA
    3 days ago
  • Causal Labs in San Francisco is seeking an experienced Infrastructure Engineer to build high-throughput inference systems for large-scale evaluation and backtesting against historical physical observations. You will design techniques to improve latency and throughput, optimize... 

    Causal Labs

    San Francisco, CA
    15 hours ago
  • Sail Research in San Francisco is seeking a talented engineer to design and implement robust systems that ensure fast and cost-efficient AI inference at global scale. You will be responsible for building high-performance schedulers and optimizing global routing while focusing... 

    Sail Research

    San Francisco, CA
    3 days ago
  • Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high-throughput systems, latency-optimized serving, and distributed orchestration across Kubernetes, Ray, and Slurm... 

    Kindredventures

    San Francisco, CA
    3 days ago
  • A tech startup focused on AI workloads is seeking a Member of Technical Staff to design and optimize inference systems. The role involves managing KV cache allocation and...  ...Ideal candidates should have strong software engineering skills and experience with ML inference... 

    Gimlet Labs

    San Francisco, CA
    15 hours ago
  • A leading technology company located in San Francisco is seeking a Machine Learning Research Engineer to design and develop safe AI benchmarking methodologies. This role involves collaboration with various teams to implement responsible evaluation techniques. Candidates... 

    Apple Inc.

    San Francisco, CA
    2 days ago
  • $315k

    A leading AI research company in San Francisco is seeking a mid-senior GPU Performance Engineer. In this role, you'll architect systems that enhance GPU performance for groundbreaking AI models. Responsibilities include developing optimizations, collaborating with teams... 

    Anthropic

    San Francisco, CA
    2 days ago
  • A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU kernels and using advanced techniques to boost computation efficiency. Ideal... 

    Baseten

    San Francisco, CA
    2 days ago
  •  ...team of YC and unicorn founders and senior engineers with deep expertise in 3D, generative...  ...'re looking for a Founding Engineer, ML Inference with deep expertise in high-performance...  ...while maintaining accuracy Profile and benchmark model performance to identify computational... 
    Relocation
    Visa sponsorship
    Relocation package

    Reactor.am

    San Francisco, CA
    4 days ago
  • $325k

    A leading AI research company in San Francisco seeks an engineer to optimize their powerful AI models for high-volume production environments. The ideal candidate has over 5 years of software engineering experience, strong familiarity with ML architectures, and experience... 

    Jobleads-US

    San Francisco, CA
    3 days ago
  •  ...a Tokens-as-a-Service (TaaS) Engineer to help build the systems that...  ...will work across performance benchmarking, tokenomics, model porting,...  ...Experience with GPU clusters, AI infrastructure, performance benchmarking...  ...with model porting, inference/training workloads, token... 
    Full time

    OpenAI

    San Francisco, CA
    1 day ago
  • Databricks Mosaic AI is seeking an experienced backend/infrastructure engineer to build platforms powering AI workloads, including model training, serving, and vector search. You will join a high-visibility team connected to research, product, and enterprise use cases,... 

    Databricks

    San Francisco, CA
    5 days ago
  •  ...Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence...  ...us and help build the platform engineers turn to to ship AI products....  .... Experience running low-level benchmarks to "qualify" new hardware clusters... 
    Full time
    Flexible hours

    Baseten

    San Francisco, CA
    1 day ago
  • OpenAI in San Francisco is seeking an experienced systems generalist to build an automated inference optimization platform across hardware, compiler, and runtime contexts. You will design the OpenAI-hosted control plane and partner-side software, focusing on reliable long... 

    Slope

    San Francisco, CA
    5 days ago
  • $170k - $245k

     ..., have Ray in their tech stacks to accelerate the progress of AI applications out into the real world.With Anyscale, we’re building...  ...50+ million raised to date.About the roleAs a Distributed LLM Inference Engineer, you will help systems and optimizations that push the... 
    Work at office

    Anyscale

    San Francisco, CA
    4 days ago
  • $160k - $230k

    About the RoleAt Together.ai, we are building state-of-the-art infrastructure to enable efficient and scalable inference for large language models (LLMs). Our mission is to optimize...  ...anInference Frameworks and Optimization Engineer to design, develop, and optimize... 
    Full time

    Together AI

    San Francisco, CA
    1 day ago
  • $102.3k - $161.76k

     ...matters and so do you. About the Role As a Staff AI FinOps Governance Lead, you will lead...  .... You will partner closely with Engineering, Product, Finance, Procurement, and Cloud...  ...drivers of Generative AI, including LLMs, inference services, vector databases, GPUs, and AI... 

    Ultimate Software

    San Francisco, CA
    4 days ago
  • $190k - $210k

     ...About The RoleSenior Developer Success Engineers (Sr. DSEs) own the post-sale technical relationship...  ...includes a competitive base salary benchmarked against real-time market data, as well...  ....We may use artificial intelligence (AI) tools to support parts of the hiring process... 
    Full time
    Work at office
    2 days per week

    Finch

    San Francisco, CA
    2 days ago
  • $216k - $270k

    About Scale AIScale AI is the data foundation for AI, helping organizations build and...  ...problems remains one of the hardest engineering challenges.As a Senior Frontier Agent Engineer...  ...evaluation frameworks using offline benchmarks, online A/B experiments, golden datasets... 
    Full time

    Scale AI

    San Francisco, CA
    2 days ago
  • $200.8k - $251k

    A leading AI technology company in San Francisco seeks a team member to build and optimize a machine learning framework for large...  ...should have system optimization experience and solid software engineering skills, particularly in tools like CUDA and Pytorch. This full-... 
    Full time

    Scale AI

    San Francisco, CA
    2 days ago
  •  ...NEAR AI was started by Illia Polosukhin, co-author of the landmark paper Attention Is...  ...high‑performance LLM serving systems and inference optimization. In this role, you will push...  ...debugging and optimizing major inference engines such as SGLang, vLLM, or TensorRT. Deep knowledge... 

    NEAR.AI

    San Francisco, CA
    1 day ago
  •  ...Dynamo AI is seeking a candidate to lead LLM evaluation and benchmarking in San Francisco, California. You will generate high-quality data and develop innovative methods for assessing the safety and helpfulness of LLMs. The role requires domain knowledge in evaluation... 

    Capitolis

    San Francisco, CA
    2 days ago
  •  ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by one...  ...a Senior Systems GPU Engineer - AI & Robotics, you will be responsible for iteratively...  ...foundational models to real-time onboard inference—while serving as a core contributor to... 
    Local area
    Worldwide
    Flexible hours

    Intuitive Surgical

    San Francisco, CA
    1 day ago
  • $300 per month

     ...intelligence. As the only vertically integrated AI infrastructure company built from the...  ...the Role:At Crusoe, our Production Engineering team ensures the reliability and scalability...  ...to support distributed AI pipelines and inference servicesDefine, measure, and improve... 
    Temporary work

    Crusoe

    San Francisco, CA
    2 days ago
  • $141k - $184k

     ...with the latest advancements in AI and IoTWho we areThe...  ...a motivated Computer Vision Engineer to help build the next generation...  ...conversion, quantization, and inference acceleration.Build tools for...  ...data processing, visualization, benchmarking, evaluation, and monitoring.Conduct... 
    Full time
    Internship
    Local area
    Remote work
    Worldwide

    Pano AI

    San Francisco, CA
    5 days ago
  •  ...Overview \ We're hiring a Machine Learning Engineer to lead our progression from...  ...and semi -autonomous. \ \ With proven AI approaches and long distance teleoperation...  ...to edge compute devices with real -time inference constraints. \\ Collaborate with Robotics... 
    Full time
    Remote work
    Worldwide
    Relocation
    Long distance
    Flexible hours

    Avatar Robotics

    San Francisco, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff AI Inference Benchmarking Engineer. Be the first to apply!