Staff AI Inference Benchmarking Engineer
Artificial Analysis
Artificial Analysis, the leading AI benchmarking company, is hiring a Member of Technical Staff to own serverless inference coverage and drive performance benchmarks across provider ecosystems. You will extend methodologies, collaborate with neoclouds, and help shape how the market measures serving performance. You will work with engineers and leadership to direct the roadmap of the benchmarking platform, analyze throughput, time-to-first-token, and cost-per-token economics, and contribute to #J-18808-Ljbffr Artificial Analysis
- Artificial Analysis, Inc. is seeking a Member of Technical Staff (Hardware) to design and run AI hardware benchmarks for GPUs, TPUs and custom silicon in San Francisco... .../data-analysis skills, and deep knowledge of inference economics. #J-18808-Ljbffr Artificial Analysis,...Suggested
- An innovative company is seeking a talented software engineer to join their dynamic Inference team. This role involves designing and implementing infrastructure... ...researchers and product teams to push the boundaries of AI technology, ensuring reliable production services. If...Suggested
- Crusoe in San Francisco is seeking a Staff Technical Program Manager to lead the Managed Inference platform team. You will ensure end-to-end program delivery for LLM workloads, driving innovation in AI infrastructure powered by clean energy. The ideal candidate has extensive...Suggested
- Dynamo AI in San Francisco seeks an ML Engineer focused on evaluating LLMs, benchmarking, and ensuring safe, responsible AI in production. You will own end-to-end evaluation pipelines and contribute to synthetic data generation. You will collaborate with researchers to...Suggested
- Anyscale is seeking a Distributed LLM Inference Engineer in San Francisco, California. This pivotal role involves pushing the boundaries of performance for ML inference at scale. You'll work closely with product teams to deliver end-to-end solutions while leveraging open...Suggested
- Causal Labs in San Francisco is seeking an experienced Infrastructure Engineer to build high-throughput inference systems for large-scale evaluation and backtesting against historical physical observations. You will design techniques to improve latency and throughput, optimize...
- Sail Research in San Francisco is seeking a talented engineer to design and implement robust systems that ensure fast and cost-efficient AI inference at global scale. You will be responsible for building high-performance schedulers and optimizing global routing while focusing...
- Kindredventures is recruiting infrastructure engineers to scale large-scale inference and evaluation around a physics-based LPM initiative. You will work on high-throughput systems, latency-optimized serving, and distributed orchestration across Kubernetes, Ray, and Slurm...
- A tech startup focused on AI workloads is seeking a Member of Technical Staff to design and optimize inference systems. The role involves managing KV cache allocation and... ...Ideal candidates should have strong software engineering skills and experience with ML inference...
- A leading technology company located in San Francisco is seeking a Machine Learning Research Engineer to design and develop safe AI benchmarking methodologies. This role involves collaboration with various teams to implement responsible evaluation techniques. Candidates...
$315k
A leading AI research company in San Francisco is seeking a mid-senior GPU Performance Engineer. In this role, you'll architect systems that enhance GPU performance for groundbreaking AI models. Responsibilities include developing optimizations, collaborating with teams...- A leading AI acceleration company in San Francisco is seeking a GPU Kernel Engineer to optimize performance for machine learning models. You will be responsible for designing high-performance GPU kernels and using advanced techniques to boost computation efficiency. Ideal...
- ...team of YC and unicorn founders and senior engineers with deep expertise in 3D, generative... ...'re looking for a Founding Engineer, ML Inference with deep expertise in high-performance... ...while maintaining accuracy Profile and benchmark model performance to identify computational...RelocationVisa sponsorshipRelocation package
$325k
A leading AI research company in San Francisco seeks an engineer to optimize their powerful AI models for high-volume production environments. The ideal candidate has over 5 years of software engineering experience, strong familiarity with ML architectures, and experience...- ...a Tokens-as-a-Service (TaaS) Engineer to help build the systems that... ...will work across performance benchmarking, tokenomics, model porting,... ...Experience with GPU clusters, AI infrastructure, performance benchmarking... ...with model porting, inference/training workloads, token...Full time
- Databricks Mosaic AI is seeking an experienced backend/infrastructure engineer to build platforms powering AI workloads, including model training, serving, and vector search. You will join a high-visibility team connected to research, product, and enterprise use cases,...
- ...Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence... ...us and help build the platform engineers turn to to ship AI products.... .... Experience running low-level benchmarks to "qualify" new hardware clusters...Full timeFlexible hours
- OpenAI in San Francisco is seeking an experienced systems generalist to build an automated inference optimization platform across hardware, compiler, and runtime contexts. You will design the OpenAI-hosted control plane and partner-side software, focusing on reliable long...
$170k - $245k
..., have Ray in their tech stacks to accelerate the progress of AI applications out into the real world.With Anyscale, we’re building... ...50+ million raised to date.About the roleAs a Distributed LLM Inference Engineer, you will help systems and optimizations that push the...Work at office$160k - $230k
About the RoleAt Together.ai, we are building state-of-the-art infrastructure to enable efficient and scalable inference for large language models (LLMs). Our mission is to optimize... ...anInference Frameworks and Optimization Engineer to design, develop, and optimize...Full time$102.3k - $161.76k
...matters and so do you. About the Role As a Staff AI FinOps Governance Lead, you will lead... .... You will partner closely with Engineering, Product, Finance, Procurement, and Cloud... ...drivers of Generative AI, including LLMs, inference services, vector databases, GPUs, and AI...$190k - $210k
...About The RoleSenior Developer Success Engineers (Sr. DSEs) own the post-sale technical relationship... ...includes a competitive base salary benchmarked against real-time market data, as well... ....We may use artificial intelligence (AI) tools to support parts of the hiring process...Full timeWork at office2 days per week$216k - $270k
About Scale AIScale AI is the data foundation for AI, helping organizations build and... ...problems remains one of the hardest engineering challenges.As a Senior Frontier Agent Engineer... ...evaluation frameworks using offline benchmarks, online A/B experiments, golden datasets...Full time$200.8k - $251k
A leading AI technology company in San Francisco seeks a team member to build and optimize a machine learning framework for large... ...should have system optimization experience and solid software engineering skills, particularly in tools like CUDA and Pytorch. This full-...Full time- ...NEAR AI was started by Illia Polosukhin, co-author of the landmark paper Attention Is... ...high‑performance LLM serving systems and inference optimization. In this role, you will push... ...debugging and optimizing major inference engines such as SGLang, vLLM, or TensorRT. Deep knowledge...
- ...Dynamo AI is seeking a candidate to lead LLM evaluation and benchmarking in San Francisco, California. You will generate high-quality data and develop innovative methods for assessing the safety and helpfulness of LLMs. The role requires domain knowledge in evaluation...
- ...millions of patients worldwide.We’re a team of engineers, clinicians, and innovators united by one... ...a Senior Systems GPU Engineer - AI & Robotics, you will be responsible for iteratively... ...foundational models to real-time onboard inference—while serving as a core contributor to...Local areaWorldwideFlexible hours
$300 per month
...intelligence. As the only vertically integrated AI infrastructure company built from the... ...the Role:At Crusoe, our Production Engineering team ensures the reliability and scalability... ...to support distributed AI pipelines and inference servicesDefine, measure, and improve...Temporary work$141k - $184k
...with the latest advancements in AI and IoTWho we areThe... ...a motivated Computer Vision Engineer to help build the next generation... ...conversion, quantization, and inference acceleration.Build tools for... ...data processing, visualization, benchmarking, evaluation, and monitoring.Conduct...Full timeInternshipLocal areaRemote workWorldwide- ...Overview \ We're hiring a Machine Learning Engineer to lead our progression from... ...and semi -autonomous. \ \ With proven AI approaches and long distance teleoperation... ...to edge compute devices with real -time inference constraints. \\ Collaborate with Robotics...Full timeRemote workWorldwideRelocationLong distanceFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff AI Inference Benchmarking Engineer. Be the first to apply!
- software engineer staff San Francisco, CA
- assistant engineer San Francisco, CA
- engineering aide San Francisco, CA
- staff engineer San Francisco, CA
- staff security engineer San Francisco, CA
- assistant mechanical engineer San Francisco, CA
- assistant engineering manager San Francisco, CA
- senior staff systems engineer San Francisco, CA
- technology administrator San Francisco, CA
- project engineer assistant project manager San Francisco, CA



