GPU Performance Engineer: Scale AI Inference
$315kAnthropic
A leading AI research company in San Francisco is seeking a mid-senior GPU Performance Engineer. In this role, you'll architect systems that enhance GPU performance for groundbreaking AI models. Responsibilities include developing optimizations, collaborating with teams, and troubleshooting performance issues. The position offers a competitive salary range of $315,000—$560,000, equity opportunities, and a supportive work environment that values communication and collaboration. #J-18808-Ljbffr Anthropic
$250k
...opportunities? Join a rapidly scaling AI cloud infrastructure... ...a next-generation GPU platform designed for... ...experimentation, and inference at scale. The company... ...Site Reliability Engineer to support and scale large... ...reliability, scalability, and performance of HPC and cloud...PerformanceFull timeRemote work$175k - $250k
...We're a well-funded AI infrastructure startup... ...artificial intelligence, high-performance computing, and... ...We're looking for an engineer to help build and maintain a high-performance inference library designed to support... ...ROCm, Triton, or similar GPU/accelerator programming...PerformanceLocal area$190k - $250k
...Sciforium is an AI infrastructure company developing... ...-on support from AMD engineers the team is scaling rapidly to build the... ...a highly skilled GPU Kernel Engineer who... ...pushing the limits of performance on modern... ...large-scale training and inference. This role is ideal...PerformanceFull timeFlexible hours- ...powers mission-critical inference for the world's most dynamic AI companies, like... ...build the platform engineers turn to to ship AI... ...multi-modal workloads scale, the network is the... ...to lead our GPU Networking efforts,... ...validate networking performance on bleeding-edge clusters...PerformanceFull timeFlexible hours
$140k - $200k
...pioneering the future of Physical AI. Our advanced vision... ...Data Infrastructure / Quality Engineering role will play a crucial... ...validation, and production-scale release Experience architecting... ...and optimizing high-performance GPU cloud inference services, with specific expertise...PerformanceFull timeWork experience placementLocal area- ...the world’s largest AI infrastructure networks... ...highly available GPU infrastructure for AI training and inference workloads.About the... ...Infrastructure Operations Engineer to operate and improve the large-scale Ethernet fabrics... ...infrastructure.Perform root-cause analysis...PerformancePermanent employment
$188k - $275k
...Essential Cloud for AI™. Built for... ...innovators to build and scale AI with confidence... ...infrastructure performance with deep technical... ...Do: The Field Engineering organization at CoreWeave... ...can train and inference on at scale,... ...lifecycle: leading new GPU cluster bring-up...PerformancePermanent employmentFull timeContract workTemporary workCasual workWork at officeFlexible hours$106.9k - $200.6k
...opportunity We are seeking an AI Systems Engineer to own the delivery, model-... ...pipelines, operating high-performance inference (GPUs, model servers,... ...sink), DCGM exporter for GPU telemetry, processor batching... ...NIM) on GPUs at production scale. Deep observability...PerformanceFull timeSummer holidayFlexible hours- ...About the Team Our Inference team brings OpenAI’s most... ...our state-of-the-art AI models, allowing them... ...before. We focus on performant and efficient model inference... ...Role We’re hiring engineers to scale and optimize OpenAI’s... ...across emerging GPU platforms. You’ll work...PerformanceFull time
$350k
...Join a rapidly growing AI infrastructure... ...provider delivering large-scale compute solutions for AI training and inference across global cloud and GPU environments. The... ...build reliable, high-performance platforms supporting... ...Site Reliability Engineer to lead the reliability...PerformanceFull time$185k - $260k
...integrate with advanced AI, and are ultimately... ...Working with scientists and engineers, we combine simulation,... .... You will define performance objectives and develop,... ...efficient enough for large-scale offline processing.... ...profiling and improving GPU performance and memory...PerformanceFull time$229.5k - $255k
...the future of physical AI. Powered by a proprietary... ...autonomy teams and engineers overcome is the sim-to-... ...geometrically coherent, correctly scaled and aligned, and usable... ...implicit. Improve Performance, Throughput, and Cost - Profile and optimize GPU and distributed...PerformanceFull timeWork at office3 days per week$125.5k - $261.6k
...We are seeking AI Systems Engineers to build and operate... ...from bare-metal and GPU infrastructure through... ...fabric, and multi-tenant scaling. You will be responsible... ...secure execution and inference: Ray Serve, vLLM/NIM... ...rewarded based on your performance and recognized for...PerformanceFull timeContract workSummer holidayFlexible hours$125k - $165k
..., delivering power to AI data centers in months... ...proprietary energy architecture engineered entirely in-house. The... ..., and the operational scale to lead it.... ...characteristics, and performance parameters — and serve... ...employment information, and inferences drawn from your PI. We...PerformanceFull time- ...Head of GPU Cloud About the Company Developing... ...platform for large-scale AI infrastructure.... ...Cloud to spearhead the engineering team dedicated to the... ...model for large-scale AI inference services, with a direct... ...platform's reliability and performance. Close collaboration...Performance
$175k - $250k
...Senior Cloud Infrastructure Engineer Location: San Francisco... ...with generative AI. They are the team behind... ...and maintaining large-scale distributed systems that... ...ensuring scalability, performance, and reliability across... ...scale Manage and automate GPU compute clusters using...PerformanceFull timeRemote workRelocationRelocation package- ...compute infrastructure scales efficiently to... ...increasingly sophisticated AI models.We’re... ...with Capacity Systems Engineering, Infrastructure,... ...Research to optimize inference capacity across our global GPU fleet. This role... ...infrastructure investments, performance-efficiency trade-...Performance
$180k - $250k
...the next generation of AI products. We build the... ..., and do it at scale without compromise. For... ...unified platform where high-performance inference, orchestration, and... ...experienced software engineer who thrives on building... ...orchestration, scheduling, GPU autoscaling, large...PerformanceFull timeCurrently hiringRemote workRelocation package$200k
...high-speed networks powering the AI era? Join a trailblazing leader in GPU-accelerated computing, designing and... ...real-time visibility into the performance of thousands of interconnected nodes... ...Terraform, or SaltStack. High-Scale Networking: A strong foundation in...PerformanceFull time- ...deeply in Generative AI — pioneering... ...Learning Systems Engineer (P60) to lead technical... ...on building and scaling the systems that power... ...reliable, high-performance infrastructure.Working... ...model training, inference pipelines, or... ...performance computing, or GPU optimization....PerformanceWork at officeLocal area
- ...-a-Service (TaaS) Engineer to help build the... ...that convert large-scale infrastructure capacity... ...will work across performance benchmarking,... ...infrastructure stack, ensuring GPU capacity can be... ...GPU clusters, AI infrastructure,... ...model porting, inference/training workloads...PerformanceFull time
$165k - $200k
...vertically integrated AI infrastructure company... ...urgency, who believe in the scale of our ambition and... ...and be part of a high-performing team that believes in each... ...Network Production Engineer to support the physical... ...performance compute (HPC) and GPU-based AI infrastructure...PerformanceFull timeTemporary workRemote work$231.6k - $289.5k
...delivering modular AI infrastructure... ...factory with speed, scale and sovereignty. Named... ...a Distinguished Engineer / Technical Fellow... ...in high-density GPU orchestration and... ...such as in-orbit inference, adaptive learning... ...a culture of high-performance engineering and architectural...PerformanceRemote workFlexible hours$161.3k - $241.9k
...frontier agentic AI, an enterprise-grade... ...support. We’re scaling fast and defining... ...interaction, every model inference, and every... ...for a Production Engineer to help build and... ...reliability, and performance. Improve compute... ...Experience operating GPU fleets, high-performance...PerformanceFull time$155k - $269k
...Description Waabi, founded by AI visionary Raquel... ...Scientists and Engineers building the content backbone... ...and curate assets at the scale of tens of thousands of... ..., and debugging GPU jobs in the cloud (AWS,... ...incentive awards and an annual performance bonus. Perks/...PerformanceFull timeWork at officeWork from homeFlexible hours$300k
...startup building out their AI and cloud platform, powered... ...ready for experimentation, full-scale model training, or inference. As a Platform Engineer/Senior Site Reliability... ...you’ll own the reliability, performance, and automation of this GPU-powered infrastructure, ensuring...PerformancePermanent employment$342k
...demands of advanced AI workloads. The... ...About the RoleAs an Engineer on our hardware optimization... ...and performance. You will work with... ...efficient training and inference on our models. If... ...decisions on scale up, scale out, front... ...understanding of GPU and/or other AI acceleratorsExperience...PerformanceWork at officeLocal areaRelocation packageFlexible hours$106.9k - $200.6k
...opportunity We are seeking AI Systems Engineers to own the security and... ...lifecycle at production scale across multiple environments... ...challenges of confidential AI inference (models/secrets inside... ...be rewarded based on your performance and recognized for the value...PerformanceFull timeWork experience placementSummer holidayRemote workFlexible hours$250k
...Ready to architect AI infrastructure... ...building a serverless inference platform,... ...Inference Platform Engineer at an early stage... ...systems to maximise GPU utilisation and minimise... ...best practices in performance and efficiency.... ...experience building large-scale, fault-tolerant...PerformanceFull time- ...Senior Systems Engineer San Francisco, California Onsite or Remote... ...posture, database performance, AI inference infrastructure: you can cover... ...platforms (vLLM, Cloud Run, GPU provisioning), latency SLOs, cost optimization, auto-scaling characteristics, and capacity...PerformanceRemote workWork from home
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to GPU Performance Engineer: Scale AI Inference. Be the first to apply!
- performance food group San Francisco, CA
- senior performance tester San Francisco, CA
- application performance engineer San Francisco, CA
- performance test architect San Francisco, CA
- performance test engineer San Francisco, CA
- senior performance engineer San Francisco, CA
- performance improvement consultant San Francisco, CA
- acting performance San Francisco, CA
- cognitive performance San Francisco, CA
- IT performance management San Francisco, CA



