Member of Technical Staff, TPU & AMD GPU Performance Engineering
$200k - $400kInferact
Role Description
We're looking for a TPU and AMD GPU performance engineer to make vLLM a first-class inference engine across non-NVIDIA accelerators. Frontier inference cannot be locked to one hardware stack. As AMD GPUs, TPUs, and other accelerators become increasingly important, vLLM needs backend paths that are fast, correct, benchmarked, and maintainable across heterogeneous hardware platforms.
- Build and optimize AMD GPU and TPU backends, kernels, compiler integrations, runtime paths, and benchmarking infrastructure.
- Work at the boundary of inference systems, kernels, compilers, and hardware architecture.
- Improve paths such as attention, GEMM, sampling, KV-cache, communication-heavy operations, and model serving on non-NVIDIA hardware.
- Your work will directly impact how broadly and efficiently the world can run AI inference with vLLM.
Qualifications
- Bachelor's degree or equivalent experience in computer science, engineering, machine learning systems, hardware systems, compilers, or similar.
- Hands-on experience optimizing workloads on AMD GPUs, TPUs, or another non-NVIDIA accelerator stack.
- Experience with AMD ecosystem tools such as ROCm, HIP, Triton, CK, AITER, or equivalent GPU performance libraries and tooling.
- Experience with TPU, XLA, JAX, Pallas, or related compiler and runtime tooling for accelerator workloads.
- Ability to optimize ML inference paths such as attention, GEMM, sampling, KV-cache, fused kernels, backend runtimes, or communication-heavy operations.
- Strong performance profiling and benchmarking discipline, including tokens/second, latency, throughput, correctness parity, hardware counters, and reproducible measurement methodology.
- Ability to navigate immature tooling, incomplete documentation, backend-specific rough edges, and cross-platform performance differences without getting stuck.
Requirements
- Experience with vLLM, SGLang, TensorRT-LLM, ATOM, JAX-based serving framework, or other LLM inference systems.
- Deep understanding of inference architecture and serving tradeoffs, including batching, KV-cache, decoding, prefill/decode scheduling, and backend performance constraints.
- Experience with compiler technologies such as XLA, MLIR, LLVM, Triton, Pallas, or other compiler/kernel DSLs, including lowering, fusion, and backend code generation.
- Knowledge of quantization techniques such as MXFP8, MXFP4, mixed precision, or hardware-specific numeric formats, and the ability to reason about accuracy/performance tradeoffs.
- Experience with distributed inference performance, including communication, memory movement, hardware topology, and scale-out bottlenecks across multi-accelerator workloads.
- Open-source contributions to vLLM, JAX/XLA, ROCm, Triton, PyTorch, compiler projects, or related ML systems infrastructure.
Benefits
- Generous health, dental, and vision benefits.
- 401(k) company match.
Logistics
- Location: This role is based in San Francisco, California. Will consider remote in the US for exceptional candidates.
- Compensation: Depending on background, skills, and experience, the expected annual salary range for this position is $200,000 - $400,000 USD + equity.
- Visa sponsorship: We sponsor visas on a case-by-case basis.
- Role Description We're looking for an AMD GPU performance engineer to make vLLM a first-class inference engine across the AMD accelerator ecosystem. You'll build and optimize AMD GPU backends, kernels, runtime paths, and benchmarking infrastructure using ROCm, HIP, Triton...PerformanceFull time
$150k - $300k
...On-site Department Engineering Building Open Superintelligence... ...system with performance engineering at its... ...reliable at scale. Core Technical Responsibilities... ...heterogeneous hardware (CPU, GPU, TPU) Platform... ...development and encourage team members to contribute to the...PerformanceFull timeWork at officeRemote workVisa sponsorshipRelocation packageFlexible hours- Member of Technical Staff - Engineer, RL and Control Position Summary You will help build and refine the learned... ...tools for evaluating controller performance in simulation and on hardware. Collaboration... ...-of-the-art machine learning and GPU programming frameworks, such as...PerformanceWork from homeFlexible hours
- ...Consulting Member Of Technical Staff As a Consulting Member of Technical Staff, you will be a... ...technologies. Your expertise in data platform engineering and service development will drive... ...decisions where analysis of data, performance, privacy, security, and healthcare...PerformanceImmediate startRemote workWorldwide
- ...inference) push hardware to its limits. We're looking for an engineer to own the GPU and ML-systems layer that frontier AI teams run on: ~... ...(e.g. Kueue, KAI, KServe) ~You've done real ML-systems performance work - tell us about a bottleneck you hunted down (a stalled...PerformanceFull time
$250k - $300k
...built a platform that deploys GPU clusters into third-party... ...to GB300. As part of the engineering team, you'll help shape a platform... ...designing and building high-performance systems spanning storage, networking... ...A track record of impressive technical work you can speak to in...PerformanceFull timeRemote work- Job Role: Member of Technical Staff (Austin, TX Remote) The Member of Technical Staff (MTS) - Systems... ...and expertise for complex systems engineering projects. MTS engineers serve as... ...patches. Ensure kernel stability and performance. Feature Development Design and implement...PerformanceRemote work
$110.4k - $165.5k
...problems and providing unmatched technical expertise. As the operator... ...for an RF/Microwave Engineer who wants their work to... ...complex systems – Generate performance budgets and run circuit simulations... ...to be Successful – Senior Member of the Technical Staff Minimum Requirements:...PerformanceFull timeFor contractorsWork at officeImmediate startRemote workRelocation packageFlexible hours$110.4k - $165.5k
...RF/Microwave Engineer The Aerospace Corporation is the trusted... ...and providing unmatched technical expertise. As the operator... ...systems – Generate performance budgets and run circuit simulations... ...to be Successful – Senior Member of the Technical Staff Minimum Requirements:...PerformanceFull timeFor contractorsWork at officeImmediate startRemote workRelocation packageFlexible hours- ...enterprise finance workflows. As a Member of Technical Staff in Finance Research, you will develop... ...agentic systems, working across research, engineering, and product teams in a remote, full... ...turning findings into measurable AI performance improvements. Develop and curate...PerformanceFull timeRemote work
$180k - $238.1k
...Member of Technical Staff (Performance Engineering) New York, NY Category-defining tech. Career-defining work. Lots of tech companies disrupt. But, many fail when they try to scale. We're different. CockroachDB makes it easier for companies to build and scale...PerformanceLocal areaRemote workWorldwideFlexible hours- Member of Technical Staff: Mechanical Engineering Position Summary: We are hiring a Mechanical Engineer to design, build, and ship high-performance actuators for our robot. In this role, you will own the design of defined actuator mechanical components and subassemblies...PerformanceWork from homeFlexible hours
$95.2k - $165.5k
...problems and providing unmatched technical expertise. As the operator of a... ...MPD, the Communication Systems Engineering Department (CSED) performs analysis, modeling and simulation... ...Communications Systems Engineer (Member of Technical Staff/Senior Member of Technical Staff...PerformanceFull timeImmediate startRemote workRelocation packageFlexible hours- ...About the job TL;DR: A founding-style, full-stack engineer who ships end to end across a high-performance Go backend and a cross-platform iOS and Android app (React Native and Expo), with AI agents as your force multiplier. If mobile UI/UX is your strength, you can...PerformanceRemote workFlexible hoursShift work
- ...Pixeltable Inc. Member of Technical Staff San Francisco, CA·Full time Apply for... ...As a founding member of the engineering team, you will impact the design... ...compute resources (CPU and GPU) efficiently? What data... ...to building a high-performing, inclusive team with a professional...Full timePart timeWork at officeWork from homeFlexible hours2 days per week
- ...Role Overview Lead technical ownership at the intersection of research, data, and deployed... ...AI systems to improve model and agent performance through rigorous evaluation, failure... ...domains such as finance, healthcare, and engineering. Key Responsibilities Own...PerformanceFull timeRemote work
- ...systems. We are seeking a Senior Member of Technical Staff to build core systems, solve challenging engineering problems, and contribute... ...internal tooling. ~Solve complex performance, reliability, and scalability... ...or scale-up experience. ~GPU or data-intensive systems....PerformanceFull time
- Member of the Technical Staff, Systems Location: North America Remote / San Francisco... ...and billions of GPU-hours supported, on everything... ...We are looking for strong engineers with experience and interest... ...designing and building high performance systems across, but not...PerformanceFull timeRemote work
- ...step-function improvements in performance and efficiency. Customers deploy... ...an Infrastructure platform Engineer to design, build, and operate... ...and operate large‑scale CPU, GPU, and accelerator clusters powering... ...environments across NVIDIA, AMD, Intel, ARM, or emerging accelerators...Performance
- ...new physics at scale. We are seeking engineers to build platform infrastructure at... ...physics, at scale. We are seeking a Member of Technical Staff, Performance & Capacity to answer two questions honestly... ...Looking For Five or more years with GPU and large-scale compute workloads,...PerformanceRemote work
- Mount Thor is hiring a software engineer with deep networking expertise to build and operate... ...(macOS and Apple Silicon) available and performant at datacenter scale for AI workloads. We... ...role you will Own the architecture and technical roadmap for Mount Thor’s production...Performance
- ...building AI systems to discover new physics at scale. We are seeking engineers to build platform infrastructure at the intersection of... ...dynamics, FEA, quantum, or comparable), with comfort in high‑performance computing environments. You can read and write production code...PerformanceRemote work
$230k
...over 10 times faster than GPU-based hyperscale cloud inference... ...multiple openings for Sr. Member of Technical Staff. Title: Sr. Member of... ...latency and scalable system performance. Develop Python-based scripts... ...Security Analyst, Software Engineer, Sr. Member of Technical...PerformanceRemote work- ...is a team of researchers, engineers, designers, and more, who are... ...focused team, breakthrough performance doesn’t require... ..., and join the team. As a member of technical staff with a focus on multimodal... ...Experience in writing efficient GPU kernels using CUDA, optimising...PerformanceFull timeWork at officeLocal areaRemote workHome office
- Member of Technical Staff - AI Cloud Infrastructure Location Bay Area; Boston; Washington... ...a senior infrastructure engineer to architect and stand up... ...AI platform, whether at a GPU cloud, an internal machine... ...clouds. Lead high-performance storage strategy. Deploy and...PerformanceFull timeWork from homeFlexible hours2 days per week
$150k - $300k
...training stack. Core Technical Responsibilities LLM Serving... ...across our cloud GPU fleets. GPU‑Aware... ...Inference Optimization & Performance Framework Development... ...PyTorch: LLM Inference engine development and... ...development and encourage team members to contribute to the...PerformanceWork at officeRemote workVisa sponsorshipRelocation packageFlexible hoursShift work- ...at About the Role As a Member of Technical Staff, you will help invent and build... ...with hands‑on software engineering experience and are excited... ...computing Systems for AI or HPC Performance optimization Excellent... .... Experience with GPU systems, accelerators, or performance...PerformanceWork from homeFlexible hours2 days per week
$150k - $300k
...tuning runs on managed GPU clusters with a single... ...runs the jobs. Core Technical Responsibilities Hosted... ...We're looking for engineers who are fluent across... ...networking, namespaces, performance tuning Programming &... ...development and encourage team members to contribute to the...PerformanceWork at officeLocal areaRemote workVisa sponsorshipRelocation packageFlexible hours- ...systems. We are seeking a Staff level Member of Technical Staff to lead high-impact... ...initiatives and solve complex engineering problems spanning multiple... ..., reliability, and performance bottlenecks. ~Establish... ...experience (Nice to Have). ~GPU or data-intensive systems...PerformanceFull time
- ...Description Job Description Member of Technical Staff, Machine Learning,... ...Technical Staff role is for engineers who want to develop strong... ...data. - Debug model issues, performance problems, and production... ...improvement. - Tech Stack: GPU, JAX, ML, Machine Learning,...PerformanceRemote workWork from home
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Member of Technical Staff, TPU & AMD GPU Performance Engineering. Be the first to apply!
- operations support technician Remote
- product support technician Remote
- senior technical analyst Remote
- systems support technician Remote
- user support analyst Remote
- junior IT service management analyst Remote
- technical support specialist Remote
- remote support technician Remote
- help desk assistant Remote
- personal computer support technician Remote



