Frontier AI Workloads - Performance and Scalability Engineer
Advanced Micro Devices Inc
WHAT YOU DO AT AMD CHANGES EVERYTHING At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career. THE ROLE:AMD is looking for a Principal Engineer to serve as a hands-on technical team lead driving the performance and scalability of frontier AI workloads on AMD GPUs, including large language models, mixture-of-experts architectures, and diffusion models. You will lead a team of engineers, define the long-term technical vision, make critical architecture decisions, and tackle the hardest performance challenges across the stack from GPU kernels and to serving frameworks and distributed systems.THE PERSON:The ideal candidate is a deep technical expert with a track record of solving industry-hard problems at the intersection of GPU architecture, AI systems, and high-performance software. You understand the full stack from hardware micro-architecture to model architecture, inference paradigms, and system-level design. You lead through technical depth, influence, and by example, staying hands-on while setting direction for your team. If you want to shape how the world runs AI on AMD hardware, this role is for you.KEY RESPONSIBILITIES:Lead a small team of engineers: set technical direction, prioritize work, and ensure delivery while remaining deeply hands-onDefine and drive the long-term technical strategy for AI workload performance on AMD GPUsOwn the most complex cross-stack performance challenges, from kernel optimization to framework-level architecture decisionsLead the design and implementation of novel GPU kernels, compiler optimizations, and framework featuresEstablish performance methodology and roofline analysis practices that set the standard for the teamInfluence upstream roadmaps in major open-source AI frameworks (e.g., vLLM, SGLang, PyTorch)Drive architecture decisions for emerging inference paradigms (e.g., prefill-decode disaggregation, speculative decoding, distributed serving)Identify and close fundamental performance gaps between AMD and competitor platformsServe as a technical authority across the organization, advising leadership on technical direction and feasibilityMentor engineers and raise the technical bar across the broader engineering organizationRepresent AMD externally through publications, conference talks, and open-source contributionsPREFERRED EXPERIENCE:Deep software development experience in GPU computing, HPC, or AI systemsDeep understanding of GPU micro-architecture, memory hierarchy, instruction scheduling, and performance tradeoffsDeep understanding of end-to-end AI systems: model architectures, inference paradigms, and system/rack-level designUnderstanding of multi-GPU communication: scale-up (NVLink, xGMI, Infinity Fabric) and scale-out (RDMA, RCCL/NCCL) topologies and performance characteristicsExperience designing and optimizing across the full stack: from low-level GPU kernels to frameworks and distributed serving systemsStrong background in performance engineering, including profiling, roofline analysis, and bottleneck diagnosis at scaleExperience with one or more of: HIP, CUDA, OpenCL, Triton/Gluon, CUTLASS, CKExperience with GPU compiler toolchains (e.g., LLVM) and intermediate representations (e.g., MLIR, LLVM IR, Triton IR) is a plusHands-on experience contributing to or architecting major open-source AI frameworks (e.g., vLLM, SGLang, xDiT, Megatron LM, PyTorch)Strong proficiency in C++ (C++17 or later) and PythonExperience leading small technical teams while remaining a hands-on contributorTrack record of influencing technical direction across teams and organizationsStrong Linux systems knowledgeExcellent written and verbal English communication skillsPublished research or significant open-source contributions in GPU computing, HPC, or AI systems is a plusPREFERRED ACADEMIC CREDENTIALS:Master's or PhD in Computer Science, Computer Engineering, Electrical Engineering, or equivalent. PhD strongly preferred.LOCATION:San Jose, CA preferredThis role is not eligible for visa sponsorship.#LI-G11#LI-HYBRIDBenefits offered are described: AMD benefits at a glance.AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.This posting is for an existing vacancy.
$2,000 per month
...building hardware for frontier intelligence. We... ...and decode workloads. Our first products... ...staffed by leading engineers, Etched is... ...engineering, and high-performance computing (HPC),... ...simulation, and AI model deployment... ...Kubernetes to enhance scalability and...PerformanceWork at officeRelocation package$124k - $195.5k
...recruiting a Senior Inference Performance Engineer to push NVIDIA's performance limits on large-scale AI inference benchmarks. This... ...improvement of AI inference workloads that methodically increase throughput... ...advance the public Pareto frontier while maintaining strict...PerformanceFull time$184k - $287.5k
...unlimited potential of AI to define the next era... ...an AI Compiler Engineer with deep expertise in... ...measurable improvements in performance and efficiency, and advancing... ...on representative workloads and benchmark suites.... ...for reliability, scalability, and performance.What...PerformanceFull timeRemote work- Bright Vision Technologies is seeking an AI Systems Engineer to design, build, and operate the... ...large-scale AI training and inference workloads. The role emphasizes GPU clusters, distributed... ...frameworks, scheduling, storage performance, and developer experience for ML engineers...PerformanceRemote job
$266.05k - $396k
...potential of their data, from AI to multicloud.... ...Summary Distinguished Engineer - AI Infrastructure... ...experience with high-performance inference engines (TensorRT... ...optimized for AI workloads (checkpoints, model artifacts... ...design for highly scalable distributed storage and...PerformanceWork at officeLocal areaImmediate start- ...in Sunnyvale, CA, seeks a Senior Software Engineer for Identity Management Services. Join a high-impact team building scalable identity workflows across Apple platforms with... ...deploying large-scale systems, improving performance, and collaborating with cross-functional partners...Performance
$219k - $351k
...communities.Job Title: Principal Engineer, AI System Architect (Hardware)The Architecture... ...step-function improvements in performance, efficiency, and scalability. We are seeking a Principal AI... ...Technical Lead role in bridging AI workloads, system architecture, and hardware...PerformanceWork at officeFlexible hours$163k - $253k
...level bottlenecks in modern AI, particularly in memory... ...function improvements in performance, efficiency, and scalability. We are seeking a Staff... ...key role in bridging AI workloads, system architecture, and... ...interconnect, and system engineering to align modeling insights...PerformanceWork at officeFlexible hours$184k - $287.5k
We are looking for a Senior System Software Engineer, Software Defined Networking to design, build, and operate highly performant and scalable SDN solutions for NVIDIA's AI Clouds hosting GPU-accelerated workloads — including hyperscale multi-node training, inference, cloud...PerformanceFull time$184k - $287.5k
...skilled and motivated software engineers to join us and build AI inference systems that... ...and implement high-performance inference stacks, optimize... ...industry benchmarks, and scale workloads across multi-GPU, multi-... ...teams to push the frontier of accelerated computing...PerformanceFull time- ...job Senior ASIC Verification Engineer (AI Hardware) Senior... ...the quality, reliability, and performance of their next-generation AI-... ...designs used in AI/compute workloads Build and execute detailed... ...validation approaches Develop scalable verification environments...PerformanceRemote workVisa sponsorship
$120k - $280k
...potential of their data, from AI to multicloud.... ...experienced Systems Software Engineers across multiple NetApp... ...Cloud Platforms, and Performance Engineering.... ...systems, performance, scalability, reliability, and data... ...everything from traditional workloads to enterprise AI with...PerformancePart timeWork at officeLocal area- ...globally for innovation, performance and quality. Sandisk... ...establish and lead an AI Systems & Performance... ...bottlenecks across real‑world workloads, software stacks, and... ...reproducibility, and scalability. Establish best... ...highly skilled team of engineers and researchers. Drive...PerformanceTemporary workRemote workFlexible hoursShift work
- ...generation of supercomputing, high-performance computing, cloud, and AI. Whether you’re designing... ...is seeking an AI Systems Engineer to help develop and optimize machine learning workloads on next-generation AMD AI... ...technical insights into scalable solutions that improve...PerformanceWorldwide
$172.8k
As a Principal Systems Engineer, you will lead the... ...solutions for high-density AI data centers in 2026... ...requirements of generative AI workloads. Influence cross... ...roadmaps into scalable DC/DC topologies and system... ...ability to balance cost/performance trade-offs and present...PerformanceLocal area$124k - $208.4k
...IP) that is applied to high-performance computing devices (mobile,... ...enable efficient execution of AI workloads on Samsung’s premium mobile... ..., efficiency, and scalability across a variety of ML workloads... ...with GPU architects, software engineers, and hardware teams to...PerformanceHourly payFull timeRelocation$100k
...leading the industry on cutting-edge AI technology, revolutionizing performance expectations, ease of use, and... ...seniorities.We are seeking an Senior Engineer to develop and optimize the... ...product teams to deliver robust, scalable, and high-performance software solutions...PerformancePermanent employment$152k - $241.5k
...deep learning ignited modern AI — the next era of computing —... ...seeking top-tier AI Compiler Engineers to drive innovation within our... ...of what is possible in AI performance and help build the technology... ...compilation problems for AI workloads (both inference and training)...PerformanceFull time$122.6k - $185k
...the backbone of Generative AI cloud at AWS? Do you want... ...continuous price performance improvements in the cloud... ...enable high performance and scalability in AI/ML and HPC workloads.You are intrigued by the... ...like you. The AWS Hardware Engineering team creates server designs...PerformanceLocal areaFlexible hours$152k - $241.5k
...hiring experienced software engineers to help scale up its AI Infrastructure. We expect... ...systems that enable large scalable GPU clusters to be used for a variety of AI workloads.Designing and developing a... ...diagnose and remediate non-performant GPU assets.Working with teams...PerformanceFull timeRemote work$2,000 per month
...building hardware for frontier intelligence. We co-design... ...prefill and decode workloads. Our first products... ...and staffed by leading engineers, Etched is redefining... ...NoCs — are robust, high-performance, and silicon-ready. This... ...-focused frontier AI system. Our addressable...PerformanceWork at officeRelocation packageNight shift$184k - $287.5k
...scale GPU infrastructure for AI research and production workloads. We are looking for Senior Software Engineers to help build the... ...make GPU clusters reliable, scalable, and safe to run. This role... ...Artificial Intelligence, High-Performance Computing and Visualization...PerformanceFull time- ...recognized globally for innovation, performance and quality. Sandisk has two... ...forward. Job Description Focus: AI workload-driven product development, test... ...automated validation Partner with Test Engineering to deliver scalable and robust production test...PerformanceTemporary workRemote workFlexible hoursShift work
$142.8k - $274.8k
...and Infrastructure Engineering (SCHIE) is the... ...organization is developing AI-native silicon... ...generation of frontier AI models. The... ...custom silicon, high-performance networking,... ...frameworks, stress workloads, and validation tools... ...into scalable tooling and workload...PerformanceOngoing contractWork at officeLocal areaWorldwide3 days per week$184.5k - $249.6k
...seeking a 'Principal FAE - Edge AI' to support strategic... ...relationships with architects, engineering leaders, and key decision-... ...to meet power, performance, scalability, and time-to-market objectives... ...solutions, including AI/ML workloads on embedded and edge platforms...PerformanceWork at officeLocal area$176k - $276k
Production engineering is a field that involves crafting, building... ...storage architecture, high-performance distributed storage, data management... ...are reliable, scalable, and efficient. They optimize... ...latency data access for HPC and AI/ML workloads.Storage Production...PerformanceFull timeFlexible hours$190.2k - $360.5k
...are looking for a Principal AI Systems Engineer with deep C++ expertise to... ...quality C++ components for performance-sensitive, cross-platform environments... ..., testable, and scalable across multiple product surfaces... ...Experience with frontier model APIs such as GPT, Claude...PerformanceFull timeTemporary workLocal areaRemote workWorldwide$2,000 per month
...building hardware for frontier intelligence. We co-design... ...prefill and decode workloads. Our first products... ...and staffed by leading engineers, Etched is redefining... ...product generations in high-performance or high-volume silicon... ...-focused frontier AI system, betting early...PerformanceWork at officeRelocation package$255k - $340k
...Superintelligence Cloud, is a leader in AI cloud infrastructure... ...currently Tuesday.Hardware Engineering at Lambda is responsible... ...work running at the frontier of AI compute, at a scale few... ...deployment, quality, compatibility, performance, and scalability of new systems.Serve as the...PerformanceWork at officeLocal areaWork from homeFlexible hours$90k - $100k
...AI Systems Engineer – Remote Bright Vision Technologies is a technology... ...AI training and inference workloads. The role focuses on GPU clusters... ..., scheduling, storage performance, and developer experience for... ...stacks. Exposure to frontier model training operations....PerformanceFull timeH1bLocal areaImmediate startRemote workVisa sponsorship
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Frontier AI Workloads - Performance and Scalability Engineer. Be the first to apply!
- performance test architect San Jose, CA
- performance food service San Jose, CA
- performance improvement consultant San Jose, CA
- senior performance engineer San Jose, CA
- system performance engineer San Jose, CA
- IT performance management San Jose, CA
- acting performance San Jose, CA
- application performance engineer San Jose, CA
- performance testing San Jose, CA
- performance windows San Jose, CA



