Advanced AI Workloads - Performance and Scalability Engineer
Advanced Micro Devices Inc
WHAT YOU DO AT AMD CHANGES EVERYTHING At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career. THE ROLE: AMD is looking for a Principal Engineer to serve as a hands-on technical team lead driving the performance and scalability of frontier AI workloads on AMD GPUs, including large language models, mixture-of-experts architectures, and diffusion models. You will lead a team of engineers, define the long-term technical vision, make critical architecture decisions, and tackle the hardest performance challenges across the stack from GPU kernels and to serving frameworks and distributed systems.THE PERSON:The ideal candidate is a deep technical expert with a track record of solving industry-hard problems at the intersection of GPU architecture, AI systems, and high-performance software. You understand the full stack from hardware micro-architecture to model architecture, inference paradigms, and system-level design. You lead through technical depth, influence, and by example, staying hands-on while setting direction for your team. If you want to shape how the world runs AI on AMD hardware, this role is for you.KEY RESPONSIBILITIES:Lead a small team of engineers: set technical direction, prioritize work, and ensure delivery while remaining deeply hands-onDefine and drive the long-term technical strategy for AI workload performance on AMD GPUsOwn the most complex cross-stack performance challenges, from kernel optimization to framework-level architecture decisionsLead the design and implementation of novel GPU kernels, compiler optimizations, and framework featuresEstablish performance methodology and roofline analysis practices that set the standard for the teamInfluence upstream roadmaps in major open-source AI frameworks (e.g., vLLM, SGLang, PyTorch)Drive architecture decisions for emerging inference paradigms (e.g., prefill-decode disaggregation, speculative decoding, distributed serving)Identify and close fundamental performance gaps between AMD and competitor platformsServe as a technical authority across the organization, advising leadership on technical direction and feasibilityMentor engineers and raise the technical bar across the broader engineering organizationRepresent AMD externally through publications, conference talks, and open-source contributionsPREFERRED EXPERIENCE:Deep software development experience in GPU computing, HPC, or AI systemsDeep understanding of GPU micro-architecture, memory hierarchy, instruction scheduling, and performance tradeoffsDeep understanding of end-to-end AI systems: model architectures, inference paradigms, and system/rack-level designUnderstanding of multi-GPU communication: scale-up (NVLink, xGMI, Infinity Fabric) and scale-out (RDMA, RCCL/NCCL) topologies and performance characteristicsExperience designing and optimizing across the full stack: from low-level GPU kernels to frameworks and distributed serving systemsStrong background in performance engineering, including profiling, roofline analysis, and bottleneck diagnosis at scaleExperience with one or more of: HIP, CUDA, OpenCL, Triton/Gluon, CUTLASS, CKExperience with GPU compiler toolchains (e.g., LLVM) and intermediate representations (e.g., MLIR, LLVM IR, Triton IR) is a plusHands-on experience contributing to or architecting major open-source AI frameworks (e.g., vLLM, SGLang, xDiT, Megatron LM, PyTorch)Strong proficiency in C++ (C++17 or later) and PythonExperience leading small technical teams while remaining a hands-on contributorTrack record of influencing technical direction across teams and organizationsStrong Linux systems knowledgeExcellent written and verbal English communication skillsPublished research or significant open-source contributions in GPU computing, HPC, or AI systems is a plusPREFERRED ACADEMIC CREDENTIALS:Master's or PhD in Computer Science, Computer Engineering, Electrical Engineering, or equivalent. PhD strongly preferred.LOCATION:San Jose, CA preferredThis role is not eligible for visa sponsorship.#LI-G11#LI-HYBRIDBenefits offered are described: AMD benefits at a glance.AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.This posting is for an existing vacancy.
$147k - $202.5k
...global leader in materials engineering solutions used to... ...virtually every new chip and advanced display in the world.... ...our world - like AI and IoT. If you want... ...collect and analyze data, perform hardware... ...performance optimization and scalability. Collaborate with partner...PerformanceFull time$184k - $287.5k
...unlimited potential of AI to define the next... ...an AI Compiler Engineer with deep expertise... ...measurable improvements in performance and efficiency, and advancing LLM-enabled... ...on representative workloads and benchmark suites... ...standards for reliability, scalability, and performance....PerformanceFull timeRemote work- ...innovations in Flash and advanced memory technologies, our... ...globally for innovation, performance and quality. Sandisk has... ...Description Focus: AI workload-driven product... ...validation Partner with Test Engineering to deliver scalable and robust production test...PerformanceTemporary workRemote workFlexible hoursShift work
$170k - $412.5k
...Technologist, Private Cloud AI - Applied & Agentic... ...-to-cloud company advancing the way people live... ..., ensuring performant, reliable, and trustworthy... ...and partner with engineering to take POCs into scalable, production‑grade... ...integrating AI workloads into production‑grade...PerformanceFull timeWork experience placementWork at officeLocal areaImmediate start2 days per week- ...innovations in Flash and advanced memory... ...for innovation, performance and quality. Sandisk... ...establish and lead an AI Systems &... ...across real‑world workloads, software stacks,... ...reproducibility, and scalability. Establish best... ...skilled team of engineers and researchers....PerformanceTemporary workRemote workFlexible hoursShift work
- ADVANCE YOUR CAREER. ADVANCE THE WORLD.At AMD,... ...supercomputing, high-performance computing, cloud, and AI. Whether you’re... ...seeking an AI Systems Engineer to join our AMD IT... ...GPU clusters, and AI workload schedulers. THE... ...You want to build scalable and highly performant...Performance
$124k - $208.4k
...SummarySamsung, a world leader in advanced semiconductor... ...is applied to high-performance computing devices (... ...execution of AI workloads on Samsung’s premium... ...performance, efficiency, and scalability across a variety of ML... ...architects, software engineers, and hardware teams...PerformanceHourly payFull timeRelocation$122.6k - $185k
...backbone of Generative AI cloud at AWS? Do... ...continuous price performance improvements in the... ...high performance and scalability in AI/ML and HPC workloads.You are intrigued... .... The AWS Hardware Engineering team creates server... ...mentorship and other career-advancing resources here to...PerformanceLocal areaFlexible hours$266.05k - $396k
...potential of their data, from AI to multicloud.... ...Summary Distinguished Engineer - AI Infrastructure... ...experience with high-performance inference engines (TensorRT... ...optimized for AI workloads (checkpoints, model artifacts... ...design for highly scalable distributed storage...PerformancePart timeWork at officeLocal area$2,000 per month
...both prefill and decode workloads. Our first products... ...staffed by leading engineers, Etched is redefining... ...quality assurance for our advanced AI hardware builds—... ...Etched’s standards for performance and reliability.... ...also establishing the scalable quality processes needed...PerformanceContract workWork at officeOverseasRelocation package$138k - $206k
...jobs within a 6-month period. Advancing the World’s Technology... ...Title Senior Physical Design AI/ML Engineer, Logic Pathfinding LabWhat... ...throughput in predicting design performance, power, and area (PPA).... ...qualificationsFamiliarity with state-of-the-art AI workloads and their compute and...PerformanceWork at officeLocal areaFlexible hours$184.5k - $249.6k
...Principal FAE - Edge AI' to support strategic... ...relationships with architects, engineering leaders, and key... ...to meet power, performance, scalability, and time-to-market objectives... ..., system IP, or advanced SoC architectures.... ...solutions, including AI/ML workloads on embedded and edge...PerformanceWork at officeLocal area- ...experiences—from AI and data centers,... ...beyond. Together, we advance your career. THE... ...influential software engineer who is passionate... ...improving the performance of key applications... ...the most demanding workloads.Innovate Across Hardware... ...complex, scalable systems using modern...Performance
$184k - $287.5k
...seeking a Senior Developer Technology Engineer, Artificial Intelligence! Would you... ...researching parallel algorithms to accelerate AI workloads on advanced computer architectures? Is it... ...bottlenecks to achieve the best possible performance of computer hardware? Could you be...PerformanceFull timeWork experience placement$184k - $287.5k
...Senior Developer Technology Engineer!NVIDIA's Developer... ...optimizing large application workloads, eliminating system... ...platform including advanced CPUs, GPUs and interconnects... ...with key customers to perform in-depth analysis and... ...existing vacancy. NVIDIA uses AI tools in its recruiting...PerformanceFull timeWork experience placementRemote work- ...building the future of AI powered digital... ...Unit (TCU) that combines advanced hardware security, AI... ...takes more than great engineering—it takes a team of exceptional... ...Senior QA Engineer - Performance & Reliability to lead... ...of TCU/BMC systems.Workload Analysis: Analyze system...Performance
$240k
...world's largest AI chip, 56 times larger... ...transforming key workloads with ultra high-... ...regression, and performance test strategies... ...Kubernetes to validate scalable SaaS deployments.... ...tools, and engineering best practices.Document... ...groundbreaking advancements in AI!Cerebras...Performance- ...in Sunnyvale, CA, seeks a Senior Software Engineer for Identity Management Services. Join a high-impact team building scalable identity workflows across Apple platforms with... ...deploying large-scale systems, improving performance, and collaborating with cross-functional partners...Performance
$126.94k - $174.58k
...combines analog, digital, AI, and software... ...world, and help drive advancements in automation and robotics... ...regulators deliver on-demand scalable power with the speed,... ...experienced Product Engineer with strong data... ...qualifies for a discretionary performance-based bonus which is...PerformancePermanent employmentFull timeWork at officeDay shift$120k - $200k
...Site Reliability Engineer (Prisma Access) 2... ...highly reliable, scalable, and secure cloud... ...systems are robust and performant. This includes... ...the adoption of AI tools. Develop... ...networking, and container workloads. Strong Linux... ...threats, and advancements through...PerformanceRotating shift- ...Description About us nEye.ai, a well-funded optical... ...hyperscale data centers enhanced performance, efficiency, and scalability. Job Overview We are... ...Systems Packaging & Assembly Engineer to lead the development and integration of advanced packaging and assembly...PerformanceContract work
- ...revolutionizing power for AI-driven data centers to... ...Process Technology Engineer to join our team in... ...project execution to ensure performance and design compliance.... ...development and advancing system integration while... ...enable cost-effective, scalable deployment across...PerformanceFull timeFor contractorsWork at officeWorldwide
$147k - $202.5k
...global leader in materials engineering solutions used to... ...virtually every new chip and advanced display in the world.... ...our world - like AI and IoT. If you want... ...collect and analyze data, perform hardware... ...improve device efficiency, scalability, and manufacturability...PerformanceFull time$161k - $221k
...global leader in materials engineering solutions used to... ...virtually every new chip and advanced display in the world.... ...our world - like AI and IoT. If you want... ...smart glasses and high‑performance optical interconnects.... ...will architect robust, scalable integration flows that...PerformanceFull time$220.92k - $311.89k
...Description: Intel’s AI SoC organization is driving... ...Formal Verification Engineer, you will play a... ...digital designs using advanced formal methods. This position... ...and develop scalable, reusable verification... ...design, analyze power/performance, and uncover bugs.Debug...PerformanceFull timeInternshipLocal areaImmediate startShift work$117k - $234k
...organization, designs, engineers, and operates the... ...services, and Kubernetes workloads at multi-petabyte... ...improving reliability and performance, and enabling self-... ..., observability, and AI/AIOps-driven... ...workloads while continuously advancing scalability, resiliency, utilization...PerformanceFull timeTemporary workPart time- ...Job TitleGenerative AI continues to accelerate... ...redefining interconnect performance with dramatically... ...efficiency, and greater scalability. In this role, you will... ...support of Lumilens' advanced photonics platform.Translate... ..., work instructions, engineering reports, and training...PerformanceShift work
- ...computing experiences—from AI and data centers, to... ...beyond. Together, we advance your career. THE ROLE:... ...Systems Design Engineer to join our growing team... ...will drive balanced, scalable, and automated solutions... ...accurate resultsOptimize the performance, such as fusion and...Performance
$168k - $270.25k
...Experience (NVEX) Solutions Engineering team is looking for an... ...will apply the latest AI technologies to triage... ...issues and AI/ML workloads in huge datacenters of... .... Expertise analyzing performance of distributed GPU-... ...customers to resolve or advance customer issues.Work with...PerformanceFull timeWeekend work$124k - $195.5k
...continue to grow, the intersection of advanced deep learning architectures,... ...team to help unlock maximum hardware performance for emerging AI workloads. You will be a crucial member of a... ...degree in Computer Science, Computer Engineering, Electrical Engineering, or related...PerformanceFull time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Advanced AI Workloads - Performance and Scalability Engineer. Be the first to apply!
- acting performance San Jose, CA
- performance specialist San Jose, CA
- high performance computing engineer San Jose, CA
- performance windows San Jose, CA
- system performance engineer San Jose, CA
- performance improvement specialist San Jose, CA
- performance testing San Jose, CA
- performance coach San Jose, CA
- application performance engineer San Jose, CA
- lead performance test engineer San Jose, CA




