Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Advanced AI Workloads - Performance and Scalability Engineer

Advanced Micro Devices Inc

WHAT YOU DO AT AMD CHANGES EVERYTHING At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career. THE ROLE: AMD is looking for a Principal Engineer to serve as a hands-on technical team lead driving the performance and scalability of frontier AI workloads on AMD GPUs, including large language models, mixture-of-experts architectures, and diffusion models. You will lead a team of engineers, define the long-term technical vision, make critical architecture decisions, and tackle the hardest performance challenges across the stack from GPU kernels and to serving frameworks and distributed systems.THE PERSON:The ideal candidate is a deep technical expert with a track record of solving industry-hard problems at the intersection of GPU architecture, AI systems, and high-performance software. You understand the full stack from hardware micro-architecture to model architecture, inference paradigms, and system-level design. You lead through technical depth, influence, and by example, staying hands-on while setting direction for your team. If you want to shape how the world runs AI on AMD hardware, this role is for you.KEY RESPONSIBILITIES:Lead a small team of engineers: set technical direction, prioritize work, and ensure delivery while remaining deeply hands-onDefine and drive the long-term technical strategy for AI workload performance on AMD GPUsOwn the most complex cross-stack performance challenges, from kernel optimization to framework-level architecture decisionsLead the design and implementation of novel GPU kernels, compiler optimizations, and framework featuresEstablish performance methodology and roofline analysis practices that set the standard for the teamInfluence upstream roadmaps in major open-source AI frameworks (e.g., vLLM, SGLang, PyTorch)Drive architecture decisions for emerging inference paradigms (e.g., prefill-decode disaggregation, speculative decoding, distributed serving)Identify and close fundamental performance gaps between AMD and competitor platformsServe as a technical authority across the organization, advising leadership on technical direction and feasibilityMentor engineers and raise the technical bar across the broader engineering organizationRepresent AMD externally through publications, conference talks, and open-source contributionsPREFERRED EXPERIENCE:Deep software development experience in GPU computing, HPC, or AI systemsDeep understanding of GPU micro-architecture, memory hierarchy, instruction scheduling, and performance tradeoffsDeep understanding of end-to-end AI systems: model architectures, inference paradigms, and system/rack-level designUnderstanding of multi-GPU communication: scale-up (NVLink, xGMI, Infinity Fabric) and scale-out (RDMA, RCCL/NCCL) topologies and performance characteristicsExperience designing and optimizing across the full stack: from low-level GPU kernels to frameworks and distributed serving systemsStrong background in performance engineering, including profiling, roofline analysis, and bottleneck diagnosis at scaleExperience with one or more of: HIP, CUDA, OpenCL, Triton/Gluon, CUTLASS, CKExperience with GPU compiler toolchains (e.g., LLVM) and intermediate representations (e.g., MLIR, LLVM IR, Triton IR) is a plusHands-on experience contributing to or architecting major open-source AI frameworks (e.g., vLLM, SGLang, xDiT, Megatron LM, PyTorch)Strong proficiency in C++ (C++17 or later) and PythonExperience leading small technical teams while remaining a hands-on contributorTrack record of influencing technical direction across teams and organizationsStrong Linux systems knowledgeExcellent written and verbal English communication skillsPublished research or significant open-source contributions in GPU computing, HPC, or AI systems is a plusPREFERRED ACADEMIC CREDENTIALS:Master's or PhD in Computer Science, Computer Engineering, Electrical Engineering, or equivalent. PhD strongly preferred.LOCATION:San Jose, CA preferredThis role is not eligible for visa sponsorship.#LI-G11#LI-HYBRIDBenefits offered are described: AMD benefits at a glance.AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.This posting is for an existing vacancy.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Advanced AI Workloads - Performance and Scalability Engineer in San Jose, CA vacancy
  • ADVANCE YOUR CAREER. ADVANCE THE WORLD. At AMD...  ...to powering AI and the technologies...  ...seeking an AI Systems Engineer to join our AMD IT...  ...of High-Performance Computing (HPC) infrastructure...  ...clusters, and AI workload schedulers. THE...  ...You want to build scalable and highly performant... 
    Performance

    AMD

    San Jose, CA
    2 days ago
  • $147k - $202.5k

     ...global leader in materials engineering solutions used to...  ...virtually every new chip and advanced display in the world....  ...our world - like AI and IoT. If you want...  ...collect and analyze data, perform hardware...  ...performance optimization and scalability. Collaborate with partner... 
    Performance
    Full time

    Applied Materials

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...unlimited potential of AI to define the next...  ...an AI Compiler Engineer with deep expertise...  ...measurable improvements in performance and efficiency, and advancing LLM-enabled...  ...on representative workloads and benchmark suites...  ...standards for reliability, scalability, and performance.... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    3 days ago
  • $136.3k - $199.9k

     ...About KLA We provide advanced inspection tools,...  ...customer requirements into scalable, maintainable software...  ...a strong emphasis on performance, scalability, and...  ...and interdisciplinary engineering teams to deliver integrated...  ...process. Use of AI Statement At KLA, our... 
    Performance
    Minimum wage
    Temporary work

    Jobleads-US

    Milpitas, CA
    1 day ago
  •  ...Technologist, Private Cloud AI - Applied & Agentic...  ...-to-cloud company advancing the way people live...  ..., ensuring performant, reliable, and trustworthy...  ...and partner with engineering to take POCs into scalable, production‑grade...  ...integrating AI workloads into production‑grade... 
    Performance
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start
    2 days per week

    Hewlett Packard Enterprise

    San Jose, CA
    3 days ago
  • $136.3k - $199.9k

     ...technologies like AI, data centers, automotive...  ...wafers, reticles, advanced packaging, process...  ...impact yield and performance.Behind these...  ...a highly skilled engineer to architect, develop...  ..., and support scalable Kubernetes environments...  ...critical to HPC workloads.Create and... 
    Performance
    Minimum wage
    Full time
    Work experience placement

    KLA-Tencor

    Milpitas, CA
    3 days ago
  •  ...Senior ASIC Verification Engineer (AI Hardware)...  ...quality, reliability, and performance of their next-generation...  ...verification strategies for advanced ASIC designs used in AI/compute workloads Build and execute detailed...  ...approaches Develop scalable verification... 
    Performance
    Remote work
    Visa sponsorship

    4 Staffing Corp

    San Jose, CA
    3 days ago
  • $2,000 per month

     ...both prefill and decode workloads. Our first products...  ...staffed by leading engineers, Etched is redefining...  ...quality assurance for our advanced AI hardware builds—...  ...Etched’s standards for performance and reliability....  ...also establishing the scalable quality processes needed... 
    Performance
    Contract work
    Work at office
    Overseas
    Relocation package

    The Consensus

    San Jose, CA
    1 day ago
  •  ...ADVANCE YOUR CAREER. ADVANCE THE WORLD. At AMD,...  ...discovery to powering AI and the technologies...  ...AI Training Systems & Performance Engineering PhD Intern/Co-Op to...  ...cutting-edge AI training workloads on AMD Instinct™ GPUs...  ...that improves scalability, efficiency, and developer... 
    Performance
    Full time
    Summer work
    Internship
    Summer internship
    Worldwide

    AMD

    San Jose, CA
    11 hours ago
  • ADVANCE YOUR CAREER. ADVANCE THE WORLD. At AMD...  ...to powering AI and the technologies...  ...Senior AI Software Engineer to join our team....  ...computing, and system performance profiling within...  ...applications, ensuring scalability, low latency, and...  ...and training workloads.System Performance... 
    Performance

    AMD

    San Jose, CA
    4 days ago
  • $124k - $208.4k

     ...SummarySamsung, a world leader in advanced semiconductor...  ...is applied to high-performance computing devices (...  ...execution of AI workloads on Samsung’s premium...  ...performance, efficiency, and scalability across a variety of ML...  ...architects, software engineers, and hardware teams... 
    Performance
    Hourly pay
    Full time
    Relocation

    Samsung Semiconductor

    San Jose, CA
    11 hours ago
  • $152k - $241.5k

     ...experienced software engineers with kubernetes...  ...help scale up its AI Infrastructure. We...  ...better. You will help advance NVIDIA's capacity...  ...that enable large scalable GPU clusters to be...  ...a variety of AI workloads. This includes...  ...consistently with maximum performance. Evaluating system... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    11 hours ago
  • $138k - $206k

     ...jobs within a 6-month period. Advancing the World’s Technology...  ...Title Senior Physical Design AI/ML Engineer, Logic Pathfinding LabWhat...  ...throughput in predicting design performance, power, and area (PPA)....  ...qualificationsFamiliarity with state-of-the-art AI workloads and their compute and... 
    Performance
    Work at office
    Local area
    Flexible hours

    Samsung Semiconductor

    San Jose, CA
    4 days ago
  • $141.15k - $211.4k

     ...Across enterprise, cloud and AI, and carrier architectures, our...  ...power the next generation of performance‑driven systems.Within this ecosystem, Marvell’s central advanced packaging organization plays...  ...Can ExpectMarvell's Central Engineering team is seeking a highly motivated... 
    Performance
    Permanent employment
    Internship
    Work from home

    Marvell

    Santa Clara, CA
    11 hours ago
  • $184.5k - $249.6k

     ...Principal FAE - Edge AI' to support strategic...  ...relationships with architects, engineering leaders, and key...  ...to meet power, performance, scalability, and time-to-market objectives...  ..., system IP, or advanced SoC architectures....  ...solutions, including AI/ML workloads on embedded and edge... 
    Performance
    Work at office
    Local area

    ARM

    San Jose, CA
    2 days ago
  • $184k - $287.5k

     ...Senior Developer Technology Engineer!NVIDIA's Developer...  ...optimizing large application workloads, eliminating system...  ...platform including advanced CPUs, GPUs and interconnects...  ...with key customers to perform in-depth analysis and...  ...existing vacancy. NVIDIA uses AI tools in its recruiting... 
    Performance
    Full time
    Work experience placement
    Remote work

    Nvidia

    Santa Clara, CA
    3 days ago
  • $266.05k - $396k

     ...potential of their data, from AI to multicloud....  ...Summary Distinguished Engineer - AI Infrastructure...  ...experience with high-performance inference engines (...  ...engines optimized for AI workloads (checkpoints, model...  ...system design for highly scalable distributed storage... 
    Performance
    Work at office
    Local area

    NetApp

    San Jose, CA
    1 day ago
  • $184k - $287.5k

     ...seeking a Senior Developer Technology Engineer, Artificial Intelligence! Would you...  ...researching parallel algorithms to accelerate AI workloads on advanced computer architectures? Is it...  ...bottlenecks to achieve the best possible performance of computer hardware? Could you be... 
    Performance
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    3 days ago
  •  ...San Jose leads research on scalable AI/ML workloads. The team designs systems...  ...on energy efficiency and performance. The role requires extensive...  ...experience in performance engineering, AI systems, distributed...  .... Join a world-class team advancing AGI-enabled computing in a... 
    Performance

    Jobleads-US

    San Jose, CA
    3 days ago
  • $224k - $356.5k

     ...exceptional hands-on engineer to help our partners and...  ...robotics simulation workloads run exceptionally well...  ...someone who merely used advanced systems, but someone...  ...establish and execute performance strategies for important...  ...vacancy. NVIDIA uses AI tools in its recruiting... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $152k - $241.5k

     ...Senior Developer Technology Engineer, CPU Performance!Would you enjoy researching...  ...on NVIDIA’s family of advanced CPU platforms.Work directly...  ...database and data analytics workloads to ensure the best possible...  ...existing vacancy. NVIDIA uses AI tools in its recruiting processes... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    3 days ago
  • ADVANCE YOUR CAREER. ADVANCE THE WORLD. At AMD, we believe...  ...discovery to powering AI and the technologies...  ...seeking a motivated Quality Engineer to join our team and...  ...workflows using scalable processes through innovative...  ...Investigate non-conformances, perform root cause analysis,... 
    Performance

    AMD

    San Jose, CA
    4 days ago
  • $285k - $355k

     ...for a Distinguished Engineer (DE) reporting into...  ...excellence, and designing scalable architectures. It is...  ..., reliability, performance, and scalability), developer...  ...of the most advanced, most difficult, most...  ...from traditional workloads to enterprise AI with unmatched performance... 
    Performance
    Work at office
    Local area

    NetApp

    San Jose, CA
    11 hours ago
  • $152k - $228k

     ...AstraZeneca achieve 20x performance results and...  ...The Platform Engineering team builds, secures...  ...secures, and operates scalable infrastructure...  ..., and creating AI Ops workflows for...  ...vulnerability detection, and workload protection across...  ...opportunity, and advancement for all while... 
    Performance
    Permanent employment
    Contract work
    Work at office
    Remote work

    NightDragon Acquisition Corp.

    Santa Clara, CA
    1 day ago
  •  ...Advanced Micro Devices, Inc. is seeking a Principal or Fellow level software engineer to advance AI infrastructure, focusing on performance, reliability and scalable workloads for model training and inference. You will collaborate with customers and cross‑functional... 
    Performance

    Jobleads-US

    San Jose, CA
    2 days ago
  •  ...building the future of AI powered digital...  ...Unit (TCU) that combines advanced hardware security, AI...  ...takes more than great engineering—it takes a team of exceptional...  ...Senior QA Engineer - Performance & Reliability to lead...  ...of TCU/BMC systems.Workload Analysis: Analyze system... 
    Performance

    Axiado Corporation

    San Jose, CA
    3 days ago
  • $195k - $220k

     ...Center Operations Engineer Location: San Jose...  ...Tier provider of advanced server, storage, and...  ...to meet scalability, availability, and performance requirements, ensure...  ...to meet demanding workload requirements Direct...  ...~ Familiarity with AI/ML infrastructure requirements... 
    Performance
    Worldwide

    Jobleads-US

    San Jose, CA
    1 day ago
  • $126.94k - $174.58k

     ...combines analog, digital, AI, and software...  ...world, and help drive advancements in automation and robotics...  ...regulators deliver on-demand scalable power with the speed,...  ...experienced Product Engineer with strong data...  ...qualifies for a discretionary performance-based bonus which is... 
    Performance
    Permanent employment
    Full time
    Work at office
    Day shift

    Analog Devices

    Milpitas, CA
    3 days ago
  • ADVANCE YOUR CAREER. ADVANCE THE WORLD. At AMD, we believe...  ...discovery to powering AI and the technologies...  ...Photonic Systems Engineer with deep expertise and...  ...behavior, and optimize performance across complex...  ...for building accurate, scalable models that capture optical... 
    Performance

    AMD

    San Jose, CA
    1 day ago
  • $232k - $406k

     ...to learn, communicate and advance faster than ever.The DASG...  ...accelerate technology pathfinding, AI and system workload analysis, memory subsystem...  ..., and teamwork across engineering organizations and external...  ..., DRAM replay, power and performance characterization, and data... 
    Performance
    Full time
    Local area
    Immediate start
    Remote work

    Micron

    San Jose, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Advanced AI Workloads - Performance and Scalability Engineer. Be the first to apply!