Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Advanced AI Workloads - Performance and Scalability Engineer

Advanced Micro Devices Inc

WHAT YOU DO AT AMD CHANGES EVERYTHING At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career. THE ROLE: AMD is looking for a Principal Engineer to serve as a hands-on technical team lead driving the performance and scalability of frontier AI workloads on AMD GPUs, including large language models, mixture-of-experts architectures, and diffusion models. You will lead a team of engineers, define the long-term technical vision, make critical architecture decisions, and tackle the hardest performance challenges across the stack from GPU kernels and to serving frameworks and distributed systems.THE PERSON:The ideal candidate is a deep technical expert with a track record of solving industry-hard problems at the intersection of GPU architecture, AI systems, and high-performance software. You understand the full stack from hardware micro-architecture to model architecture, inference paradigms, and system-level design. You lead through technical depth, influence, and by example, staying hands-on while setting direction for your team. If you want to shape how the world runs AI on AMD hardware, this role is for you.KEY RESPONSIBILITIES:Lead a small team of engineers: set technical direction, prioritize work, and ensure delivery while remaining deeply hands-onDefine and drive the long-term technical strategy for AI workload performance on AMD GPUsOwn the most complex cross-stack performance challenges, from kernel optimization to framework-level architecture decisionsLead the design and implementation of novel GPU kernels, compiler optimizations, and framework featuresEstablish performance methodology and roofline analysis practices that set the standard for the teamInfluence upstream roadmaps in major open-source AI frameworks (e.g., vLLM, SGLang, PyTorch)Drive architecture decisions for emerging inference paradigms (e.g., prefill-decode disaggregation, speculative decoding, distributed serving)Identify and close fundamental performance gaps between AMD and competitor platformsServe as a technical authority across the organization, advising leadership on technical direction and feasibilityMentor engineers and raise the technical bar across the broader engineering organizationRepresent AMD externally through publications, conference talks, and open-source contributionsPREFERRED EXPERIENCE:Deep software development experience in GPU computing, HPC, or AI systemsDeep understanding of GPU micro-architecture, memory hierarchy, instruction scheduling, and performance tradeoffsDeep understanding of end-to-end AI systems: model architectures, inference paradigms, and system/rack-level designUnderstanding of multi-GPU communication: scale-up (NVLink, xGMI, Infinity Fabric) and scale-out (RDMA, RCCL/NCCL) topologies and performance characteristicsExperience designing and optimizing across the full stack: from low-level GPU kernels to frameworks and distributed serving systemsStrong background in performance engineering, including profiling, roofline analysis, and bottleneck diagnosis at scaleExperience with one or more of: HIP, CUDA, OpenCL, Triton/Gluon, CUTLASS, CKExperience with GPU compiler toolchains (e.g., LLVM) and intermediate representations (e.g., MLIR, LLVM IR, Triton IR) is a plusHands-on experience contributing to or architecting major open-source AI frameworks (e.g., vLLM, SGLang, xDiT, Megatron LM, PyTorch)Strong proficiency in C++ (C++17 or later) and PythonExperience leading small technical teams while remaining a hands-on contributorTrack record of influencing technical direction across teams and organizationsStrong Linux systems knowledgeExcellent written and verbal English communication skillsPublished research or significant open-source contributions in GPU computing, HPC, or AI systems is a plusPREFERRED ACADEMIC CREDENTIALS:Master's or PhD in Computer Science, Computer Engineering, Electrical Engineering, or equivalent. PhD strongly preferred.LOCATION:San Jose, CA preferredThis role is not eligible for visa sponsorship.#LI-G11#LI-HYBRIDBenefits offered are described: AMD benefits at a glance.AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.This posting is for an existing vacancy.

Vacancy posted 4 hours ago
Similar jobs that could be interesting for youBased on the Advanced AI Workloads - Performance and Scalability Engineer in San Jose, CA vacancy
  • $147k - $202.5k

     ...global leader in materials engineering solutions used to...  ...virtually every new chip and advanced display in the world....  ...our world - like AI and IoT. If you want...  ...collect and analyze data, perform hardware...  ...performance optimization and scalability. Collaborate with partner... 
    Performance
    Full time

    Applied Materials

    Santa Clara, CA
    4 hours ago
  • $184k - $287.5k

     ...unlimited potential of AI to define the next...  ...an AI Compiler Engineer with deep expertise...  ...measurable improvements in performance and efficiency, and advancing LLM-enabled...  ...on representative workloads and benchmark suites...  ...standards for reliability, scalability, and performance.... 
    Performance
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    4 hours ago
  •  ...innovations in Flash and advanced memory technologies, our...  ...globally for innovation, performance and quality. Sandisk has...  ...Description Focus:  AI workload-driven product...  ...validation Partner with Test Engineering to deliver scalable and robust production test... 
    Performance
    Temporary work
    Remote work
    Flexible hours
    Shift work

    Sandisk

    Milpitas, CA
    22 days ago
  • $170k - $412.5k

     ...Technologist, Private Cloud AI - Applied & Agentic...  ...-to-cloud company advancing the way people live...  ..., ensuring performant, reliable, and trustworthy...  ...and partner with engineering to take POCs into scalable, production‑grade...  ...integrating AI workloads into production‑grade... 
    Performance
    Full time
    Work experience placement
    Work at office
    Local area
    Immediate start
    2 days per week

    Hewlett Packard Enterprise

    San Jose, CA
    2 days ago
  •  ...innovations in Flash and advanced memory...  ...for innovation, performance and quality. Sandisk...  ...establish and lead an AI Systems &...  ...across real‑world workloads, software stacks,...  ...reproducibility, and scalability. Establish best...  ...skilled team of engineers and researchers.... 
    Performance
    Temporary work
    Remote work
    Flexible hours
    Shift work

    Sandisk

    Milpitas, CA
    5 days ago
  • ADVANCE YOUR CAREER. ADVANCE THE WORLD.At AMD,...  ...supercomputing, high-performance computing, cloud, and AI. Whether you’re...  ...seeking an AI Systems Engineer to join our AMD IT...  ...GPU clusters, and AI workload schedulers. THE...  ...You want to build scalable and highly performant... 
    Performance

    AMD

    San Jose, CA
    1 day ago
  • $124k - $208.4k

     ...SummarySamsung, a world leader in advanced semiconductor...  ...is applied to high-performance computing devices (...  ...execution of AI workloads on Samsung’s premium...  ...performance, efficiency, and scalability across a variety of ML...  ...architects, software engineers, and hardware teams... 
    Performance
    Hourly pay
    Full time
    Relocation

    Samsung Semiconductor

    San Jose, CA
    2 days ago
  • $122.6k - $185k

     ...backbone of Generative AI cloud at AWS? Do...  ...continuous price performance improvements in the...  ...high performance and scalability in AI/ML and HPC workloads.You are intrigued...  .... The AWS Hardware Engineering team creates server...  ...mentorship and other career-advancing resources here to... 
    Performance
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    3 days ago
  • $266.05k - $396k

     ...potential of their data, from AI to multicloud....  ...Summary Distinguished Engineer - AI Infrastructure...  ...experience with high-performance inference engines (TensorRT...  ...optimized for AI workloads (checkpoints, model artifacts...  ...design for highly scalable distributed storage... 
    Performance
    Part time
    Work at office
    Local area

    NetApp

    San Jose, CA
    1 day ago
  • $2,000 per month

     ...both prefill and decode workloads. Our first products...  ...staffed by leading engineers, Etched is redefining...  ...quality assurance for our advanced AI hardware builds—...  ...Etched’s standards for performance and reliability....  ...also establishing the scalable quality processes needed... 
    Performance
    Contract work
    Work at office
    Overseas
    Relocation package

    Etched

    San Jose, CA
    13 days ago
  • $138k - $206k

     ...jobs within a 6-month period. Advancing the World’s Technology...  ...Title Senior Physical Design AI/ML Engineer, Logic Pathfinding LabWhat...  ...throughput in predicting design performance, power, and area (PPA)....  ...qualificationsFamiliarity with state-of-the-art AI workloads and their compute and... 
    Performance
    Work at office
    Local area
    Flexible hours

    Samsung Semiconductor

    San Jose, CA
    4 hours ago
  • $184.5k - $249.6k

     ...Principal FAE - Edge AI' to support strategic...  ...relationships with architects, engineering leaders, and key...  ...to meet power, performance, scalability, and time-to-market objectives...  ..., system IP, or advanced SoC architectures....  ...solutions, including AI/ML workloads on embedded and edge... 
    Performance
    Work at office
    Local area

    ARM

    San Jose, CA
    4 days ago
  •  ...experiences—from AI and data centers,...  ...beyond. Together, we advance your career. THE...  ...influential software engineer who is passionate...  ...improving the performance of key applications...  ...the most demanding workloads.Innovate Across Hardware...  ...complex, scalable systems using modern... 
    Performance

    AMD

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

     ...seeking a Senior Developer Technology Engineer, Artificial Intelligence! Would you...  ...researching parallel algorithms to accelerate AI workloads on advanced computer architectures? Is it...  ...bottlenecks to achieve the best possible performance of computer hardware? Could you be... 
    Performance
    Full time
    Work experience placement

    Nvidia

    Santa Clara, CA
    4 hours ago
  • $184k - $287.5k

     ...Senior Developer Technology Engineer!NVIDIA's Developer...  ...optimizing large application workloads, eliminating system...  ...platform including advanced CPUs, GPUs and interconnects...  ...with key customers to perform in-depth analysis and...  ...existing vacancy. NVIDIA uses AI tools in its recruiting... 
    Performance
    Full time
    Work experience placement
    Remote work

    Nvidia

    Santa Clara, CA
    4 hours ago
  •  ...building the future of AI powered digital...  ...Unit (TCU) that combines advanced hardware security, AI...  ...takes more than great engineering—it takes a team of exceptional...  ...Senior QA Engineer - Performance & Reliability to lead...  ...of TCU/BMC systems.Workload Analysis: Analyze system... 
    Performance

    Axiado Corporation

    San Jose, CA
    4 days ago
  • $240k

     ...world's largest AI chip, 56 times larger...  ...transforming key workloads with ultra high-...  ...regression, and performance test strategies...  ...Kubernetes to validate scalable SaaS deployments....  ...tools, and engineering best practices.Document...  ...groundbreaking advancements in AI!Cerebras... 
    Performance

    Cerebras Systems

    Sunnyvale, CA
    1 day ago
  •  ...in Sunnyvale, CA, seeks a Senior Software Engineer for Identity Management Services. Join a high-impact team building scalable identity workflows across Apple platforms with...  ...deploying large-scale systems, improving performance, and collaborating with cross-functional partners... 
    Performance

    Apple Inc.

    Sunnyvale, CA
    2 days ago
  • $126.94k - $174.58k

     ...combines analog, digital, AI, and software...  ...world, and help drive advancements in automation and robotics...  ...regulators deliver on-demand scalable power with the speed,...  ...experienced Product Engineer with strong data...  ...qualifies for a discretionary performance-based bonus which is... 
    Performance
    Permanent employment
    Full time
    Work at office
    Day shift

    Analog Devices

    Milpitas, CA
    4 hours ago
  • $120k - $200k

     ...Site Reliability Engineer (Prisma Access) 2...  ...highly reliable, scalable, and secure cloud...  ...systems are robust and performant. This includes...  ...the adoption of AI tools. Develop...  ...networking, and container workloads. Strong Linux...  ...threats, and advancements through... 
    Performance
    Rotating shift

    Palo Alto Networks

    Santa Clara, CA
    5 days ago
  •  ...Description About us nEye.ai, a well-funded optical...  ...hyperscale data centers enhanced performance, efficiency, and scalability. Job Overview We are...  ...Systems Packaging & Assembly Engineer to lead the development and integration of advanced packaging and assembly... 
    Performance
    Contract work

    nEye.ai

    Santa Clara, CA
    13 days ago
  •  ...revolutionizing power for AI-driven data centers to...  ...Process Technology Engineer to join our team in...  ...project execution to ensure performance and design compliance....  ...development and advancing system integration while...  ...enable cost-effective, scalable deployment across... 
    Performance
    Full time
    For contractors
    Work at office
    Worldwide

    Bloom Energy

    San Jose, CA
    4 days ago
  • $147k - $202.5k

     ...global leader in materials engineering solutions used to...  ...virtually every new chip and advanced display in the world....  ...our world - like AI and IoT. If you want...  ...collect and analyze data, perform hardware...  ...improve device efficiency, scalability, and manufacturability... 
    Performance
    Full time

    Applied Materials

    Santa Clara, CA
    1 day ago
  • $161k - $221k

     ...global leader in materials engineering solutions used to...  ...virtually every new chip and advanced display in the world....  ...our world - like AI and IoT. If you want...  ...smart glasses and high‑performance optical interconnects....  ...will architect robust, scalable integration flows that... 
    Performance
    Full time

    Applied Materials

    Santa Clara, CA
    3 days ago
  • $220.92k - $311.89k

     ...Description: Intel’s AI SoC organization is driving...  ...Formal Verification Engineer, you will play a...  ...digital designs using advanced formal methods. This position...  ...and develop scalable, reusable verification...  ...design, analyze power/performance, and uncover bugs.Debug... 
    Performance
    Full time
    Internship
    Local area
    Immediate start
    Shift work

    Intel

    Santa Clara, CA
    4 hours ago
  • $117k - $234k

     ...organization, designs, engineers, and operates the...  ...services, and Kubernetes workloads at multi-petabyte...  ...improving reliability and performance, and enabling self-...  ..., observability, and AI/AIOps-driven...  ...workloads while continuously advancing scalability, resiliency, utilization... 
    Performance
    Full time
    Temporary work
    Part time

    Walmart

    Sunnyvale, CA
    1 day ago
  •  ...Job TitleGenerative AI continues to accelerate...  ...redefining interconnect performance with dramatically...  ...efficiency, and greater scalability. In this role, you will...  ...support of Lumilens' advanced photonics platform.Translate...  ..., work instructions, engineering reports, and training... 
    Performance
    Shift work

    Lumilens

    San Jose, CA
    1 day ago
  •  ...computing experiences—from AI and data centers, to...  ...beyond. Together, we advance your career. THE ROLE:...  ...Systems Design Engineer to join our growing team...  ...will drive balanced, scalable, and automated solutions...  ...accurate resultsOptimize the performance, such as fusion and... 
    Performance

    AMD

    San Jose, CA
    3 days ago
  • $168k - $270.25k

     ...Experience (NVEX) Solutions Engineering team is looking for an...  ...will apply the latest AI technologies to triage...  ...issues and AI/ML workloads in huge datacenters of...  .... Expertise analyzing performance of distributed GPU-...  ...customers to resolve or advance customer issues.Work with... 
    Performance
    Full time
    Weekend work

    Nvidia

    Santa Clara, CA
    4 hours ago
  • $124k - $195.5k

     ...continue to grow, the intersection of advanced deep learning architectures,...  ...team to help unlock maximum hardware performance for emerging AI workloads. You will be a crucial member of a...  ...degree in Computer Science, Computer Engineering, Electrical Engineering, or related... 
    Performance
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Advanced AI Workloads - Performance and Scalability Engineer. Be the first to apply!