Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Director, AI Systems - Scale-up, Scale-out and Networking

Advanced Micro Devices

ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believe technology can change lives for the better. It can heal us, entertain us, and make us more connected, productive, and understanding of the world around us. And we’re looking for talent who feel the same: people who want to leave the planet better than they found it, those who don’t shy away from humanity’s challenges but are determined to help solve them. AMD is powering the next generation of supercomputing, high-performance computing, cloud, and AI. Whether you’re designing next-gen processors, enabling AI breakthroughs, or creating go-to-market plans, every role at AMD contributes to something bigger — technology that moves the world forward.

THE ROLE

We are seeking a seasoned technical leader to drive platform readiness and enablement for AMD's next-generation AI infrastructure and connectivity technologies. This Director-level role is responsible for shaping the strategy, execution, and cross-functional alignment required to bring industry-leading scale-up and scale-out solutions to market, with a focus on UA Link, PCI Express (PCIe), CXL, Ethernet, and high-performance AI/HPC systems. Working across silicon, firmware, software, platform engineering, validation, and ecosystem partners, you will lead the development and qualification of technologies that enable large-scale AI racks, accelerator clusters, and next-generation data center platforms. You will play a critical role in ensuring AMD delivers robust, scalable, and innovative connectivity solutions for CPUs, GPUs, accelerators, and AI infrastructure.

THE PERSON

You are an accomplished engineering leader with deep expertise in high-performance computing, AI infrastructure, and data center system architectures. You will lead a global cross functional engineering organization while fostering technical excellence, innovation, and strategic execution across AMD's AI infrastructure portfolio. You combine strong technical judgment with proven organizational leadership, enabling you to influence technical direction, align diverse engineering teams, and successfully execute complex product initiatives. You thrive in highly collaborative environments, excel at building relationships across organizations, and have a demonstrated ability to scale impact through both technical leadership and people leadership.

KEY RESPONSIBILITIES

Define and drive AMD's strategy for platform readiness, qualification, and ecosystem enablement of next-generation connectivity technologies including UA Link, PCIe, CXL, Ethernet, and future scale-up and scale-out architectures. Lead cross-functional engineering initiatives spanning silicon design, architecture, firmware, software, systems, validation, and product organizations to ensure successful product execution and market readiness. Provide technical and organizational leadership across multiple teams, guiding key architectural decisions, risk management, execution priorities, and long-term technology investments. Partner with internal engineering teams, industry consortiums, strategic customers, and ecosystem partners to accelerate adoption of AMD connectivity solutions and AI infrastructure platforms. Establish scalable validation and qualification strategies, ensuring readiness from pre-silicon development through system bring-up, production qualification, and customer deployment. Represent AMD in executive reviews, cross-functional technical forums, and strategic planning discussions, serving as a trusted technical leader for AI infrastructure and connectivity technologies.

PREFERRED EXPERIENCE

Extensive experience leading server, data center, AI, or HPC platform architecture, systems engineering, validation, or product development organizations. Deep understanding of modern connectivity and interconnect technologies including UA Link, PCIe Gen5/Gen6, CXL, Ethernet, high-speed SerDes, coherent fabrics, and large-scale distributed computing systems. Proven track record delivering complex products from architecture definition through silicon development, platform bring-up, qualification, customer enablement, and production launch. Experience collaborating across architecture, design, validation, firmware, software, manufacturing, and customer-facing organizations while influencing technical direction across multiple teams and stakeholders. Demonstrated success building, mentoring, and leading high-performing engineering teams, with the ability to drive execution through both direct leadership and organizational influence.

ACADEMIC CREDENTIALS

Bachelor's degree in Electrical Engineering, Computer Engineering, Computer Science, or related field. Master's degree, PhD, or equivalent industry experience preferred.

LOCATION

San Jose, CA This role is not eligible for visa sponsorship.

#LI-SL2

#LI-HYBRID

Benefits offered are described: AMD benefits at a glance. AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process. AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD's “Responsible AI Policy” is available here. This posting is for an existing vacancy. #J-18808-Ljbffr Advanced Micro Devices

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Director, AI Systems - Scale-up, Scale-out and Networking in San Jose, CA vacancy
  •  ...performance computing, cloud, and AI. Whether you’re designing...  ...technologies. This Director-level role is responsible for...  ...required to bring industry-leading scale-up and scale-out solutions to market, with a...  ...and high-performance AI/HPC systems.Working across silicon,... 
    Network

    AMD

    San Jose, CA
    2 days ago
  • $160k - $275k

     ...stack solutions from silicon to systems including hardware and...  ...MatX is building next-generation AI infrastructure systems for high...  ...looking for a Senior NPI TPM, Rack-Scale AI Systems to manage...  ...accelerator systems, datacenter racks, networking systems, storage systems, or... 
    Network
    Daily paid
    Full time
    Work experience placement
    Local area
    Remote work
    Monday to Friday
    Flexible hours

    MatX Inc.

    Mountain View, CA
    4 days ago
  • $272k - $431.25k

     ...technical focal point for rack-scale system SW/FW, working with CSP...  ...engineering teamsWays to stand out from the crowd:Experience with...  ...are increasingly known as “the AI computing company.” We're...  ...NVIDIA NVLink, NVIDIA InfiniBand networking, NVIDIA Grace CPUs, and a... 
    Network
    Full time
    Remote work
    Shift work

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...unlimited potential of AI to define the next era...  ...engineer for the Senior Systems Software Engineer role,...  ...on GPU Performance at Scale. At NVIDIA, this role is...  ...NVIDIA GPUs, CPUs, and networking hardware. Engage early...  ...experience.Ways to Stand Out From the CrowdEnd-to-... 
    Network
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $272k - $431.25k

     ...framework for serving generative AI and reasoning models across...  ...feel like a single system at datacenter scale. As large language models rapidly...  ...closely with GPU architecture, networking, and platform teams to...  ...customer teams.Ways to stand out from the crowd:Prior contributions... 
    Network
    Full time
    Local area
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $208k - $333.5k

     ...devices with various operating systems (Windows/Linux/Android), a...  ...requirements primarily Rack Scale AI Products.Finding Optimum Solutions...  ...and tools.Assist in roll-out and deployment of new development...  ...compute, Storage and networking.Ways to stand out from the crowd... 
    Network
    Full time
    Worldwide

    Nvidia

    Santa Clara, CA
    2 days ago
  • $184k - $287.5k

     ...unlimited potential of AI to define the next era of...  ...searching for a Senior Systems Software Engineer with deep...  ...problems at large scale and help shape how AI infrastructure...  ...such as GPU Operator, Network Operator, node-feature-...  ..., or OCI)Ways to stand out from the crowd:Strong... 
    Network
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    21 hours ago
  • $272k - $431.25k

     ...into the unlimited potential of AI to define the next era of...  ...NVIDIA, as a Principal Rack Scale Systems Infrastructure Engineer, you...  ...firmware, OS lifecycle, and networking fabrics. Your task is to compose...  ...systems.Expertise in in-band and out-of-band management... 
    Network
    Full time
    Remote work
    Shift work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $168k - $258.75k

     ...Manager (TPM) to join our Applied Systems Engineering Team to drive...  ...the next generation of NVIDIA AI supercomputing systems. This TPM...  ...lifecycle of the latest AI systems at scale, from datacenter design and...  ...instrumentationCoordinate design and fit-out of new datacenter builds,... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    4 days ago
  • $284.7k - $390.9k

     ...hyperscalers to deploy massive networks fueling AI clusters, general-purpose...  ...of networking at a scale that transforms industries...  ...You’ll Do:As a Senior Director for theHyperscaler Systems Architecture, you will lead...  ...systems roadmap for scale-out, scale-up, and scale-... 
    Network
    Full time
    Temporary work
    Work experience placement
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    3 days ago
  •  ...accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and...  ...Technical Program Manager (TPM) to lead Training at Scale Programs for AMD Instinct products. In this role, you... 

    AMD

    San Jose, CA
    21 hours ago
  • $320k

     ...unlimited potential of AI to define the next era...  ...From single node HGX/DGX systems all the way up to large...  ..., NVIDIA InfiniBand networking, NVIDIA Grace CPUs, and...  ...team responsible for rack-scale system software...  ...communications skillsWays to stand out from the crowd:... 
    Network
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    4 days ago
  • $152k - $241.5k

     ...Software Engineers to join our Fabric Networking team with a targeted focus on NVLink Rack-Scale Systems Stability & Reliability. In this...  ..., recovery, and large-scale AI infrastructure, contributing...  ...challenges at scale.Ways to stand out from the crowd:Experience with NVIDIA... 
    Network
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    4 days ago
  • Cerebras Systems builds the world’s largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras...  ...Cerebras, to deploy 750 megawatts of scale, transforming key workloads with...  ...that respects individual beliefs. Find out more about what it's like to work at... 

    Cerebras

    Sunnyvale, CA
    2 days ago
  • $232k - $320.75k

     ...Palo Alto Networks, Inc. is seeking a Director of Software Engineering to oversee the development of Cortex Platform teams in Santa...  ...extensive experience in cloud distributed systems and a strong ability to lead high-scale engineering organizations. Responsibilities... 
    Network

    Palo Alto Networks, Inc.

    Santa Clara, CA
    1 day ago
  • $240k - $333k

     ...and landing of next-gen ML TPU systems through NPI PLC process from...  ...infrastructure (e.g., TPU, GPU, or similar AI-compute clusters).Preferred...  ..., prioritize, resource, and scale programs according to...  ...for Google Cloud, Google Global Networking, Data Center operations, systems... 
    Network
    Worldwide

    Google

    Sunnyvale, CA
    1 day ago
  • $168k - $258.75k

     ...reduce risks in advanced PCB and system bring-up programs essential...  ...workflows, and processes to scale the team’s output.Run...  ...ERP platforms).Ways to stand out from the crowd:Recent experience...  ...activities on advanced compute, GPU, networking, or AI hardware platforms.Deep hands... 
    Network
    Full time
    Local area
    Shift work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $120k - $275k

     ...mission is to make the world's best AI models run as efficiently as...  ...-generation AI infrastructure systems for high-performance...  ...for a Senior Platform TPM, Rack-Scale AI Systems to drive design-phase...  ...accelerator systems, datacenter racks, networking systems, storage systems, or... 
    Network
    Full time
    Contract work
    Work experience placement
    Local area
    Remote work
    Monday to Friday
    Flexible hours

    MatX

    Mountain View, CA
    1 day ago
  • ## Sr. Director, Global IT NetworkingApplylocations...  ...you a visionary networking leader ready to...  ...automation, telemetry, and AI-driven network...  ...leadership in large-scale network...  ...an Internet-based system that compares information...  ...assistance in filling out the employment application... 
    Network
    Work experience placement
    Work at office
    Local area
    Immediate start
    Remote work
    Worldwide
    2 days per week

    Hewlett Packard Enterprise Development LP

    San Jose, CA
    1 day ago
  • $230k - $315k

     ...Credo is seeking a Senior Director, AI Interface Architecture to solve system-level interconnection challenges that connect XPUs...  ...role owns how compute elements are wired, networked, and orchestrated at node, rack, and cluster scale for efficient large-scale LLM training... 
    Network

    Jobleads-US

    San Jose, CA
    3 days ago
  • $168k - $258.75k

     ...capacity operations for large-scale AI infrastructure programs and...  ...access, storage, data movement, networking, readiness checks, migration...  ...AI/ML platforms, distributed systems, cloud infrastructure,...  ...collaborators!Ways to stand out from the crowd:Experience with... 
    Network
    Full time

    Nvidia

    Santa Clara, CA
    21 hours ago
  • $168k - $258.75k

     ...strategy for our Radio Access Network (RAN) Digital Twin...  ...portfolio. As 5G evolves toward AI-native 6G, the ability...  ..., covering end-to-end systems in both hardware and...  ...environmentWays to stand out from the crowd:Direct...  ...SDKs.Experience with large-scale simulations, large-scale... 
    Network
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • Etched, building at-scale AI inference supercomputers powered by its own chips, is seeking...  ...model for cluster-scale AI compute systems. This leader will own the end-to-end system...  ...orchestration, telemetry, provisioning, networking, to fleet reliability—working with ASIC,... 
    Network

    The Consensus

    San Jose, CA
    1 day ago
  • Payactiv in Milpitas, CA is seeking a Director, Revenue Operations to align Sales, Marketing, Channel, and Finance for predictable...  ...with Marketing Ops to improve lead handoffs and win rates while exploring AI opportunities to scale Revenue Ops. #J-18808-Ljbffr Payactiv

    Payactiv

    Milpitas, CA
    21 hours ago
  •  ...workloads. You will work across hardware, software, networking and infrastructure to deliver scalable, high-performance...  ...-generation products. The role requires strong Linux systems expertise, experience with large-scale compute infrastructure, and collaboration with cross-... 
    Network

    ASML

    San Jose, CA
    2 days ago
  • $208k - $327.75k

     ...operational future of Enterprise AI. While the NVIDIA DGX is...  ...deploy, manage, and scale their Enterprise AI...  ...-metal provisioning and network fabric configuration to...  ...the integration of DGX systems into the cloud-native ecosystem...  ...expands. Ways to Stand Out from the Crowd:... 
    Network
    Full time
    Night shift

    Nvidia

    Santa Clara, CA
    4 days ago
  • $264.9k - $358.3k

     ...-class silicon and systems IP for next-generation...  ...Technical Program Director to lead end-to-end...  ...leader in Arm’s AI transformation, you...  ...pivot into data center scale computing, AI/ML...  ...infrastructure (compute, networking, storage, power/...  ...talk to us to find out more about what... 
    Network
    Local area

    Arm

    San Jose, CA
    1 day ago
  • $185k - $215k

     ...experienced and hands‑on Director to lead the strategy,...  ...including distributed systems, development infrastructure...  ...adoption of generative AI capabilities across the...  ...record of building, scaling, and stabilizing complex...  ...expertise in Linux, networking, storage, and distributed... 
    Network
    Full time

    black.ai

    Milpitas, CA
    21 hours ago
  • Cerebras Systems is seeking an engineering leader to build and scale the Inference Model Scaling organization. You will define the technical vision and execution...  ...runtime, cloud, hardware, product management, and AI research to shape the future of AI inference at Cerebras... 

    Cerebras

    Sunnyvale, CA
    2 days ago
  • Vectra is the leader in AI-driven threat detection and response...  ...SaaS, identity, and data center networks in a single platform. Powered...  ...AI to move at the speed and scale of hybrid attackers. For more...  ...decisions for scalable AI/ML systems, data pipelines, model governance... 
    Network
    Worldwide

    Vectra AI

    San Jose, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Director, AI Systems - Scale-up, Scale-out and Networking. Be the first to apply!