Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Datacenter Technologist

Jobleads-US

Graphcore is a leading innovator in artificial intelligence computing. We develop hardware, software, and data center infrastructure that provide the specialized processing and systems capabilities needed to advance AI while improving the efficiency required for broad adoption.

As part of SoftBank Group, Graphcore works alongside companies developing advanced technologies. Our teams bring together AI researchers, silicon designers, hardware and software engineers, and systems architects to solve complex technical problems across the computing stack.

The Opportunity

As Principal Datacenter Technologist, you will turn new platform architectures into repeatable operating models for Graphcore's AI data centers. You will lead the technical definition of how Graphcore introduces, deploys, services, and sustains new compute, networking, storage, memory, power, and cooling technologies at scale.

Working within Advanced Architecture, you will connect platform design with data center engineering and site reliability operations. You will define new-product-introduction workflows, readiness criteria, diagnostics, and cross-functional handoffs so that new systems can move from architecture through deployment with clear ownership, supportability, and operational feedback.

What You Will Do

  • Own the technical operating model and new-product-introduction framework for new hardware and infrastructure technologies entering Graphcore data centers.
  • Define end‑to‑end workflows for installation, configuration, validation, deployment, service, repair, upgrade, and sustained operation across the product lifecycle.
  • Translate architecture requirements into data center readiness criteria covering software, firmware, compute, networking, storage, memory, rack integration, power, cooling, space, and site operations.
  • Create clear runbooks, interface definitions, ownership models, acceptance criteria, and escalation paths that align architecture, engineering, data center, and site reliability teams.
  • Lead operational readiness reviews and identify gaps in tooling, diagnostics, telemetry, serviceability, spares, documentation, and training before deployment.
  • Serve as a senior escalation point for complex platform issues, ensuring that teams collect the right logs, telemetry, failure evidence, and environmental data to support root‑cause analysis and disposition.
  • Define and guide the deployment of automated diagnostic and telemetry systems that identify systemic hardware, software, quality, handling, and site‑integration issues across the fleet.
  • Establish quality and operational metrics for new technologies, analyze field trends, and feed findings back into architecture, design, supplier, and deployment decisions.
  • Coordinate cross‑functional resolution of issues spanning hardware, firmware, software, networking, storage, facilities, and site operations, with clear owners, decisions, and follow‑through.
  • Provide technical leadership, review implementation plans, mentor engineers, and communicate readiness, risks, tradeoffs, and recommendations to engineering and operations leaders.

What You Will Bring

  • A bachelor's degree or equivalent experience in information technology, computer science, computer engineering, electrical engineering, or a related field, or equivalent practical experience.
  • Extensive experience in data center infrastructure build, infrastructure operations, platform deployment, or related technical fields, including leadership at Principal, Staff, or Lead Engineer scope.
  • Demonstrated experience defining processes, workflows, and readiness criteria for new‑product introduction in large‑scale data center or infrastructure environments.
  • Broad systems knowledge across compute, networking, storage, memory, firmware, operating systems, racks, power distribution, cooling, and data center operations.
  • Hands‑on Linux experience, including scripting or automation used to collect diagnostics, telemetry, configuration, and health information.
  • Experience with failure analysis and validation of AI compute, server, networking, or other complex data center equipment.
  • Ability to define telemetry, diagnostic coverage, quality metrics, and fleet‑level signals that distinguish product issues from deployment, handling, or environmental problems.
  • Proven ability to lead multidisciplinary technical work across architecture, hardware, software, facilities, data center engineering, and site reliability teams without relying on direct authority.
  • Clear written and verbal communication skills, including the ability to produce precise technical workflows and explain operational risks and decisions to engineering leaders.

Preferred Qualifications

These qualifications are helpful, not required. We encourage you to apply even if you do not meet every preferred qualification.

  • Experience leading data center engineering, platform introduction, or infrastructure readiness initiatives.
  • Experience operating or supporting high‑density AI or high‑performance computing systems at fleet scale.
  • Experience with telemetry platforms, automated diagnostics, serviceability tooling, and quality metrics for compute and networking equipment.
  • Experience with rack‑as‑a‑system architectures, direct liquid cooling, high‑power racks, and the operational controls required for safe deployment and service.
  • Experience working with suppliers, field‑service teams, manufacturing, or logistics partners on equipment lifecycle, quality, repair, or failure‑disposition processes.

Graphcore offers compensation and benefits designed to support employees' health, financial well‑being, work‑life needs, and professional growth. Benefits and programs for eligible U.S. employees may include:

  • Medical, dental, and vision coverage, with options that may extend to eligible dependents.
  • Mental health, wellness, and employee assistance resources.
  • Retirement savings benefits and company contributions where applicable.
  • Paid vacation, sick time, company holidays, and parental or family leave in accordance with applicable plans and policies.
  • Life insurance and short‑term or long‑term disability coverage.
  • Flexible working hours and hybrid working arrangements where compatible with the role and team requirements.
  • Professional development resources, learning programs, office amenities, and team‑led activities.

Benefits vary by work location, employment status, scheduled hours, and plan eligibility and are subject to the terms of the applicable plans and company policies. This overview is not a contract or guarantee of benefits.

Equal Opportunity and Accommodations

Graphcore is an equal opportunity employer. We consider qualified applicants without regard to race, color, religion, creed, sex, pregnancy, sexual orientation, gender identity or expression, national origin, ancestry, age, disability, genetic information, veteran status, or any other status protected by applicable law.

Graphcore is committed to an inclusive and accessible hiring process. If you need a reasonable accommodation to participate in the application or interview process, please let the recruiting team know.

Candidate Privacy

Personal information submitted during the recruiting process will be handled in accordance with Graphcore's applicable candidate privacy notices.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 10 hours ago
Similar jobs that could be interesting for youBased on the Principal Datacenter Technologist in Milpitas, CA vacancy
  •  ...architects. Job Summary We are looking for an experienced Principal Engineer to join our System Management team and help lead the...  ...Collaborate with Hardware, Firmware, Platform Software and Datacenter Operations teams to diagnose system-level issues and improve end... 
    Principal
    Flexible hours

    Graphcore

    Milpitas, CA
    3 days ago
  •  ...Principal Network EngineerGraphcore is a globally recognized leader in Artificial Intelligence computing systems. The company designs...  ...fabrics.Primary support of deployed internal fabrics in our AI Datacenters.Direct escalation for Hardware, Link and, Performance issues on... 
    Principal
    Remote work
    Flexible hours

    Graphcore

    Milpitas, CA
    3 days ago
  •  ...We are seeking an experienced Principal Network Engineer with 6+ years of hands-on experience in data center networking, routing, switching, and network architecture. The ideal candidate will have strong expertise with HPE Juniper networking platforms , including QFX... 
    Principal
    Contract work

    PB consulting

    Alviso, CA
    11 days ago
  • $182k - $319k

     ...for our customers.Let’s engineer the future together.The Senior Principal, Design Engineering will be responsible for architecting,...  ...specification, design, and debug of complex power delivery systems for datacenter products.Skills & Responsibilities:Engage with customers and... 
    Principal
    Local area

    Celestica

    San Jose, CA
    3 days ago
  •  ...Ll Oefentherapie is seeking a seasoned Program Manager to own and deliver GPU infrastructure programs, spanning datacenter enablement, server provisioning, deployment, and lifecycle operations. You will align priorities across engineering, networking, datacenter ops... 
    Principal

    Jobleads-US

    Santa Clara, CA
    2 days ago
  • $240k - $379.5k

     ...a company that lets us do our life’s work, at the highest level of our craft. Join our new team building the next generation of datacenter infrastructure that powers the future of AI. We are seeking highly motivated technical experts to join our new organization leading... 
    Principal
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  •  ...What You’ll Do Own and deliver GPU infrastructure programs spanning datacenter enablement, GPU server provisioning, deployment, configuration, and lifecycle operations. Align priorities and execution across multiple teams (engineering, networking, datacenter ops... 
    Principal

    Jobleads-US

    Santa Clara, CA
    2 days ago
  • $148k - $308k

     ...system architecture come together to shape the future of AI and datacenter platforms. We work at the intersection of silicon, systems,...  ...computing infrastructure, you’ll find that work here matters!As a Sr. Principal Solutions Architect for Network Interconnect Technology, you... 
    Principal
    Full time
    Work at office
    Local area
    Immediate start

    Micron

    San Jose, CA
    3 days ago
  •  ...leading model builders such as OpenAI and other frontier labs.As a Principal SRE, you will define and drive the technical architecture for...  ...and running software reliably and at scale across multiple datacenters and cloud-based solutions.Architect self-service platforms and... 
    Principal
    Shift work

    Cerebras Systems

    Sunnyvale, CA
    20 hours ago
  • $196k - $300k

     ...our Santa Clara, Ca headquarters 3-5 days per week.The Role: Principal Technical Program Manager, HW/SystemsWe are seeking a Principal...  ...chip-to-ship execution - post-silicon to product.Background in datacenter hardware development companies.Experience scaling engineering... 
    Principal
    3 days per week

    d-Matrix

    Santa Clara, CA
    2 days ago
  • $232k - $368k

     ...spans early architecture through final product delivery across Datacenter, Gaming, Robotics, Automotive, and Embedded markets. We work...  ...crowd:Prior experience as a Chip Lead, Project Tech Lead, or Principal or Distinguished Engineer on a flagship GPU, AI accelerator, CPU... 
    Principal
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $187k - $260k

    DescriptionAbout the Role Power Integrations is seeking an ambitious, highly motivated Principal AI/Data Center Systems Engineer to help define and develop next-generation high-density power solutions for AI and data center applications. In this role, you will contribute... 
    Principal
    Worldwide
    Shift work

    Power Integrations

    San Jose, CA
    2 days ago
  • $165k - $240k

     ...global connectivity.Molex is seeking a highly skilled Senior or Principal Firmware Engineer to join our Optical Systems Business Unit in...  ...OSFP and QSFP-DD form factors at 800G and 1.6T serving datacenter and telecom markets. The ideal candidate brings deep expertise... 
    Principal
    Long distance
    Flexible hours

    Molex

    Fremont, CA
    3 days ago
  • $272k - $431.25k

     ...clusters so that many accelerators feel like a single system at datacenter scale. As large language models rapidly outgrow the memory and...  ...deployment of cutting-edge LLM workloads.We are seeking a Principal Systems Engineer to define the vision and roadmap for memory management... 
    Principal
    Full time
    Local area
    Remote work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $154.68k - $231.7k

     ...position with direct influence on customer success and revenue.As a Principal Applications Engineer, you will:Serve as the primary technical...  ...Marvell’s Optical DSP solutions into next-generation AI and datacenter platformsDrive end-to-end system success, from early design-in... 
    Principal
    Permanent employment
    Internship
    Immediate start
    Work from home

    Marvell

    Santa Clara, CA
    1 day ago
  • $272k - $431.25k

    We're looking for a Principal Software Engineer to join our CSP Engagements team as the technical focal point for fleet-scale reliability...  ...we need to see:15+ years of experience in systems software at datacenter scale, or reliability engineering with focus on at-scale... 
    Principal
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $200k - $230k

     ...seek talented, passionate, and committed engineers, technologists, and business leaders to join us.Job Summary:Ready to...  ...own the brain behind the world’s fastest-growing AI datacenters? Supermicro is hiring a Principal Product Manager - DCIM Solutions to lead the... 
    Principal
    Worldwide

    Super Micro Computer

    San Jose, CA
    2 days ago
  •  ...Principal Engineer, Power Engineering About Us Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive... 
    Principal
    Remote work
    Flexible hours

    Foundation Capital

    Milpitas, CA
    1 day ago
  • $180.64k - $261.58k

     ...and developing high-frequency and high-density power converters for Al processors and data centers applications. We are seeking a Principal Power IC Chip Lead to spearhead the development of our next generation of products. In this role, you will be responsible for architecting... 
    Principal
    Permanent employment
    Full time
    Work at office
    Day shift

    Analog Devices

    Milpitas, CA
    2 days ago
  • $195k - $391k

    Hungry, Humble, Honest, with Heart.The OpportunityAs a Principal Product Manager for Nutanix Cloud Clusters (NC2) portfolio, you will...  ...strong plus5+ years of product management experience in enterprise datacenter products, virtualization, IaaS, cloud computing, or related... 
    Principal
    Work at office
    Remote work
    Relocation package
    3 days per week

    Nutanix

    San Jose, CA
    3 days ago
  • $200k - $230k

     ...community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.Job Summary:We are seeking a...  ...driven leader to own and accelerate the market growth of our Datacenter Building Block Solutions (DCBBS). This senior role is... 
    Principal
    Immediate start
    Worldwide

    Supermicro

    San Jose, CA
    1 day ago
  • $180k - $225k

     ...be responsible for driving innovation delivering best in class solutions addressing power efficiency and density challenges for datacenter products serving the most critical AI customers.  You will work directly with customers defining roadmaps aligned with their platform... 
    Principal
    Full time

    Renesas Electronics

    San Jose, CA
    14 days ago
  •  ...Manpower, Inc. seeks a Senior Principal Power Engineer in San Jose to lead power system design for networking, storage, and server products. You will architect and implement advanced power solutions, ensuring high performance and reliability across multi‑tier architectures... 
    Principal

    Jobleads-US

    San Jose, CA
    3 days ago
  •  ....What You'll DoOwn the multi-year sourcing and category strategy for the equipment and construction scopes that deliver Lambda's datacenter capacity — long-lead electrical and mechanical equipment, construction, and trade packagesPartner with Construction, Design, and... 
    Principal
    Contract work
    For contractors
    Work at office
    Local area
    Remote work
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    3 days ago
  •  ...Engineering 12 years of relevant electrical engineering work experience and 5+ years of experience in the system-level design in the datacenter, computer, networking or storage industries In-depth knowledge of hardware and associated signaling environments. Technical... 
    Principal
    Full time
    Work experience placement
    Remote work

    Jobleads-US

    Sunnyvale, CA
    2 days ago
  • $210.16k - $271.98k

    Senior Principal Systems Development EngineerHelp architect and deliver Dell's L11 rack-scale AI solutions: fully integrated racks that bring Dell IP and partner technologies together to fuel our customers' AI growth. As a Principal Systems Development Engineer, you'll... 
    Principal

    Dell Technologies

    Santa Clara, CA
    4 days ago
  • $102.3k - $209.5k

    What You’ll DoOwn and deliver GPU infrastructure programs spanning datacenter enablement, GPU server provisioning, deployment, configuration, and lifecycle operations. Align priorities and execution across multiple teams (engineering, networking, datacenter ops, supply... 
    Principal
    Temporary work
    Flexible hours

    Oracle Corporation

    Santa Clara, CA
    4 days ago
  • $160.28k - $232.09k

    About Analog DevicesAnalog Devices, Inc. (NASDAQ: ADI) is a global semiconductor leader that bridges the physical and digital worlds to enable breakthroughs at the Intelligent Edge. ADI combines analog, digital, AI, and software technologies into solutions that combat climate...
    Principal
    Permanent employment
    Full time
    Work at office
    Day shift

    Analog Devices

    Milpitas, CA
    2 days ago
  •  ...Hewlett Packard Enterprise is seeking a Senior Principal Power Hardware Engineer to design and validate high‑voltage power systems in Sunnyvale. The role emphasizes architecture development, multi‑platform coordination, and mentoring teams. This onsite position requires... 
    Principal

    Jobleads-US

    Sunnyvale, CA
    2 days ago
  • Senior Distinguished Technologist - Pre-Sales AI & Data Center Networking This role has been designated as ‘Remote/Teleworker’, which means you will primarily work from home.Who We Are:Hewlett Packard Enterprise is the global edge-to-cloud company advancing the way people... 
    Full time
    Work experience placement
    Local area
    Immediate start
    Remote work
    Work from home

    Hewlett Packard Enterprise

    San Jose, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Datacenter Technologist. Be the first to apply!