Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Firmware Engineer - Data Center Server Management

$272k - $431.25k

NVIDIA

US, CA, Santa Clara

US, Remote

Full time

JR1978573

NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern deep learning — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company.” We're looking to grow our company and establish teams with the most thoughtful people in the world. NVIDIA GH200 superchip provides performance and productivity required for strong scaling for HPC and generative AI workload. Scale out is inherent to design of this massive superchip. We are looking for expert engineers to come and help design rack level solutions for next generation scaling AI supercomputing platforms.

We are looking for a strong technical architect to own end to end manageability architecture for these products in data centers. You will work with various component leads internally and externally, drive customer use cases, align architecture with customer requirements and release best products to market.

Join us at the forefront of technological advancement.

What you’ll be doing:

  • Drive server management for large clusters and data centers deploying GPUs and Grace solution from Nvidia.

  • Work with data center architects and cloud customers to narrow down on requirements for implementation to ensure speed of light product development.

  • Work with internal teams to make sure requirements are designed and implemented in right way with each firmware and software module

  • Collaborate with other leads to design & build data center health management workflow.

  • Drive reliability and optimization in firmware architecture from a data center view point.

  • Work closely with cluster bring up team and resolve issues at Speed of Light

  • Own firmware delivered to data centers in terms of quality, reliability and telemetry performance.

What we need to see:

  • 15+ years of relevant experience working on server firmware (BMC) and platform software development

  • BS, MS, or PhD in EE/CS or related field of education or equivalent experience

  • Hands on experience with data center health management workflow. Proven record of delivering server firmware for large data centers..

  • Strong knowledge of data center management, server architecture and server manageability in data centers and strong and demonstrable skill in C/C++ and Python

  • Experience programming and debugging skills for server platforms.

  • Experience in SCM (e.g. Git, Perforce) and project management tools like Jira.

  • You should possess excellent written and oral communication skills, good work ethics, high sense of team-work, love to produce quality work and commitment to finish your tasks every single day.

  • You are a self-starter who loves to find creative solutions to complicated problems and hands on with coding.

Ways to stand out from the crowd:

  • Hands on experience with data center health management

  • Hands on with x86 or ARM system architecture.

  • Proven technical leaders to drive large complex problem with 50+ engineers working

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people on the planet working for us. If you're creative and autonomous, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD for Level 6, and 320,000 USD - 488,750 USD for Level 7.

You will also be eligible for equity and benefits ( .

Applications for this job will be accepted at least until September 5, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry.

Learn more about NVIDIA .

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Principal Firmware Engineer - Data Center Server Management in Santa Clara, CA vacancy
  • $267.8k - $288.56k

     ...mobility, healthcare, energy and data centers. With revenue of more than $...  ...Devices, Inc. Job Title: Principal Data EngineerJob Requisition...  ...and implement master data management (MDM) solutions, incorporating...  ...frameworks, and modern data engineering tools for efficient data... 
    Suggested
    Permanent employment
    Full time
    Work at office
    Remote work
    Work from home
    Day shift
    2 days per week

    Analog Devices

    San Jose, CA
    5 days ago
  •  ...career with us!Position Title: Principal Firmware EngineerEmployment Type:...  ...AI and cloud computing to data centers, telecom, and advanced manufacturing...  ...a Principal Firmware Engineer to lead architecture,...  ...Manufacturing, and Program Management to meet schedule and product... 
    Suggested
    Full time
    Live in

    Lumentum Operations

    San Jose, CA
    6 days ago
  • $219k - $351k

     ...smartphones, electric vehicles, hyperscale data centers, IoT devices, and so much more. Here,...  ...customers, partners, and communities.Principal Engineer, Architecture & Performance Research...  ...our human recruiting team and hiring managers to ensure every candidate is evaluated... 
    Suggested
    Work at office
    Flexible hours
    Shift work

    Samsung Semiconductor

    San Jose, CA
    3 days ago
  • $275k - $382k

     ...robotics automation technical solutions for data center hardware required for AI2 business...  ...within AI2, to meet the business needs.Lead engineering execution and provide technical...  ...billions of Google users worldwide. As Principal Engineer, you will define the strategic... 
    Suggested
    Remote work
    Worldwide
    Flexible hours

    Google

    Sunnyvale, CA
    7 days ago
  • $168k - $258.75k

    We are looking for a Senior Technical Program Manager (TPM) to join NVIDIA’s Server Engineering Operations Team. You will be the cross-section between execution...  ...us take on more of these unique opportunities in data-center solutions.What you will be doing:The Technical... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $230k - $260k

    Principal Embedded firmware engineer As an Principal Embedded Firmware Engineer, you'll play a central role in...  ...stacks, and build tool management. Work in a RTOS environment, including...  ...features and debugging. Review performance data from the Lunar system gathered at internal... 
    Full time
    Immediate start

    Lunar Energy

    Mountain View, CA
    5 days ago
  • $142.8k - $274.8k

     ...EngineeringDiscipline: Firmware EngineeringCompany:...  ...Infrastructure Engineering (SCHIE) is the team...  ...platform globally with our server and data center infrastructure,...  ...globalization, and manageability solutions. Our focus...  ...a highly motivated Principal Firmware Engineer with... 
    Ongoing contract
    Work at office
    Local area
    Worldwide
    3 days per week

    Microsoft

    Mountain View, CA
    4 days ago
  •  ...and real-time platform management. Our technology...  ...infrastructure—from data centers to next generation cloud...  ...takes more than great engineering—it takes a team of exceptional...  ...Sr. Staff Firmware Engineer with deep expertise...  ...with Intel Server Platform Services (SPS... 

    Axiado Corporation

    San Jose, CA
    5 days ago
  • $272k - $431.25k

     ...how you can make a lasting impact on the world!The Data Center MODS organization seeks a Principal Engineer to architect and scale next-generation L10 and L11...  ...knowledge of x86/ARM architectures, Linux OS internals, firmware (UEFI/BIOS), Redfish, HMC, BMC protocols and... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $165k - $220k

     ...everything from smartphones to data centers. As a global leader in DRAM...  ...Role: Design & implement firmware code for Flash Interface...  ...Experience in developing NAND managing algorithms or error control...  ...Computer Science or Electrical Engineering; MS is preferred.... 

    SK hynix memory solutions America Inc.

    San Jose, CA
    3 days ago
  • $272k - $431.25k

     ...are looking for an excellent Senior Engineering Manager to lead a large firmware engineering organization...  ...firmware for NVIDIA's next generation Data Center Compute Systems. This role owns HGX...  ...overall years of proven experience in server firmware, BMC/OpenBMC, MCU... 
    Full time

    Nvidia

    Santa Clara, CA
    6 days ago
  • $211k - $347k

     ...LinkedIn, our approach to flexible work is centered on trust and optimized for culture,...  ...needs of the team. LinkedIn's Data Center Engineering organization is responsible for the strategy...  ...delivery, infrastructure lifecycle management, capacity growth, and critical... 
    Contract work
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Mountain View, CA
    3 days ago
  • $2,000 per month

     ...investors and staffed by leading engineers, Etched is redefining the...  ...infrastructure. As a Data Center Engineer at Etched, you will...  ...networking layout, cabling, hardware management, day-to-day operations, and...  ...hands‑on with custom server platforms and high‑speed networking... 
    Work at office
    Relocation package
    Shift work

    The Consensus

    San Jose, CA
    3 days ago
  • $132k - $189k

    Manage end-to-end physical deployment, rack-level integration, power-up, and network...  ...systems and early-stage hardware in the engineering lab.Perform physical reworks on development...  ....2 years of experience working in a data center operations or facilities technical... 
    Work at office

    Google

    Sunnyvale, CA
    3 days ago
  • $284.9k - $427.3k

     ...Technologies, Inc. Job Area: Engineering Group, Engineering...  ...to transforming server-class computing by...  ...demanding customer and data center requirements. We are...  ...architecture, coherency management, memory systems,...  ...platform enumeration and firmware requirements Deep... 
    Work experience placement
    Work from home

    Socket.dev

    Santa Clara, CA
    3 hours ago
  • $207k - $300k

     ...software systems, programming the data plane and control plane for...  ...network infrastructure.Manage individual project priorities...  ...qualifications:Master’s degree or PhD in Engineering, Computer Science, or a...  ...Global Networking, Data Center operations, systems research,... 
    Worldwide

    Google

    Sunnyvale, CA
    6 days ago
  • $188k - $265k

     ...delivers affordable, reliable energy to industry, data centers, and the grid. Our thermal batteries turn low-cost...  .... Position Summary   Antora is seeking a Sr Manager or Principal, Interconnection & Grid Engineering to lead the technical execution of generation interconnection... 
    Remote work
    Flexible hours

    Antora Energy

    San Jose, CA
    3 days ago
  •  ...production methods.  Job Summary: The AI Server/Rack System Lead Manager leads a multidisciplinary engineering team across System, Electrical, Power, and Mechanical...  ...systems ready for deployment at hyperscale data centers. This role provides technical leadership... 
    Local area

    Foxconn-PCE Technology

    Santa Clara, CA
    3 days ago
  • $284.9k - $427.3k

    ## Lead Server Product Architect- Sr DirectorSanta Clara, California...  ..., Inc.## **Job Area:**Engineering Group, Engineering Group CPU...  ...*General Summary:**Qualcomm Data Center team is developing High performance...  ...CPU architecture, Coherency management, Memory system, Power... 
    Work experience placement
    Work from home

    Qualcomm

    Santa Clara, CA
    2 days ago
  • $133k - $150k

     ...Top Tier provider of advanced server, storage, and networking solutions for Data Center, Cloud Computing, Enterprise IT...  ...talented, passionate, and committed engineers, technologists, and business...  ...-density GPU cluster thermal management, and cooling tower management,... 
    Internship
    Worldwide

    Super Micro Computer

    San Jose, CA
    3 days ago
  • $207k - $300k

    Develop the next generation of Indexing Engine platform, Web Data Service (for rapid prototyping and...  ...Data Indexing domain.Collaborate and manage stakeholders and development counterparts...  ...From developing and maintaining our data centers to building the next generation of... 

    Google

    San Jose, CA
    3 days ago
  •  ...and new network components including but not limited to # Troubleshooting network problems # Customer Incident and request management # Configure network equipment to include but not limited to basic routing & switching, base and port configurations using... 
    Work experience placement

    ADEX

    Santa Clara, CA
    1 day ago
  • $105k - $140k

     ...This position requires presence in our San Jose Data Centers 5 days; shift workWhat You'll DoEnsure new server, storage and network infrastructure is properly racked...  ...centers, such as power distribution, air flow management, environmental monitoring, capacity planning,... 
    Work at office
    Local area
    Flexible hours
    Day shift

    Lambda Labs

    San Jose, CA
    3 days ago
  • $364k

     ...that orchestrate, validate, and manage bare-metal systems, custom AI silicon, and distributed data center environments.Design and...  ...silicon, and distributed systems engineering teams to ensure the...  ...layers, or custom kernels).As a Principal/Distinguished Engineer for AI... 
    Worldwide

    Google

    Sunnyvale, CA
    3 days ago
  • $207k - $300k

     ...of high-performance server platforms,...  ...software, and system engineering teams to drive pre-silicon Firmware (FW)/Software (SW)...  ...professional growth.Manage project priorities,...  ...of experience with data structures and algorithms...  ...Global Networking, Data Center operations, systems... 
    Worldwide

    Google

    Sunnyvale, CA
    5 days ago
  •  ...design specifications and develop firmware applications for low-power,...  ...drivers for embedded systems.Manage and maintain source code...  ...practices and seamless collaboration.Engineering for device systems spans deep...  ...Digital team to gather the data needed to derive the IoT data... 
    Full time
    For contractors
    Local area
    Immediate start
    Remote work

    Brambles

    Santa Clara, CA
    6 days ago
  • $146.7k - $339.3k

     ...expectThis role leads global network infrastructure and data center operations supporting over 10,000 employees and AI/ML training infrastructure. The manager oversees a team of 12 globally distributed network engineers and data center technicians. The team is operating in... 
    Full time
    Work at office
    Remote work

    Zoom

    San Jose, CA
    7 days ago
  • $157.3k - $212.8k

     ...support the development and management of Compute, Database,...  ...hardware, and network engineers, supply chain...  ...segment of accelerated servers.You will work closely...  ...interdisciplinary team of component, firmware, test, qualification,...  ...these servers to the data center. After launch you will... 
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    7 days ago
  • $133.5k - $272k

     ...AI era. We secure and accelerate cloud, data, and AI in real time, everywhere. Thousands...  ...Netskope One platform, its Zero Trust Engine, and the powerful NewEdge network to gain...  ...wherever it resides. As a Senior Engineering Manager, you will lead the Data Lineage team—the... 

    Netskope

    Santa Clara, CA
    3 days ago
  •  ...our lives. Here, you’ll do more than join something — you’ll add something. The AIML team is looking for an engineering program manager (EPM) to work with Data Scientists and Data Engineers to define, instrument, track and report key performance indicators for Apple Products... 

    Socket

    Cupertino, CA
    3 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Firmware Engineer - Data Center Server Management. Be the first to apply!