Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Firmware Engineer - Data Center Server Management

$272k - $431.25k

NVIDIA

NVIDIA’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern deep learning — the next era of computing — with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company.” We're looking to grow our company and establish teams with the most thoughtful people in the world. NVIDIA GH200 superchip provides performance and productivity required for strong scaling for HPC and generative AI workload. Scale out is inherent to design of this massive superchip. We are looking for expert engineers to come and help design rack level solutions for next generation scaling AI supercomputing platforms.

We are looking for a strong technical architect to own end to end manageability architecture for these products in data centers. You will work with various component leads internally and externally, drive customer use cases, align architecture with customer requirements and release best products to market.

Join us at the forefront of technological advancement.

What you’ll be doing:

  • Drive server management for large clusters and data centers deploying GPUs and Grace solution from Nvidia.

  • Work with data center architects and cloud customers to narrow down on requirements for implementation to ensure speed of light product development.

  • Work with internal teams to make sure requirements are designed and implemented in right way with each firmware and software module

  • Collaborate with other leads to design & build data center health management workflow.

  • Drive reliability and optimization in firmware architecture from a data center view point.

  • Work closely with cluster bring up team and resolve issues at Speed of Light

  • Own firmware delivered to data centers in terms of quality, reliability and telemetry performance.

What we need to see:

  • 15+ years of relevant experience working on server firmware (BMC) and platform software development

  • BS, MS, or PhD in EE/CS or related field of education or equivalent experience

  • Hands on experience with data center health management workflow. Proven record of delivering server firmware for large data centers..

  • Strong knowledge of data center management, server architecture and server manageability in data centers and strong and demonstrable skill in C/C++ and Python

  • Experience programming and debugging skills for server platforms.

  • Experience in SCM (e.g. Git, Perforce) and project management tools like Jira.

  • You should possess excellent written and oral communication skills, good work ethics, high sense of team-work, love to produce quality work and commitment to finish your tasks every single day.

  • You are a self-starter who loves to find creative solutions to complicated problems and hands on with coding.

Ways to stand out from the crowd:

  • Hands on experience with data center health management

  • Hands on with x86 or ARM system architecture.

  • Proven technical leaders to drive large complex problem with 50+ engineers working

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people on the planet working for us. If you're creative and autonomous, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD for Level 6, and 320,000 USD - 488,750 USD for Level 7.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 5, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Vacancy posted 12 hours ago
Similar jobs that could be interesting for youBased on the Principal Firmware Engineer - Data Center Server Management in Santa Clara, CA vacancy
  • $267.8k - $288.56k

     ...mobility, healthcare, energy and data centers. With revenue of more than $...  ...Devices, Inc. Job Title: Principal Data EngineerJob Requisition...  ...and implement master data management (MDM) solutions, incorporating...  ...frameworks, and modern data engineering tools for efficient data... 
    Suggested
    Permanent employment
    Full time
    Work at office
    Remote work
    Work from home
    Day shift
    2 days per week

    Analog Devices

    San Jose, CA
    2 days ago
  •  ...career with us!Position Title: Principal Firmware EngineerEmployment Type:...  ...AI and cloud computing to data centers, telecom, and advanced manufacturing...  ...a Principal Firmware Engineer to lead architecture,...  ...Manufacturing, and Program Management to meet schedule and product... 
    Suggested
    Full time
    Live in

    Lumentum Operations

    San Jose, CA
    3 days ago
  • $142.8k - $274.8k

     ...Hardware, and Infrastructure Engineering (SCHIE) is the team...  ...platform globally with our server and data center infrastructure, security...  ...operations, globalization, and manageability solutions. Our focus is...  ...seeking an accomplished Principal Firmware Engineer to join a team... 
    Suggested
    Ongoing contract
    Work at office
    Local area
    Worldwide

    Microsoft Corporation

    Santa Clara, CA
    5 days ago
  • $275k - $382k

     ...robotics automation technical solutions for data center hardware required for AI2 business...  ...within AI2, to meet the business needs.Lead engineering execution and provide technical...  ...billions of Google users worldwide. As Principal Engineer, you will define the strategic... 
    Suggested
    Remote work
    Worldwide
    Flexible hours

    Google

    Sunnyvale, CA
    4 days ago
  • $219k - $351k

     ...smartphones, electric vehicles, hyperscale data centers, IoT devices, and so much more. Here,...  ...customers, partners, and communities.Principal Engineer, Architecture & Performance Research...  ...our human recruiting team and hiring managers to ensure every candidate is evaluated... 
    Suggested
    Work at office
    Flexible hours
    Shift work

    Samsung Semiconductor

    San Jose, CA
    4 hours ago
  • $168k - $258.75k

    We are looking for a Senior Technical Program Manager (TPM) to join NVIDIA’s Server Engineering Operations Team. You will be the cross-section between execution...  ...us take on more of these unique opportunities in data-center solutions.What you will be doing:The Technical... 
    Full time

    Nvidia

    Santa Clara, CA
    4 hours ago
  • $142.8k - $274.8k

     ...EngineeringDiscipline: Firmware EngineeringCompany:...  ...Infrastructure Engineering (SCHIE) is the team...  ...platform globally with our server and data center infrastructure,...  ...globalization, and manageability solutions. Our focus...  ...a highly motivated Principal Firmware Engineer with... 
    Ongoing contract
    Work at office
    Local area
    Worldwide
    3 days per week

    Microsoft

    Mountain View, CA
    1 day ago
  • $230k - $260k

    Principal Embedded firmware engineer As an Principal Embedded Firmware Engineer, you'll play a central role in...  ...stacks, and build tool management. Work in a RTOS environment, including...  ...features and debugging. Review performance data from the Lunar system gathered at internal... 
    Full time
    Immediate start

    Lunar Energy

    Mountain View, CA
    2 days ago
  •  ...and real-time platform management. Our technology...  ...infrastructure—from data centers to next generation cloud...  ...takes more than great engineering—it takes a team of exceptional...  ...Sr. Staff Firmware Engineer with deep expertise...  ...with Intel Server Platform Services (SPS... 

    Axiado Corporation

    San Jose, CA
    2 days ago
  • $272k - $431.25k

     ...how you can make a lasting impact on the world!The Data Center MODS organization seeks a Principal Engineer to architect and scale next-generation L10 and L11...  ...knowledge of x86/ARM architectures, Linux OS internals, firmware (UEFI/BIOS), Redfish, HMC, BMC protocols and... 
    Full time

    Nvidia

    Santa Clara, CA
    4 hours ago
  • $272k - $431.25k

     ...are looking for an excellent Senior Engineering Manager to lead a large firmware engineering organization...  ...firmware for NVIDIA's next generation Data Center Compute Systems. This role owns HGX...  ...overall years of proven experience in server firmware, BMC/OpenBMC, MCU... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $211k - $347k

     ...LinkedIn, our approach to flexible work is centered on trust and optimized for culture,...  ...needs of the team. LinkedIn's Data Center Engineering organization is responsible for the strategy...  ...delivery, infrastructure lifecycle management, capacity growth, and critical... 
    Contract work
    For contractors
    Work at office
    Flexible hours

    Linkedin

    Mountain View, CA
    1 hour ago
  • $2,000 per month

     ...investors and staffed by leading engineers, Etched is redefining the...  ...infrastructure. As a Data Center Engineer at Etched, you will...  ...networking layout, cabling, hardware management, day-to-day operations, and...  ...hands‑on with custom server platforms and high-speed networking... 
    Work at office
    Relocation package
    Shift work

    Etched.ai, Inc.

    San Jose, CA
    3 days ago
  • $132k - $189k

    Manage end-to-end physical deployment, rack-level integration, power-up, and network...  ...systems and early-stage hardware in the engineering lab.Perform physical reworks on development...  ....2 years of experience working in a data center operations or facilities technical... 
    Work at office

    Google

    Sunnyvale, CA
    4 hours ago
  • $207k - $300k

     ...software systems, programming the data plane and control plane for...  ...network infrastructure.Manage individual project priorities...  ...qualifications:Master’s degree or PhD in Engineering, Computer Science, or a...  ...Global Networking, Data Center operations, systems research,... 
    Worldwide

    Google

    Sunnyvale, CA
    3 days ago
  • $133k - $150k

     ...Top Tier provider of advanced server, storage, and networking solutions for Data Center, Cloud Computing, Enterprise IT...  ...talented, passionate, and committed engineers, technologists, and business...  ...-density GPU cluster thermal management, and cooling tower management,... 
    Internship
    Worldwide

    Super Micro Computer

    San Jose, CA
    4 hours ago
  • $207k - $300k

    Develop the next generation of Indexing Engine platform, Web Data Service (for rapid prototyping and...  ...Data Indexing domain.Collaborate and manage stakeholders and development counterparts...  ...From developing and maintaining our data centers to building the next generation of... 

    Google

    San Jose, CA
    4 hours ago
  • $284.9k - $427.3k

    ## Lead Server Product Architect- Sr DirectorSanta Clara, California...  ..., Inc.## **Job Area:**Engineering Group, Engineering Group CPU...  ...*General Summary:**Qualcomm Data Center team is developing High performance...  ...CPU architecture, Coherency management, Memory system, Power... 
    Work experience placement
    Work from home

    Qualcomm

    Santa Clara, CA
    4 days ago
  •  ...and new network components including but not limited to # Troubleshooting network problems # Customer Incident and request management # Configure network equipment to include but not limited to basic routing & switching, base and port configurations using... 
    Work experience placement

    ADEX

    Santa Clara, CA
    3 days ago
  • $105k - $140k

     ...This position requires presence in our San Jose Data Centers 5 days; shift workWhat You'll DoEnsure new server, storage and network infrastructure is properly racked...  ...centers, such as power distribution, air flow management, environmental monitoring, capacity planning,... 
    Work at office
    Local area
    Flexible hours
    Day shift

    Lambda Labs

    San Jose, CA
    1 hour ago
  • $207k - $300k

     ...of high-performance server platforms,...  ...software, and system engineering teams to drive pre-silicon Firmware (FW)/Software (SW)...  ...professional growth.Manage project priorities,...  ...of experience with data structures and algorithms...  ...Global Networking, Data Center operations, systems... 
    Worldwide

    Google

    Sunnyvale, CA
    2 days ago
  • $364k

     ...that orchestrate, validate, and manage bare-metal systems, custom AI silicon, and distributed data center environments.Design and...  ...silicon, and distributed systems engineering teams to ensure the...  ...layers, or custom kernels).As a Principal/Distinguished Engineer for AI... 
    Worldwide

    Google

    Sunnyvale, CA
    4 hours ago
  •  ...design specifications and develop firmware applications for low-power,...  ...drivers for embedded systems.Manage and maintain source code...  ...practices and seamless collaboration.Engineering for device systems spans deep...  ...Digital team to gather the data needed to derive the IoT data... 
    Full time
    For contractors
    Local area
    Immediate start
    Remote work

    Brambles

    Santa Clara, CA
    3 days ago
  •  ...procurement, transport logistics, data center design and construction, equipment management, and daily operations....  ...We are seeking a Principal Kubernetes Control Plane Engineer to architect the foundational...  ...isolation mechanisms at the API server layer to support... 
    Full time
    Local area

    Bitdeer Technologies Group

    San Jose, CA
    8 days ago
  • $146.7k - $339.3k

     ...expectThis role leads global network infrastructure and data center operations supporting over 10,000 employees and AI/ML training infrastructure. The manager oversees a team of 12 globally distributed network engineers and data center technicians. The team is operating in... 
    Full time
    Work at office
    Remote work

    Zoom

    San Jose, CA
    4 days ago
  • $133.5k - $272k

     ...AI era. We secure and accelerate cloud, data, and AI in real time, everywhere. Thousands...  ...Netskope One platform, its Zero Trust Engine, and the powerful NewEdge network to gain...  ...wherever it resides. As a Senior Engineering Manager, you will lead the Data Lineage team—the... 

    Netskope

    Santa Clara, CA
    4 hours ago
  • $157.3k - $212.8k

     ...support the development and management of Compute, Database,...  ...hardware, and network engineers, supply chain...  ...segment of accelerated servers.You will work closely...  ...interdisciplinary team of component, firmware, test, qualification,...  ...these servers to the data center. After launch you will... 
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    4 days ago
  • $184k - $287.5k

     ...are searching for an outstanding Senior Firmware Engineer to join the NVIDIA System Control Firmware...  ...to dynamic power, clock, and thermal management for top-tier autonomous vehicles, AI edge devices, next-generation data centers, and advanced robotics. This role lets you... 
    Full time
    Worldwide

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

    NVIDIA is seeking a Senior Firmware Engineer to join our CSP Engagements team, focusing on system software for Datacenter...  ...doing:Design and develop firmware solutions for manageability and observability of data center servers.Actively participate in hardware bring-up... 
    Full time

    Nvidia

    Santa Clara, CA
    4 hours ago
  • $82.5k - $132k

     ...you already have a Candidate Account, please Sign-In before you apply.Job Description:Broadcom's Data Center Solutions Group (DCSG) is seeking a skilled Firmware Engineer to join our embedded storage systems development team. In this role, you will design, implement, and... 
    Full time
    Local area

    Broadcom

    San Jose, CA
    1 day ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Firmware Engineer - Data Center Server Management. Be the first to apply!