Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Principal Firmware Engineer - Server Manageability and Observability

$272k - $431.25k

NVIDIA Corporation

NVIDIA data center systems, such as DGX and HGX, have become core to NVIDIA's rapidly growing enterprise and cloud provider businesses. These platforms bring together the full power of NVIDIA GPUs, NVIDIA NVLink, NVIDIA InfiniBand networking, NVIDIA Grace CPUs, and a fully optimized NVIDIA AI and HPC software stack. We're looking for a strong technical architect to own the end-to-end architecture of these products, at the system software level.

Including firmware, kernel drivers, operating systems, and user mode drivers. You will work with component leads internally and engage with industry leading cloud service providers on taking these products to market.

What you'll be doing:
  • Serve as the primary technical point of contact for major customers, leading technological discussions, defining KPIs, gathering requirements, and addressing complex technical queries.
  • As a system software architect, lead technical innovation and strategic collaborations with major hyperscalers to architect next-generation data center products.
  • Align NVIDIA's roadmap with major customers' requirements through direct engagement.
  • Develop and drive adoption of new technologies and protocols.
  • Make critical technical decisions in ambiguous situations, mitigating risks through left-shift strategies.
What we need to see:
  • Deep expertise in scalable and performant server system architecture, focusing on SW/HW interfaces.
  • Extensive experience with complex system software for accelerators (GPUs, DPUs, FPGAs).
  • Mastery of system firmware (SBIOS, OpenBMC), embedded systems, and Linux kernel internals.
  • Proficiency in Out-of-Band and In-Band management architectures, device management protocols (e.g., MCTP, PLDM, SPDM, RDE) and system management protocols (Redfish, IPMI).
  • Extensive knowledge of networking technologies and protocols, including TCP/IP, Ethernet, InfiniBand, as well as advanced switching and routing concepts
  • Experience collaborating with platform security experts to define tradeoffs between security and ease of use.
  • Demonstrated success in leading complex, cross-functional projects to completion, showcasing the ability to influence and achieve results without direct authority in large-scale, collaborative environments. Demonstrable experience in implementing left shift strategy to de-risk program execution.
  • BS or MS degree in Computer Science, Electrical Engineering or related field (or equivalent experience).
  • 15+ years in the area of System architecture and design.
Ways to stand out from the crowd:
  • Knowledge of cloud and cluster level deployment and management systems. Participation and contributions in standards bodies such as OCP and DMTF.
  • Familiarity with NVIDIA HPC programming models and libraries (CUDA, cuDNN, DOCA)
  • Knowledge of enterprise storage architectures and distributed parallel processing paradigms
NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. We have some of the most forward-thinking and hardworking people on the planet working for us. If you're creative, passionate and self-motivated, we want to hear from you!

NVIDIA's invention of the GPU in 1999 fueled the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern deep learning - the next era of computing - with the GPU acting as the brain of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as the AI computing company." We're looking to grow our company and establish teams with the most thoughtful people in the world. Are you ready to change the next generation of computing? Join us at the forefront of technological advancement.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD.

You will also be eligible for equity and benefits .

Applications for this job will be accepted at least until September 21, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Principal Firmware Engineer - Server Manageability and Observability in Santa Clara, CA vacancy
  • $272k - $431.25k

     .... We are looking for expert engineers to come and help design rack...  ...architect to own end to end manageability architecture for these products...  ...you’ll be doing: Drive server management for large...  ...implemented in right way with each firmware and software module Collaborate... 
    Suggested
    Full time
    Work at office
    Remote work

    NVIDIA

    Santa Clara, CA
    3 days ago
  • $91k - $247k

     ...careers page to see what exciting opportunities and company perks await!Job Description:We are seeking an experienced Senior Manager - Firmware Engineering to lead our San Jose-based firmware team responsible for the development and delivery of firmware solutions for the... 
    Suggested
    Full time

    Microchip Technology

    San Jose, CA
    4 days ago
  •  ...design specifications and develop firmware applications for low-power,...  ...drivers for embedded systems.Manage and maintain source code...  ...teams in Digital and beyond. Engineering for device systems spans deep...  ...and prototypes to products.The Principal Firmware Engineer provides thought... 
    Suggested
    Full time
    Temporary work
    Local area
    Immediate start
    Remote work

    Brambles

    Santa Clara, CA
    3 days ago
  • $180k - $301k

    Firmware Principal EngineerCategorySoftwareLocationSan Jose, CAExperienceMore than 4 Years Work Expe.Job DescriptionWe are...  ...embedded firmware for advanced optical modules and optical engines. This role spans module management, CMIS implementation, PIC control algorithms,... 
    Suggested
    Temporary work

    MediaTek

    San Jose, CA
    21 hours ago
  •  ...enjoy your career with us!Position Title: Principal Firmware EngineerEmployment Type: Full-time,...  ...is seeking a Principal Firmware Engineer to lead architecture, design, and development...  ..., Manufacturing, and Program Management to meet schedule and product goals.What... 
    Suggested
    Full time
    Live in

    Lumentum Operations

    San Jose, CA
    21 hours ago
  • $168k - $231k

     ...the flexibility to do it in their own way. The Role: As a Principal Firmware Engineer, you will play a critical role in designing, developing,...  ...embedded systems. Project Leadership: Lead firmware projects, managing timelines, resources, and collaboration with hardware and... 
    Full time
    Immediate start
    Remote work
    Work from home
    Worldwide
    Flexible hours

    Logitech

    San Jose, CA
    3 days ago
  • $142.8k - $274.8k

     ...Hardware, and Infrastructure Engineering (SCHIE) is the team behind...  ...platform globally with our server and data center...  ...operations, globalization, and manageability solutions. Our focus is on...  ...are seeking an accomplished Principal Firmware Engineer to join a team of... 
    Ongoing contract
    Work at office
    Local area
    Worldwide

    Microsoft Corporation

    Santa Clara, CA
    5 days ago
  • $147k - $237.5k

     ...stronger relationships, and the kind of precision that drives great outcomes.Job SummaryWe are seeking a Principal Software Engineer to join our Machine Identity Management CyberArk team, focused on building and scaling frontend experiences that enable visibility, control,... 
    Full time
    Work at office

    Palo Alto Networks

    Santa Clara, CA
    1 day ago
  • $188k - $265k

     .... Position Summary   Antora is seeking a Sr Manager or Principal, Interconnection & Grid Engineering to lead the technical execution of generation interconnection...  ...that features flexible and inclusive holiday observance, as well as paid volunteer time off.   When it... 
    Remote work
    Flexible hours

    Antora Energy

    San Jose, CA
    22 days ago
  • $248k - $396.75k

     ...time.NVIDIA's IT Storage Engineering team architects, designs, deploys, and manages petabyte-scale storage...  ...platforms, and observability stacks, ensuring a single...  ...knowledge of bare-metal server hardware (rack units,...  ...selection, BMC/iDRAC/iLO, firmware management) and data... 
    Full time
    Shift work

    Nvidia

    Santa Clara, CA
    3 days ago
  • $140k - $215k

     ...Role:At CrowdStrike, Site Reliability Engineering (SRE) is at the forefront of ensuring the...  ...platform. In this role, you'll manage a team of talented engineers, providing...  ...Bitbucket PipelinesBuild and maintain observability frameworks including metrics, distributed... 
    Full time
    Work experience placement
    Work at office
    Local area

    CrowdStrike

    Sunnyvale, CA
    2 days ago
  • $190.9k - $334.1k

    Company DescriptionIt all started when engineer Fred Luddy wrote code that automated a...  ...Business Unit that brings together IT Service Management (ITSM), IT Operations Management (ITOM)...  ...system design for scale, reliability, observability, fault tolerance, data quality,... 
    Work at office
    Immediate start
    Remote work
    Flexible hours

    ServiceNow

    Santa Clara, CA
    21 hours ago
  • $224k - $356.5k

     ...improving efficiency, and scaling. As Technical Lead Manager, you will lead the engineering team within NVIDIA’s Dynamo organization. Your responsibility...  ...including operators, Helm charts, and GPU observability tooling (DCGM, dcgm-exporter, PyNVML).Background in competitive... 
    Full time
    Local area
    Remote work
    Worldwide

    Nvidia

    Santa Clara, CA
    4 days ago
  • $230k - $260k

    Principal Embedded firmware engineer As an Principal Embedded Firmware Engineer, you'll play a central role in shaping the intelligence inside our...  ...flash file systems, communication stacks, and build tool management. Work in a RTOS environment, including porting and... 
    Full time
    Immediate start

    Lunar Energy

    Mountain View, CA
    4 days ago
  • $166.5k - $291.4k

    Company DescriptionIt all started when engineer Fred Luddy wrote code that automated a...  ...for people.Job DescriptionOverviewAs the Manager of Software Engineering, you will lead...  ...checkpoints, exception and retry handling, and observability over what the agent did and why.... 
    Work at office
    Immediate start
    Remote work
    Flexible hours

    ServiceNow

    Santa Clara, CA
    2 days ago
  • $248k - $391k

     ...leader to build and lead a high-performance engineering organization that architects, delivers,...  ...'ll Be Doing:As a Senior Engineering Manager, you will own the technical vision,...  ...Azure).Experience with monitoring and observability tools (Prometheus, Grafana, Datadog, PagerDuty... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $220k - $245k

     ...Top Tier provider of advanced server, storage, and networking...  ..., passionate, and committed engineers, technologists, and business...  ...Senior Software Engineering Manager with exceptional networking...  ...based Controller or Network Observability & Automation software or SDN... 
    Worldwide

    Super Micro Computer

    San Jose, CA
    3 days ago
  • $177k - $226k

     ...focused on transforming the diagnosis and management of patients with serious neurological...  ...:The Senior Manager, Applied AI Engineering is a senior individual contributor role...  ...CI/CD pipelines, secrets management, observability, and security controls Translate business... 
    For contractors
    Work at office
    Local area
    Immediate start
    Remote work

    Ceribell

    Sunnyvale, CA
    1 day ago
  • $239.6k - $324.1k

    AWS Hardware Engineering is looking for a senior leader to deliver development, implementation...  ...engineers and technical program managers, to manage ODMs and work with internal stakeholders...  ...supporting high volume enterprise servers.- Hands-on experience with system-level... 
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    4 days ago
  •  ...transportation on a global scale. Commercial Software Engineering builds digital products and platforms that help fleet customers manage their operations and benefit from connected...  ...debt reduction, production readiness, observability, incident response, and service-level... 
    Full time
    Local area
    Work from home
    Relocation package

    General Motors

    Sunnyvale, CA
    1 day ago
  • $164k - $329k

     ...obsessed with delivering exceptional customer experience. As an Engineering Manager in the DevEx (Development Extension) organization, you will...  ...Drive the design and development of serviceability tools, observability frameworks (logging, metrics, tracing), and automation to... 
    Fixed term contract
    Work at office
    Remote work
    Relocation package
    3 days per week

    Nutanix

    San Jose, CA
    21 hours ago
  • $184k - $287.5k

    NVIDIA is seeking a Senior Firmware Engineer to join our CSP Engagements team, focusing on system software for Datacenter...  ...doing:Design and develop firmware solutions for manageability and observability of data center servers.Actively participate in hardware bring-up... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $165k - $220k

     ...About the Role: Design & implement firmware code for Flash Interface Layer of SSD....  ...operations. ~ Experience in developing NAND managing algorithms or error control coding. ~...  ...either Computer Science or Electrical Engineering; MS is preferred.     COMPENSATION... 

    SK hynix memory solutions America Inc.

    San Jose, CA
    22 days ago
  •  ...global scale, come make a difference at Fiserv.Job TitleSenior Engineering Manager (Backend & Platform)About Your RoleFiserv is a global...  ...governance standards, non-functional requirements (NFRs), and observability protocols to guarantee platform reliability, zero-downtime... 
    Full time
    Contract work
    Work at office
    Worldwide
    Monday to Friday

    Fiserv

    Sunnyvale, CA
    2 days ago
  • $233.9k - $330.4k

     ...globally.Within CHG, the Systems & Optics Team (SaO) serves as the execution engine for Cisco's most strategic hardware initiatives. We partner across silicon, engineering, product management, operations, supply chain, manufacturing, and executive leadership to transform... 
    Full time
    Temporary work
    Work at office
    Local area
    Flexible hours

    CISCO Systems

    San Jose, CA
    1 day ago
  • $190.9k - $334.1k

    Company DescriptionIt all started when engineer Fred Luddy wrote code that automated a...  ...in this roleWe are looking for a Senior Manager of Machine Learning Engineering to lead...  ...Drive operational excellence, reliability, observability, and performance optimization across... 
    Temporary work
    Work at office
    Immediate start
    Remote work
    Flexible hours

    ServiceNow

    Santa Clara, CA
    21 hours ago
  • $201k - $402k

    Hungry Humble HonestThe OpportunityWe are looking for a Senior Engineering Manager to lead the design, development, and scaling of a next-...  ...delivering reliable, production-grade systemsExperience with SLOs, observability, incident management, and lifecycle operationsCollaboration... 
    Work at office
    Local area
    Remote work
    Relocation package
    3 days per week

    Nutanix

    San Jose, CA
    1 day ago
  •  ...Tuesday.About the RoleWe are seeking a Senior Software Engineer to join our Managed Kubernetes (Mk8s) team. You will play a crucial role in...  ...pieces of large-scale Kubernetes clustersExperience with observability at scale: Prometheus, Grafana, distributed tracing, and... 
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    1 day ago
  •  ....THE ROLE :AMD is looking for an experienced software engineer to help build Fleet Manager, a secure control plane for operating large-scale AMD...  ...Kueue, JobSet, container runtimes, storage systems, and observability platforms.Help evolve Fleet Manager into a portable... 

    AMD

    San Jose, CA
    3 days ago
  • $200k - $322k

    We are looking for a highly skilled Senior Software Engineer to design and develop AIOps & Observability platforms at NVIDIA. The platforms are used by internal...  .... You will work with a team of engineers, product managers, and partners to define the observability strategy,... 
    Full time

    Nvidia

    Santa Clara, CA
    21 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Principal Firmware Engineer - Server Manageability and Observability. Be the first to apply!