Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

System Software Engineer, Distributed Systems

$152k - $241.5k

Jobleads-US

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. The VLSI Productivity and Infrastructure team supports 1000+ chip design engineers by building tools and platforms that supercharge their everyday work. Our mission: make chip designers faster. We build and operate long shelf-life systems spanning build automation, observability, analytics, automated error detection/remediation, and codebase modernization—with a strong commitment to stability. Our core workflow infrastructure runs as userspace software on bare-metal Linux hosts (no sudo, no containers). We coordinate shared state and artifacts via NFS, launch long-running, compute-heavy workflows on IBM LSF, and provide adjacent services for APIs and observability. This is a high-ownership environment where you'll often be the expert on what you build. We are looking for a pragmatic and versatile systems engineer who enjoys working near the metal and building tools that empower other engineers. This is a generalist role with an emphasis on distributed systems and operational excellence in a “below containers” world: coordination, reliability, performance, and safe evolution of legacy systems (including incremental modernization of large codebases into Go). This isn't a CI/CD pipeline configuration role; you will be writing the userspace software that manages state, concurrency, and reliability at scale.

What you will be doing:

  • Design, build, and deliver core components of our next-generation productivity platforms
  • Develop reliable userspace infrastructure for long-running engineering workflows at scale on bare-metal Linux hosts
  • Build state coordination over NFS (atomicity, idempotency/dedup, partial-write recovery, without privileged ops)
  • Build and improve orchestration around IBM LSF (submission/tracking, retries/cancel, log capture, fairness/backpressure)
  • Convert legacy codebases into modern powerhouses using incremental migration techniques (e.g., Perl to Go), with stage gates, parity strategies, and strong observability
  • Debug and improve performance and reliability across Linux and Kubernetes, including operational tooling
  • Collaborate with engineering users to turn ambiguous workflows into durable production systems

What we need to see:

  • B.S. CS/EE (or equivalent experience)
  • 5+ years developing and operating production software in Go and/or Python, ideally in large codebases
  • Strong Linux fundamentals: processes, filesystems, permissions, synchronization/locks, concurrency, and debugging
  • Solid distributed-systems thinking: failures, retries/timeouts, backoff, idempotency, and operational rigor
  • Experience building long-runtime automation or services on shared compute clusters (batch schedulers, build systems)
  • Ability to translate ambitious, high-level goals into a safe delivery plan (instrumentation, staged rollout, measurable outcomes)

Ways to stand out from the crowd:

  • Hands-on experience with shared filesystems at scale (NFS), or coordination patterns on eventually-consistent storage
  • Experience with batch job scheduling, shared compute fleets, or build systems
  • Track record of incremental modernization (tests, shadow runs, canaries, rollback plans)
  • Experience partitioning/optimizing metadata-heavy systems and reducing I/O or R/W hot spots
  • Strong incident/debug tactics: clear root-cause analysis, remediation, and guardrails as well as rapid comprehension and ownership of unfamiliar codebases in any language (including LLM-generated code) to implement high-leverage changes

With competitive salaries and a generous benefits package, we are widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us and, due to unprecedented growth, our exclusive engineering teams are rapidly growing. If you're a creative and autonomous engineer with a real passion for technology, we want to hear from you. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4. You will also be eligible for equity and benefits .

Applications for this job will be accepted at least until October 10, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer.

As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

#J-18808-Ljbffr Jobleads-US
Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the System Software Engineer, Distributed Systems in Santa Clara, CA vacancy
  • $152k - $241.5k

     ...supports 1000+ chip design engineers by building tools and platforms...  ...and operate long shelf-life systems spanning build automation,...  ...infrastructure runs as userspace software on bare-metal Linux hosts (...  ...role with an emphasis on distributed systems and operational... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    6 hours ago
  • $272k - $431.25k

     ...fabs, equipment manufacturers, and software providers to make inspection, metrology...  ...data is limited and proprietary, distributions shift across tools and fabs, production...  ....We’re seeking a Principal Systems Software Engineer for Semiconductor Inspection in Santa... 
    Suggested
    Full time
    Local area
    Shift work

    Nvidia

    Santa Clara, CA
    2 days ago
  • $128k - $184k

     ...meet user needs by working with engineers to refine code and tools, and...  ...practical experience in software development.4 years of experience...  ...software development and system design.Experience with front-...  ...experience designing and optimizing distributed applications.2 years of... 
    Suggested

    Google

    Sunnyvale, CA
    6 hours ago
  •  ...Description ThisWay Global is looking for a Distributed Systems Engineer in a remote role within the United States. ThisWay Global, Inc. is...  ...NVL72/GB300 GPU clusters. Amalgamy.ai — AI Orchestration Software: An enterprise AI orchestration platform focused on GPU... 
    Suggested
    Full time
    Remote work

    GrabJobs

    San Jose, CA
    1 day ago
  •  ...the latest in a line of cutting-edge systems that have been helping customers and...  ...looking for a deeply hands‑on Senior Distributed Systems Engineer to join the team building IonQ’s Network...  ...~6+ years of backend software engineering experience, with a strong... 
    Suggested
    Work at office

    Canaan Partners

    Santa Clara, CA
    1 day ago
  •  ...operates ultra-scale GPU supercomputing systems to train next-generation foundation models...  ...— driving communication performance, distributed reliability, and cross-layer optimization...  ...We are looking for a deeply technical engineer to co-design and optimize the communication... 

    Lever

    Sunnyvale, CA
    4 days ago
  • Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs...  ...speed inference.Cerebras is seeking a Distributed Systems Security Engineer to help build out the next...  ...CerebrasPeople who are serious about software make their own hardware. At Cerebras... 

    Cerebras Systems

    Sunnyvale, CA
    2 days ago
  • $168k - $270.25k

     ...Experience (NVEX) Solutions Engineering team is looking for an experienced...  ...contribute to products and software tooling. You must have...  ...Expertise analyzing performance of distributed GPU-accelerated workloads is...  ...multi-GPU platformsStrong system software (firmware, BIOS,... 
    Full time
    Weekend work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $172k - $229.6k

     ...monitor, fix, and improve the reliability of large-scale distributed systems. This platform supports the telemetry and reliability needs...  ...Prometheus, ClickHouse, OpenTelemetry, and related tools.As a Software Engineer on this team, you will design and build scalable... 
    Immediate start
    Remote work
    Visa sponsorship

    eBay

    San Jose, CA
    4 days ago
  • $59.15k - $106.93k

     ...helping you grow your career is good business. Leidos is seeking Distribution Engineers in Maui, HI and Oahu, HI who are passionate about electric...  ...one of the nation's most unique electric utility systems. Your greatest work is ahead! Location & Relocation These positions... 
    Local area
    Immediate start
    Relocation
    Relocation package
    Flexible hours
    Night shift

    Leidos

    San Jose, CA
    2 days ago
  •  ...Google LLC in Sunnyvale, CA is seeking a Web Solutions Engineer for gUP Distribution Technology to design, develop, and maintain distribution platforms at scale. You will collaborate with partners and product teams to deliver robust web features and APIs using modern backend... 

    Jobleads-US

    Sunnyvale, CA
    2 days ago
  •  ...Google is hiring a Web Solutions Engineer for gTech Users and Products (gUP) Distribution Technology in Sunnyvale. You will design, develop, enhance, and maintain distribution platforms to scale Google products through partner channels. You will work with cross-functional... 

    Jobleads-US

    Sunnyvale, CA
    1 day ago
  • $184k - $287.5k

     ...building a scalable and modular software stack that powers advanced driver-assistance systems across a diverse range of...  ...motivated Senior Software Systems Engineer with a strong foundation in software...  ....Familiarity with parallel/distributed systems and low-level system profiling... 
    Full time

    Nvidia

    Santa Clara, CA
    16 hours ago
  • $175k - $275k

    Cerebras Systems builds the world's largest AI chip, 56 times larger...  ...RoleAs part of the Embedded Software team, you will help build the...  ...the Cerebras Wafer Scale Engine (WSE)—the world’s largest AI...  ...systems, platform engineering, and distributed system enablement. As our... 

    Cerebras Systems

    Sunnyvale, CA
    1 day ago
  • $184k - $287.5k

     ...together cutting‑edge hardware and software innovation to deliver industry‑...  ...a group of forward‑thinking engineers tackling some of the globe’s...  .... We’re searching for a Senior Systems Software Engineer with deep expertise in distributed systems, Kubernetes, containers... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    16 hours ago
  • $165.2k - $223.6k

    AWS designs custom SoCs (System on Chips) that power the world's...  ...both the SoCs and the low-level software stack that brings these chips...  ...for a Systems Software Engineer who wants to work at the boundary...  ...collective communication libraries or distributed systems primitives (MPI, NCCL... 
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    4 days ago
  • $184k - $287.5k

     ...acceleration. We're looking for an outstanding software engineer to apply their skills in the...  ...accelerate other libraries and database systems. In this role, you will research,...  ...from the crowd:Experience developing distributed algorithms and running on distributed... 
    Full time

    Nvidia

    Santa Clara, CA
    16 hours ago
  • $184k - $287.5k

    NVIDIA is looking for a hardworking Sr. Systems Software Engineer to work on platform software based on open-source container runtimes and Kubernetes...  ...GO and Kubernetes, experience with Systems Software and Distributed systems, as well as excellent communication and planning... 
    Full time
    Work experience placement
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

    We are seeking a Sr System Software Engineer to help us build out our scientific computing platform workflows on Cloud. This Cloud based scientific...  ....10+ years experience working on building and operating distributed compute and data intensive platform as a service on... 
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $224k - $356.5k

    We are now looking for a Senior System Software Engineer to work on Dynamo. NVIDIA is hiring software engineers for its GPU-accelerated deep...  ...task costs for self-hosted LLMs.Build and evolve Dynamo’s distributed inference frontend across vLLM, SGLang, and TensorRT-LLM,... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $184k - $287.5k

     ...make a lasting impact on the world.NVIDIA is seeking a Sr. Systems Software Engineer for the Apache Spark Acceleration group. Over the past...  ...file formats such as Parquet, ORC and JSONCollaborate with distributed systems teams to craft solutions to distributed processing... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $184k - $287.5k

     ...sits at the intersection of research and systems engineering, where we turn new ideas into reliable...  ...experience.8+ years of relevant software engineering experience.Experience in one...  ...with real-time or low-latency inference, distributed training or inference, performance... 
    Full time

    Nvidia

    Santa Clara, CA
    4 days ago
  • $184k - $287.5k

     ...learning, supercomputing, gaming, and visualization. As a Senior System Software Engineer on the NvSci team, you will play an integral role in...  ....Ability to work effectively in cross-functional, distributed teams.Your base salary will be determined based on your location... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    4 days ago
  • $184k - $287.5k

    NVIDIA is growing a senior engineering team focused on making our compute software stack first-class on NVIDIA CPU platforms...  ...are looking for an experienced systems software engineer who can lead...  ...Perforce-scale development, or distributed build systems.Experience turning... 
    Full time
    Remote work

    Nvidia

    Santa Clara, CA
    1 day ago
  • $184k - $287.5k

     ...reliable production workflows? Our Engineering Workflow Platform team...  ...artifacts, EDA tools, distributed jobs, validation checks, and...  ...architecture with hands-on systems engineering to solve challenges...  ...least 12 years of relevant software engineering experience.Production... 
    Full time
    Immediate start

    Nvidia

    Santa Clara, CA
    4 days ago
  • $152k - $241.5k

    NVIDIA Solutions Engineering team is searching for engineers to help develop and bring NVIDIA...  ...their best work. We are looking for a System Software Engineer with expertise in embedded...  ...part of an internationally distributed team with locations in US, Europe, APAC... 
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • $207k - $300k

    Write and test product or system development code.Review code developed by other engineers and provide feedback to ensure best...  ...technical field.Google's software engineers develop the next-generation...  ...information retrieval, distributed computing, large-scale system... 
    Worldwide

    Google

    Sunnyvale, CA
    1 day ago
  •  ...Google LLC in Sunnyvale, CA is seeking a Senior Software Engineer, Infrastructure, Spanner to design and optimize distributed storage systems at scale. You will work on core Spanner features, collaborating with product teams to deliver high-availability solutions for global... 

    Jobleads-US

    Sunnyvale, CA
    1 day ago
  • $227k - $303k

     ...role We are looking for a Principal Engineer to provide technical leadership across...  ...across multiple teams, and solve complex distributed-systems problems in security-critical...  ...experience designing and building production software, including substantial experience with... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Remote work
    Flexible hours

    CoreWeave

    Sunnyvale, CA
    25 days ago
  • $149k - $224k

     ...part of the Core Products Business Unit (CPBU) Systems Software team, you will join a high-leverage engineering initiative dedicated to transforming upgrade workflows...  ...in C++ and Python. Solid foundations in distributed systems, high availability (HA), storage... 
    For contractors
    Local area
    Flexible hours
    Shift work

    Pentair

    Santa Clara, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to System Software Engineer, Distributed Systems. Be the first to apply!