System Software Engineer, Distributed Systems
$152k - $241.5kJobleads-US
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. The VLSI Productivity and Infrastructure team supports 1000+ chip design engineers by building tools and platforms that supercharge their everyday work. Our mission: make chip designers faster. We build and operate long shelf-life systems spanning build automation, observability, analytics, automated error detection/remediation, and codebase modernization—with a strong commitment to stability. Our core workflow infrastructure runs as userspace software on bare-metal Linux hosts (no sudo, no containers). We coordinate shared state and artifacts via NFS, launch long-running, compute-heavy workflows on IBM LSF, and provide adjacent services for APIs and observability. This is a high-ownership environment where you'll often be the expert on what you build. We are looking for a pragmatic and versatile systems engineer who enjoys working near the metal and building tools that empower other engineers. This is a generalist role with an emphasis on distributed systems and operational excellence in a “below containers” world: coordination, reliability, performance, and safe evolution of legacy systems (including incremental modernization of large codebases into Go). This isn't a CI/CD pipeline configuration role; you will be writing the userspace software that manages state, concurrency, and reliability at scale.
What you will be doing:
- Design, build, and deliver core components of our next-generation productivity platforms
- Develop reliable userspace infrastructure for long-running engineering workflows at scale on bare-metal Linux hosts
- Build state coordination over NFS (atomicity, idempotency/dedup, partial-write recovery, without privileged ops)
- Build and improve orchestration around IBM LSF (submission/tracking, retries/cancel, log capture, fairness/backpressure)
- Convert legacy codebases into modern powerhouses using incremental migration techniques (e.g., Perl to Go), with stage gates, parity strategies, and strong observability
- Debug and improve performance and reliability across Linux and Kubernetes, including operational tooling
- Collaborate with engineering users to turn ambiguous workflows into durable production systems
What we need to see:
- B.S. CS/EE (or equivalent experience)
- 5+ years developing and operating production software in Go and/or Python, ideally in large codebases
- Strong Linux fundamentals: processes, filesystems, permissions, synchronization/locks, concurrency, and debugging
- Solid distributed-systems thinking: failures, retries/timeouts, backoff, idempotency, and operational rigor
- Experience building long-runtime automation or services on shared compute clusters (batch schedulers, build systems)
- Ability to translate ambitious, high-level goals into a safe delivery plan (instrumentation, staged rollout, measurable outcomes)
Ways to stand out from the crowd:
- Hands-on experience with shared filesystems at scale (NFS), or coordination patterns on eventually-consistent storage
- Experience with batch job scheduling, shared compute fleets, or build systems
- Track record of incremental modernization (tests, shadow runs, canaries, rollback plans)
- Experience partitioning/optimizing metadata-heavy systems and reducing I/O or R/W hot spots
- Strong incident/debug tactics: clear root-cause analysis, remediation, and guardrails as well as rapid comprehension and ownership of unfamiliar codebases in any language (including LLM-generated code) to implement high-leverage changes
With competitive salaries and a generous benefits package, we are widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us and, due to unprecedented growth, our exclusive engineering teams are rapidly growing. If you're a creative and autonomous engineer with a real passion for technology, we want to hear from you. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4. You will also be eligible for equity and benefits .
Applications for this job will be accepted at least until October 10, 2026.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer.
As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
#J-18808-Ljbffr Jobleads-US$152k - $241.5k
...supports 1000+ chip design engineers by building tools and platforms... ...and operate long shelf-life systems spanning build automation,... ...infrastructure runs as userspace software on bare-metal Linux hosts (... ...role with an emphasis on distributed systems and operational...SuggestedFull time$272k - $431.25k
...fabs, equipment manufacturers, and software providers to make inspection, metrology... ...data is limited and proprietary, distributions shift across tools and fabs, production... ....We’re seeking a Principal Systems Software Engineer for Semiconductor Inspection in Santa...SuggestedFull timeLocal areaShift work$128k - $184k
...meet user needs by working with engineers to refine code and tools, and... ...practical experience in software development.4 years of experience... ...software development and system design.Experience with front-... ...experience designing and optimizing distributed applications.2 years of...Suggested- ...Description ThisWay Global is looking for a Distributed Systems Engineer in a remote role within the United States. ThisWay Global, Inc. is... ...NVL72/GB300 GPU clusters. Amalgamy.ai — AI Orchestration Software: An enterprise AI orchestration platform focused on GPU...SuggestedFull timeRemote work
- ...the latest in a line of cutting-edge systems that have been helping customers and... ...looking for a deeply hands‑on Senior Distributed Systems Engineer to join the team building IonQ’s Network... ...~6+ years of backend software engineering experience, with a strong...SuggestedWork at office
- ...operates ultra-scale GPU supercomputing systems to train next-generation foundation models... ...— driving communication performance, distributed reliability, and cross-layer optimization... ...We are looking for a deeply technical engineer to co-design and optimize the communication...
- Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs... ...speed inference.Cerebras is seeking a Distributed Systems Security Engineer to help build out the next... ...CerebrasPeople who are serious about software make their own hardware. At Cerebras...
$168k - $270.25k
...Experience (NVEX) Solutions Engineering team is looking for an experienced... ...contribute to products and software tooling. You must have... ...Expertise analyzing performance of distributed GPU-accelerated workloads is... ...multi-GPU platformsStrong system software (firmware, BIOS,...Full timeWeekend work$172k - $229.6k
...monitor, fix, and improve the reliability of large-scale distributed systems. This platform supports the telemetry and reliability needs... ...Prometheus, ClickHouse, OpenTelemetry, and related tools.As a Software Engineer on this team, you will design and build scalable...Immediate startRemote workVisa sponsorship$59.15k - $106.93k
...helping you grow your career is good business. Leidos is seeking Distribution Engineers in Maui, HI and Oahu, HI who are passionate about electric... ...one of the nation's most unique electric utility systems. Your greatest work is ahead! Location & Relocation These positions...Local areaImmediate startRelocationRelocation packageFlexible hoursNight shift- ...Google LLC in Sunnyvale, CA is seeking a Web Solutions Engineer for gUP Distribution Technology to design, develop, and maintain distribution platforms at scale. You will collaborate with partners and product teams to deliver robust web features and APIs using modern backend...
- ...Google is hiring a Web Solutions Engineer for gTech Users and Products (gUP) Distribution Technology in Sunnyvale. You will design, develop, enhance, and maintain distribution platforms to scale Google products through partner channels. You will work with cross-functional...
$184k - $287.5k
...building a scalable and modular software stack that powers advanced driver-assistance systems across a diverse range of... ...motivated Senior Software Systems Engineer with a strong foundation in software... ....Familiarity with parallel/distributed systems and low-level system profiling...Full time$175k - $275k
Cerebras Systems builds the world's largest AI chip, 56 times larger... ...RoleAs part of the Embedded Software team, you will help build the... ...the Cerebras Wafer Scale Engine (WSE)—the world’s largest AI... ...systems, platform engineering, and distributed system enablement. As our...$184k - $287.5k
...together cutting‑edge hardware and software innovation to deliver industry‑... ...a group of forward‑thinking engineers tackling some of the globe’s... .... We’re searching for a Senior Systems Software Engineer with deep expertise in distributed systems, Kubernetes, containers...Full timeRemote work$165.2k - $223.6k
AWS designs custom SoCs (System on Chips) that power the world's... ...both the SoCs and the low-level software stack that brings these chips... ...for a Systems Software Engineer who wants to work at the boundary... ...collective communication libraries or distributed systems primitives (MPI, NCCL...Local areaFlexible hours$184k - $287.5k
...acceleration. We're looking for an outstanding software engineer to apply their skills in the... ...accelerate other libraries and database systems. In this role, you will research,... ...from the crowd:Experience developing distributed algorithms and running on distributed...Full time$184k - $287.5k
NVIDIA is looking for a hardworking Sr. Systems Software Engineer to work on platform software based on open-source container runtimes and Kubernetes... ...GO and Kubernetes, experience with Systems Software and Distributed systems, as well as excellent communication and planning...Full timeWork experience placementRemote work$184k - $287.5k
We are seeking a Sr System Software Engineer to help us build out our scientific computing platform workflows on Cloud. This Cloud based scientific... ....10+ years experience working on building and operating distributed compute and data intensive platform as a service on...Full time$224k - $356.5k
We are now looking for a Senior System Software Engineer to work on Dynamo. NVIDIA is hiring software engineers for its GPU-accelerated deep... ...task costs for self-hosted LLMs.Build and evolve Dynamo’s distributed inference frontend across vLLM, SGLang, and TensorRT-LLM,...Full time$184k - $287.5k
...make a lasting impact on the world.NVIDIA is seeking a Sr. Systems Software Engineer for the Apache Spark Acceleration group. Over the past... ...file formats such as Parquet, ORC and JSONCollaborate with distributed systems teams to craft solutions to distributed processing...Full time$184k - $287.5k
...sits at the intersection of research and systems engineering, where we turn new ideas into reliable... ...experience.8+ years of relevant software engineering experience.Experience in one... ...with real-time or low-latency inference, distributed training or inference, performance...Full time$184k - $287.5k
...learning, supercomputing, gaming, and visualization. As a Senior System Software Engineer on the NvSci team, you will play an integral role in... ....Ability to work effectively in cross-functional, distributed teams.Your base salary will be determined based on your location...Full timeRemote work$184k - $287.5k
NVIDIA is growing a senior engineering team focused on making our compute software stack first-class on NVIDIA CPU platforms... ...are looking for an experienced systems software engineer who can lead... ...Perforce-scale development, or distributed build systems.Experience turning...Full timeRemote work$184k - $287.5k
...reliable production workflows? Our Engineering Workflow Platform team... ...artifacts, EDA tools, distributed jobs, validation checks, and... ...architecture with hands-on systems engineering to solve challenges... ...least 12 years of relevant software engineering experience.Production...Full timeImmediate start$152k - $241.5k
NVIDIA Solutions Engineering team is searching for engineers to help develop and bring NVIDIA... ...their best work. We are looking for a System Software Engineer with expertise in embedded... ...part of an internationally distributed team with locations in US, Europe, APAC...Full time$207k - $300k
Write and test product or system development code.Review code developed by other engineers and provide feedback to ensure best... ...technical field.Google's software engineers develop the next-generation... ...information retrieval, distributed computing, large-scale system...Worldwide- ...Google LLC in Sunnyvale, CA is seeking a Senior Software Engineer, Infrastructure, Spanner to design and optimize distributed storage systems at scale. You will work on core Spanner features, collaborating with product teams to deliver high-availability solutions for global...
$227k - $303k
...role We are looking for a Principal Engineer to provide technical leadership across... ...across multiple teams, and solve complex distributed-systems problems in security-critical... ...experience designing and building production software, including substantial experience with...Permanent employmentFull timeTemporary workCasual workWork at officeRemote workFlexible hours$149k - $224k
...part of the Core Products Business Unit (CPBU) Systems Software team, you will join a high-leverage engineering initiative dedicated to transforming upgrade workflows... ...in C++ and Python. Solid foundations in distributed systems, high availability (HA), storage...For contractorsLocal areaFlexible hoursShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to System Software Engineer, Distributed Systems. Be the first to apply!
- system programmer Santa Clara, CA
- IT system engineer Santa Clara, CA
- systems software developer Santa Clara, CA
- senior linux systems engineer Santa Clara, CA
- senior staff systems engineer Santa Clara, CA
- operations support system engineer Santa Clara, CA
- advanced systems engineer Santa Clara, CA
- software system engineer Santa Clara, CA
- electronic systems engineer Santa Clara, CA
- system engineer remote Santa Clara, CA



