Senior System Reliability Engineer
$168k - $264.5kNVIDIA
NVIDIA has continuously reinvented itself over two decades. Our invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing — with the GPU acting as the brains of computers, robots, and self-driving cars that can perceive and understand the world. Today, we are increasingly known as “the AI computing company.” We're looking to grow our company and build our teams with the most thoughtful people in the world. Join us at the forefront of technological advancement. GPU Servers are one of the fastest-growing segments for NVIDIA and the Artificial Intelligence industry. As the computational power increases with every GPU generation, developing efficient and reliable systems is an imperative. We are looking for a System Reliability Engineer to join NVIDIA's existing Reliability Engineering team, involved in NVIDIA's diverse system product range specifically Graphics and High-Performance Computing printed circuit boards and Data Center Servers.What you'll be doing:Represent Product Reliability Engineering in development teams.Develop and execute reliability test plans for product qualification. Collaborate with peer reliability groups for successful implementation.Define product reliability tests for products ranging from embedded, automotive, graphics cards to server, rack, and cluster products.Lead reliability testing, failure analysis, and root cause investigations; drive corrective actions to improve design and manufacturing quality.Establish and continuously improve product reliability standards, metrics, and methodologies.Participate in product and engineering design reviews, assess the reliability budget and influence changes that enhance product reliability.Collaborate cross-functionally with engineering teams, suppliers, and partners to achieve reliability targets using Design for Reliability (DfR) methods including FMEA and DoE approaches.Provide reliability predictions to access and drive product reliability to meet product requirements.What we need to see:Bachelor’s or Master’s degree in Electrical Engineering, Mechanical Engineering, or equivalent experience.8+ years of experience in hardware reliability or hardware engineering from datacenter, systems, or computer industries.Hands-on experience in theoretical and practical Reliability concepts as it relates to high-tech electronic enterprise and consumer products.Have a strong understanding of statistical concepts and how they relate to product reliability & life analysis.Good verbal and writing skills as well as the ability to communicate at a high level.Good project management skills and ability to balance multiple simultaneous projects during development.Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 168,000 USD - 264,500 USD.You will also be eligible for equity and benefits.Applications for this job will be accepted at least until June 7, 2026.This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.SummaryLocation: US, CA, Santa ClaraType: Full time
$166k - $244k
Overview Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both our internally critical and our externally-visible systems—have...SeniorFull time$116k - $184k
...build the next era of computing!We're seeking an outstanding Senior HTOL Reliability Engineer to join our Santa Clara lab. This role requires deep... ...develop and implement improvements to burn-in boards, HTOL systems, and thermal interface materials.What we need to see:...SeniorFull time$167.3k - $284.4k
...hands without us. KLA invents systems and solutions for the... ...expert teams of physicists, engineers, data scientists and problem-... ...application development engineers, and senior product technology process... .../Preferred QualificationsSr. Reliability Engineer - SEM SystemsJoin a...SeniorMinimum wageFull timeWork experience placementFlexible hours$210.16k - $271.98k
Senior Principal Systems Development EngineerHelp architect and deliver Dell's L11 rack-scale AI solutions... .... As a Principal Systems Development Engineer, you'll own the system-level... ...requirements.Ensure the manufacturability, reliability, and serviceability of rack designs,...Senior- Senior Principal Systems Development EngineerHelp architect and deliver Dell's L11 rack-scale AI solutions: fully integrated racks that bring Dell... ...customers' AI growth. As a Principal Systems Development Engineer, you'll own the system-level architecture and engineering...Senior
$114k - $171k
...opportunities to work on revolutionary systems that impact people's lives around the world... ...solutions for global security. Our Engineering and Sciences (E&S) organization pushes... ...naval surface ships and submarines. The Reliability and System Safety Engineering Department...Full timeWork experience placementRelocation packageShift work$136k - $218.5k
We are seeking Systems Quality and Reliability Engineer to join our LPU team!NVIDIA has continuously reinvented itself over two decades. Our invention of the GPU in 1999 fueled the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel...Full time$332k
...great technology—and amazing people.NVIDIA's hardware reliability is foundational to some of the world's most... ...critical conditions. As Sr. Director of Board and System Level Reliability, you will set the engineering standard for how NVIDIA's products perform and thrive...SeniorFull timeWork at office$106.9k
...jobs that impact everyone's life. Build the future nobody’s dreamed of yet... are you ready to rethink the impossible? As a Senior Engineer Systems in our Research & Development team, you'll have the opportunity to merge creativity with your technical expertise by...SeniorLocal area$155.5k - $248.9k
...deliver exceptional results.Opportunity OverviewTeradyne's Memory Test Division is seeking a Senior Reliability Design Engineer to help develop next-generation semiconductor test systems by driving reliability into products from concept through qualification and release. In...SeniorLocal areaRelocationFlexible hours$220k - $255k
A leading electric mobility company in Palo Alto is seeking a Reliability Engineer to drive the reliability of electric mobility vehicles during their development cycle. The ideal candidate will have experience in reliability engineering, strong technical knowledge of...SeniorFlexible hours$136k - $218.5k
...Automotive, and Embedded markets. As a Silicon Speed Features Engineer, you will co-design system-level speed features, build the validation and... ...system architects, hardware, firmware/software, process/reliability, and operations teams to co-design system-level speed features...Full time$188k - $274k
...dynamic thermal and power envelopes on fleet reliability.Build health state monitoring and... ...thermal solutions and power delivery/battery systems. Drive Failure Modes and Effects... ...qualifications:Bachelor’s degree in Reliability Engineering, Data Science, Mechanical/Electrical...Worldwide- A leading semiconductor equipment manufacturer in Santa Clara, CA, seeks a Systems Engineer IV to design, integrate, and optimize complex systems driving the semiconductor industry forward. Candidates should have a strong background in optics and electrical engineering...Senior
$147k - $202.5k
Applied Materials, Inc. is seeking a Mechanical Engineer IV in Santa Clara, CA. You will lead mechanical system design for semiconductor equipment, ensuring integration of fluid handling and performance. Ideal candidates should have a PhD or master's in mechanical engineering...Senior$113.67k - $153.8k
...Senior Mechanical Engineer Structural Integrity Associates (SIA) is seeking a Senior Mechanical Engineer with expertise in thermo-fluid systems to join our Energy Services Group within the Turbine Generator Services team. This role is a senior individual contributor...SeniorTemporary workCasual workFlexible hours- ...technologies—like the da Vinci surgical system and Ion—have transformed how care is delivered... ...of patients worldwide.We’re a team of engineers, clinicians, and innovators united by one... ...DescriptionPrimary Function of Position Senior Systems Analysts (Robotic Control Engineers...SeniorLocal areaWorldwideFlexible hours
- ...computing. About the Team: SK HMS Systems engineering / Quality assurance group is looking for a motivated individual to join our reliability testing team. Job Description:... ...record working as a contractor or senior consultant, capable of rapidly onboarding...SeniorContract workFor contractors
- ...robotic-assisted surgery, is seeking a Senior Mechanical Engineer in Sunnyvale, CA. The role leads... ...maintain and improve the da Vinci surgical system, solving complex electro-mechanical... ...while ensuring manufacturability and reliability. You will design components, perform...Senior
- Senior Site Reliability Engineer ILocationSan Jose, Costa Rica - RemoteSummary of roleOwn availability, the most important product feature, by continually... ...Sumo’s developers to deliver features more rapidly.Scale systems sustainably through mechanisms like automation, and...SeniorFlexible hours
- ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building... ...includes the Lambda website, cloud APIs and systems as well as internal tooling for system... ...networking teams to improve service reliability and deployment workflowsDeploy and maintain...SeniorWork at officeLocal areaWork from homeFlexible hours
- ...We are seeking a Senior Database Reliability Engineer (DBRE) to design, operate, and improve reliable, scalable, secure, and highly available database... ...engineering, site reliability engineering, Linux systems administration, and infrastructure automation. The ideal...Senior
- ...automate, simplify, and accelerate revenue.We are looking for a Senior Site Reliability Engineer to lead the strategic evolution of our cloud... ...our environment. Your mission is to take a high-velocity system and implement the best practices, guardrails, and automated...SeniorFull timeWork at office2 days per week
- ...work from home day is currently Tuesday.Engineering at Lambda is responsible for building... ...includes the Lambda website, cloud APIs and systems as well as internal tooling for system... ...services, workloads, and platform reliability.You6+ years of experience in a SRE, operations...SeniorWork at officeLocal areaWork from homeFlexible hours
- ...provider located in Santa Clara, California is seeking an experienced Opto-Mechanical Engineer. The ideal candidate will contribute to the development of precision measurement systems, requiring expertise in mechanical design, optics, and proficiency in CAD and FEA tools...SeniorFlexible hours
- NVIDIA Corporation is looking for a Senior CAD Engineer in Santa Clara, CA. This role involves developing innovative tools for PCB electrical... ...with databases and automation in multiple operating systems. The competitive salary range for this position varies based...Senior
$90k - $180k
...serve people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale... ...building and maintaining the resilient backbone for systems where failure is not an option, and where our success directly...SeniorRemote work$267k - $356k
...is currently Tuesday.Lambda's Storage Engineering team is the backbone behind our world-class... ...services—from low-level storage systems to the APIs and tooling our customers build... ...workloads in the industry, which means reliability and performance aren't just goals—they'...SeniorWork experience placementWork at officeLocal areaWork from homeFlexible hours$168k - $270.25k
...of artificial intelligence.Join our team at NVIDIA as a Senior Site reliability engineer focused on HPC storage and play a crucial role in designing... ...technology evaluations, related to distributed file systems.Collaborate across teams to better understand developers'...SeniorFull time$183k - $237k
...Senior Principal System Development Engineer Infrastructure Solutions Group (ISG) builds the products that power infrastructure, solutions, and data management our customers need most. Our teams design and develop the hardware and software that connect infrastructure...SeniorFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior System Reliability Engineer. Be the first to apply!
- system engineer remote Santa Clara, CA
- senior windows systems engineer Santa Clara, CA
- senior linux systems engineer Santa Clara, CA
- ground systems engineer Santa Clara, CA
- advanced systems engineer Santa Clara, CA
- system verification engineer Santa Clara, CA
- wireless systems engineer Santa Clara, CA
- senior staff systems engineer Santa Clara, CA
- mission system engineer Santa Clara, CA
- sr systems engineer Santa Clara, CA


