Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Reliability Engineer

$122.44k - $232.19k

Intel

Job Details:Job Description: Join us to help build the next generation of AI hardware solutions. You will be part of a highly skilled, agile team developing cutting-edge hardware for the AI domain, where we push the boundaries of what silicon can do for emerging AI workloads. With a startup-like culture, we move quickly and give engineers the opportunity to drive significant technical and business impact. We are continuously developing modern and effective working methods, including hands-on adoption of AI tools throughout the chip development flow. Mission: Define and own the pod-level reliability specifications that ensure the availability, resilience, and serviceability of a large-scale data center across hardware, thermal, and operational dimensions.Responsibilities:Define and maintain pod-level reliability/availability specs and targets (MTBF, AFR, RAS) for compute, memory, storage, network, power, and cooling subsystems.Translate system/SLA requirements into pod and subsystem level reliability specs; flow requirements down to silicon, platform, and facilities teams.Lead FMEA, root-cause analysis, and pod fleet failure-data analytics to drive corrective actions and spec updates.Architect RAS features (ECC, memory mirroring, predictive failure, telemetry) and graceful degradation/redundancy against pod-level specs.Partner with facilities on pod power/cooling redundancy (N+1, 2N), thermal margins, and disaster-recovery readiness.Establish HALT/HASS, burn-in, qualification processes; track field returns and KPIs against pod spec.Qualifications:Minimum Qualifications:BS/MS/PhD in EE/ME Reliability or related; and/or at least 4-6 yrs experience.Experience authoring and owning reliability specs and requirement flow-down.Strong RAS, FMEA, statistical reliability (Weibull, FIT) skills.Experience with large-scale fleet telemetry and thermal/power redundancy.Preferred Qualifications: AI cluster operations, data analytics (Python/SQL).Job Type:Experienced HireShift:Shift 1 (United States of America)Primary Location: US, Massachusetts, Beaver BrookAdditional Locations:US, California, Santa ClaraPosting Statement:All qualified applicants will receive consideration for employment without regard to race, color, religion, religious creed, sex, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, military and veteran status, marital status, pregnancy, gender, gender expression, gender identity, sexual orientation, or any other characteristic protected by local law, regulation, or ordinance.Position of TrustN/ABenefitsWe offer a total compensation package that ranks among the best in the industry. It consists of competitive pay, stock bonuses, and benefit programs which include health, retirement, and vacation. Find out more about the benefits of working at Intel. Annual Salary Range for jobs which could be performed in the US: $122,440.00-232,190.00 USDThe range displayed on this job posting reflects the minimum and maximum target compensation for the position across all US locations. Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training. Your recruiter can share more about the specific compensation range for your preferred location during the hiring process.Work Model for this RoleThis role will require an on-site presence. * Job posting details (such as work model, location or time type) are subject to change.*ADDITIONAL INFORMATION: Intel is committed to Responsible Business Alliance (RBA) compliance and ethical hiring practices. We do not charge any fees during our hiring process. Candidates should never be required to pay recruitment fees, medical examination fees, or any other charges as a condition of employment. If you are asked to pay any fees during our hiring process, please report this immediately to your recruiter.SummaryLocation: US, Massachusetts, Beaver Brook; US, California, Santa ClaraType: Full time

Vacancy posted 13 hours ago
Similar jobs that could be interesting for youBased on the Reliability Engineer in Santa Clara, CA vacancy
  • $168k - $264.5k

     ...innovator in computer graphics, PC gaming, and accelerated computing, as we step into the next era shaped by AI. As a Senior Reliability Engineer, you'll work within a focused team, developing groundbreaking products that form the future. This unique position provides... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    13 hours ago
  •  ...of the most challenging global health issues.Play a meaningful role in the development of new life science technology as a reliability engineer in the R&D Impact Engineering group. Provide hands-on, analytical, and reliability expertise to support engineering programs... 
    Suggested
    Full time

    BD, Becton and Dickinson

    San Jose, CA
    1 day ago
  •  ...people from diverse backgrounds and industries to create a safer, sustainable and more connected world. Job OverviewA Quality & Reliability Engineer Supervisor uses their engineering skills to assist in issues related to the Quality of the product. This job also involves... 
    Suggested
    Remote work

    TE connectivity

    San Jose, CA
    22 hours ago
  • $168k - $264.5k

     ...human inventiveness and intelligence. Make the choice to join us today. We are seeking an outstanding candidate for Silicon Reliability Engineer to drive and utilize cutting edge technologies to deliver high performance products while ensuring world class reliability.What... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $168k - $264.5k

     ...industry. As the computational power increases with every GPU generation, developing efficient and reliable systems is an imperative. We are looking for a System Reliability Engineer to join NVIDIA's existing Reliability Engineering team, involved in NVIDIA's diverse system... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    2 days ago
  • $116k - $184k

     ...profoundly impacting society. Come join the team and help build the next era of computing!We're seeking an outstanding Senior HTOL Reliability Engineer to join our Santa Clara lab. This role requires deep device-circuitry knowledge and hands-on hardware development. You will... 
    Full time

    Nvidia

    Santa Clara, CA
    5 days ago
  •  ...we advance your career. THE ROLE:We are looking for a Quality Engineer II to support component- and board-level failure analysis for...  ...fault isolation, physical failure analysis, board-level analysis, reliability stress support, and clear technical reporting. This is a hands... 

    AMD

    San Jose, CA
    3 days ago
  • $120k - $250k

     ...humanoid robots with human level intelligence. Its robots are engineered to perform a variety of tasks in the home and commercial...  ...5 days/week in-office collaboration. We are looking for a Reliability Test Engineer to design and execute test plans for our humanoid... 
    Full time
    Contract work
    Work at office

    Figure

    San Jose, CA
    22 hours ago
  • At Bloom Energy, our vision for a world powered by clean, reliable, and affordable energy is more than just a dream—we’re making it reality...  ...of the 21st century.We are looking for a Staff Reliability Engineer to join our team in one of today’s most exciting technologies.... 
    Full time
    Work experience placement
    Remote work
    Worldwide

    Bloom Energy

    San Jose, CA
    22 hours ago
  • $110.5k - $152k

    Who We AreApplied Materials is a global leader in materials engineering solutions used to produce virtually every new chip and advanced...  ..., applies, revises, maintains and/ or tests quality/ reliability standards to ensure alignment with customer expectations.Designs... 
    Full time

    Applied Materials

    Santa Clara, CA
    3 days ago
  • $130.17k - $188.49k

     ...software technologies into solutions that combat climate change, reliably connect humans and the world, and help drive advancements in...  ...on LinkedIn and X.Analog Devices, Inc. (ADI) is seeking a Staff Engineer to join our Reliability Hardware & Systems Development team in... 
    Permanent employment
    Full time
    Work at office
    Day shift

    Analog Devices

    San Jose, CA
    2 days ago
  • $136k - $218.5k

    We are seeking Systems Quality and Reliability Engineer to join our LPU team!NVIDIA has continuously reinvented itself over two decades. Our invention of the GPU in 1999 fueled the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel... 
    Full time

    Nvidia

    Santa Clara, CA
    22 hours ago
  • $332k

     ...fueled by great technology—and amazing people.NVIDIA's hardware reliability is foundational to some of the world's most demanding...  ...Director of Board and System Level Reliability, you will set the engineering standard for how NVIDIA's products perform and thrive across... 
    Full time
    Work at office

    Nvidia

    Santa Clara, CA
    1 day ago
  • $155.5k - $248.9k

     ...drive innovation, and deliver exceptional results.Opportunity OverviewTeradyne's Memory Test Division is seeking a Senior Reliability Design Engineer to help develop next-generation semiconductor test systems by driving reliability into products from concept through... 
    Local area
    Relocation
    Flexible hours

    Universal Robots

    San Jose, CA
    22 hours ago
  • $124.36k - $146.3k

     ...career goals: partnering with our customers, our communities, and each other. Job DescriptionResponsibilitiesAs a senior-level Reliability Engineer specializing in observability, this role partners closely with product owners, application engineering teams, SRE teams, and... 
    Full time
    Work experience placement
    Local area
    3 days per week

    US Bank

    Cupertino, CA
    13 hours ago
  • $136k - $218.5k

     ...Automotive, and Embedded markets. As a Silicon Speed Features Engineer, you will co-design system-level speed features, build the validation...  ...with system architects, hardware, firmware/software, process/reliability, and operations teams to co-design system-level speed features... 
    Full time

    Nvidia

    Santa Clara, CA
    3 days ago
  • $122.5k - $196k

     ...cutting-edge technology, Teradyne may be the place for you. Opportunity OverviewTeradyne is seeking a talented, experienced Reliability Engineer to join our Memory Test Division in San Jose, California. In this role, you will partner with instrumentation, automation, system... 
    Local area
    Relocation
    Flexible hours

    Universal Robots

    San Jose, CA
    22 hours ago
  • $138k - $183.5k

    Who We AreApplied Materials is a global leader in materials engineering solutions used to produce virtually every new chip and advanced...  ...about our benefits. Key ResponsibilitiesEvaluates, from a reliability standpoint, the materials, properties and techniques used in... 
    Full time

    Applied Materials

    Santa Clara, CA
    1 day ago
  • $116k - $159.5k

    Who We AreApplied Materials is a global leader in materials engineering solutions used to produce virtually every new chip and advanced...  ...our winning team that is focused on Quality Engineering and Reliability Engineering. We are working with engineers and scientists across... 
    Full time
    Worldwide

    Applied Materials

    Santa Clara, CA
    3 days ago
  • $157.3k - $212.8k

    The Trainium Manufacturing, Quality and Reliability (MQR) Team is part of AWS Annapurna Labs focused on Machine Learning products that designs...  ...’s largest Cloud Services provider. As a Senior Reliability Engineer you will engage with an experienced cross-disciplinary staff... 
    Work experience placement
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    2 days ago
  •  ...solutions that will define the future of computing. About the Team: SK HMS Systems engineering / Quality assurance group is looking for a motivated individual to join our reliability testing team. Job Description: SK Hynix Memory Solutions has a key design... 
    Contract work
    For contractors

    SK hynix memory solutions America Inc.

    San Jose, CA
    more than 2 months ago
  •  ...someone who might be interested in, please feel free to reach me at (***) ***-****/ ****@*****.***. Title: Reliability Engineering Support Position Type: Full Time Permanent/ Direct Hire Background This role supports the hardware development group by executing... 
    Permanent employment
    Full time

    Confidencial

    Cupertino, CA
    4 days ago
  • A leading technology firm is seeking a Reliability Engineering Support specialist in Cupertino, CA. This full-time position involves executing reliability testing and failure analysis on cutting-edge hardware technologies. The ideal candidate has a BS or MS in Engineering... 
    Full time

    Confidencial

    Cupertino, CA
    4 days ago
  • NVIDIA Gruppe is seeking a Silicon Speed Features Engineer to lead validation and automation infrastructure for silicon issues. You will work across teams to ensure product quality and performance in a dynamic environment. This role requires an MS in EE or equivalent,... 

    NVIDIA Gruppe

    Santa Clara, CA
    5 days ago
  • Lab Engineer III Duration: 24 months Pay Rate: $60-75/hr on W2 + 15 days PTO (Without benefits) About the Team The Reliability Engineering Team plays a critical role in ensuring Client's products meet the highest standards of reliability and performance. We work closely... 
    Contract work

    US Tech Solutions

    Sunnyvale, CA
    4 days ago
  • $167.7k - $245.2k

     ...guardrails that ensure AI agents behave as intended, improving reliability and reducing risks. This unified approach empowers our...  ...enhanced observability and control.As a Senior Site Reliability Engineer (SRE), you will build, operate, and continuously improve the reliability... 
    Full time
    Temporary work
    Local area
    Flexible hours
    2 days per week

    CISCO Systems

    Sunnyvale, CA
    5 days ago
  • $90k - $180k

     ...and branded generic medicines. Our 115,000 colleagues serve people in more than 160 countries.About the RoleThis Senior Site Reliability Engineer position works on-site out of our Sylmar, CA or Sunnyvale, CA location in the Cardiac Rhythm Management Division.We are... 
    Remote work

    Abbott

    Sunnyvale, CA
    3 days ago
  • $147k - $210k

     ...product or system development code.Review code developed by other engineers and provide feedback to ensure best practices (e.g., style...  ...analyzing, and troubleshooting large-scale distributed systems. Site Reliability Engineering (SRE) is what you get when you treat operations as... 

    Google

    Sunnyvale, CA
    1 day ago
  •  ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering...  ...software, platform, and networking teams to improve service reliability and deployment workflowsDeploy and maintain network... 
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    3 days ago
  •  ...Lambda’s designated work from home day is currently Tuesday.Engineering at Lambda is responsible for building and scaling our cloud offering...  ...and SLIs for Kubernetes services, workloads, and platform reliability.You6+ years of experience in a SRE, operations engineer, or... 
    Work at office
    Local area
    Work from home
    Flexible hours

    Lambda Labs

    San Jose, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Reliability Engineer. Be the first to apply!