Senior HPC and LSF Operations Engineer
$152k - $241.5kNVIDIA
As a member of the Hardware Infrastructure EDA Compute team, you will optimize, scale, and support workload scheduling systems that directly impact design velocity and infrastructure efficiency. Success in this role requires both operational precision along with developing and supporting forward-looking resource management solutions that address evolving compute demands. Beyond day-to-day operations, the role drives improvements in observability, service reliability, and automation, ensuring the EDA compute environment remains resilient, measurable, and aligned with long-term engineering demands.What you'll be doing:Manage, scale, and optimize job scheduling systems (LSF, Slurm, etc.) in a large-scale, multi-site environment supporting EDA and other compute-intensive workloadsAnalyze scheduler and infrastructure performance data to identify systemic bottlenecks and drive measurable improvements in utilization, throughput, and turnaround timeLead problem solving across scheduler, OS, and workload layers, ensuring timely resolution of service-impacting issuesIdentify recurring operational challenges and implement targeted automation or process improvements to reduce manual effort and prevent repeat incidentsHelp define and track reliable metrics and SLOs for service performance and reliability, partnering with customers to ensure expectations are realistic and measurableContribute to operational standards, documentation, and best practices to improve consistency across sitesPartner directly with customer teams to clarify requirements, translate technical tradeoffs, and drive issues to closureWhat we need to see:Bachelor’s degree in Computer Science or related field, or equivalent experienceMinimum 5+ years of experience operating and supporting large-scale Linux-based compute infrastructureStrong hands-on experience supporting and tuning job scheduling systems (LSF, Slurm, etc.) in HPC or silicon design environmentsProficiency in Linux systems administration (CentOS/RHEL)Strong problem solving skills and the ability to independently analyze complex system behavior under loadClear and effective communication skills, including the ability to articulate technical tradeoffs and reliability metrics to engineering stakeholdersWays to stand out from the crowd:Experience implementing reliability engineering practices within HPC scheduling environmentsDeep knowledge of job scheduling systems (LSF, Slurm, etc.) configuration tuning, scheduler internals, and advanced troubleshooting techniquesExperience building or enhancing observability systems, including metrics collection, monitoring pipelines, alerting strategies, and performance dashboardsBackground with container technologies such as Docker, Singularity, or Podman in HPC environmentsExperience influencing adoption of new infrastructure standards across multiple teams or sitesNVIDIA offers highly competitive salaries and a comprehensive benefits package. We have some of the most forward-thinking and hardworking people in the world on our team and our collaborative talent continues to drive NVIDIA's growth. We are seeking creative and independent engineers with real passion for technology!#LI-Hybrid Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.You will also be eligible for equity and benefits.Applications for this job will be accepted at least until July 24, 2026.This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes.NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.SummaryLocation: US, CA, Santa Clara; US, MA, Westford; US, TX, Austin; US, NC, DurhamType: Full time
$124k - $195.5k
As an HPC Operations Engineer at NVIDIA, you will play a pivotal role in ensuring the flawless operation of our high-performance computing (HPC)... ...operational tasksSolid understanding of workload schedulers such as LSF, Slurm, or similar systemsStrong grasp of network computing...SuggestedFull time- A leading tech company in Austin, Texas is seeking an experienced professional to research and design scalable distributed storage services for high-performance computing workloads. The ideal candidate will have a Bachelor's degree in Computer Science and 8+ years of experience...Senior
$166k - $249k
General InformationJob TitleSenior/Principal ASIC Digital Design Engineer (HPC IP) - 18387Job ID18387CityAustinState/ProvinceTexasDate Posted13-Aug-2026Job CategoryEngineeringJob SubcategoryASIC Digital DesignHire TypeEmployeeRemote EligibleNoBase Salary Range: $166000...SeniorWorldwide$184k - $287.5k
...environment runs millions of cores across federated LSF cells, and every simulation, synthesis run... ...LSF platform, and we are looking for an engineer who knows LSF at the level of its... ...Engineering, or equivalent experience.8+ years in HPC or large-scale batch compute, with 5+...SeniorFull timeRemote work$152k - $241.5k
...world.We are seeking a highly skilled and experienced HPC Cluster Engineer to design, deploy, and operate GPU Compute Clusters for EDA (Electronic Design... ...HPC job schedulers and orchestrators, such as Slurm, LSF, PBS or K8s. Applied experience with AI/HPC workflows...SeniorFull time$136k - $218.5k
NVIDIA is looking for a Senior CPU Tooling and Design Automation Engineer! Do you want to help drive the development of CPU technology for architectures used... ...AI) / deep learning (DL), high-performance computing (HPC), cloud service providers (CSP), gaming, virtual reality...SeniorFull timeWork experience placementNight shift- AMD is seeking an Enterprise AI/HPC GPU architect to join the Datacenter System Architecture and Engineering team to develop world-class products around Instinct GPUs. In this role you will engage with customers and internal teams to define end-to-end architecture spanning...Senior
- Job DescriptionThe Senior Director, Physical Design & Backend Engineering is accountable for end-to-end backend execution across HPC SoC and MCU programs, including implementation, signoff... ..., and area (PPA) targets. The role operates at the intersection of execution,...Senior
- NextSilicon, a leader in HPC acceleration, is seeking an experienced HPC and AI application field support engineer to join our pre-sales engineering team. You will port, benchmark and optimize applications for CPUs, GPUs and FPGAs, and present value to customers across...Senior
- NVIDIA is seeking a Senior CPU Tooling and Design Automation Engineer to advance CPU technology for AI, HPC, cloud, gaming, and autonomous vehicles. You will help shape the toolchain and automation for CPU design and validation. You will collaborate with Architects, RTL...SeniorNight shift
$184k - $287.5k
...foundation for its EDA compute farm, and we need an automation engineer to own it end to end. You are joining at the point where this is... ...you'll be doing:Designing and owning the configuration schema for LSF cell deployment, so that a policy change is written once,...SeniorFull time$120k - $207k
NVIDIA is seeking a Senior HPC Support Engineer to provide customer support for cutting-edge networking solutions. The role involves both onsite and remote support, with duties including troubleshooting and resolving technical issues for customers. The ideal candidate will...SeniorRemote work- NVIDIA is seeking a Senior Software Engineer in Westford, Massachusetts to improve their HPC infrastructure. The role includes designing scalable systems and supporting multi-cloud environments. The ideal candidate will have 10+ years of experience, strong software development...Senior
$184k - $356.5k
A leading technology company in Austin, Texas, is seeking a Senior Software Architect to enhance communication performance in AI and HPC applications. Candidates should have at least 5 years of experience, a relevant Master's or Ph.D., and expertise in C/C++ programming...Senior$184k - $287.5k
...NVSHMEM, and UCX that are crucial for scaling Deep Learning and HPC. We're seeking a Senior Software Architect to help co-design next-gen data center... ...NCCL, NVSHMEM, OpenSHMEM, UCX, UCC).Deep understanding of operating systems, computer and system architecture.Solid in...SeniorFull timeRemote work- ...thinking individuals to join our team.We are looking to add a Senior Production Test Engineer I to our team. If you enjoy working in a startup... ...and reliability through direct support of manufacturing operations. You will support root cause analysis, help with design-...SeniorPermanent employmentFull timeContract workWork experience placementLocal area
$120k - $202.5k
Who We Are Looking For State Street's Cyber Data & Analytics (CyberDNA) team is seeking a Sr.Platform Operations Engineer to help shape the next generation of cybersecurity data, analytics, and AI-powered platforms. Partnering closely with Global Cyber Security, Infrastructure...SeniorFull timeTemporary workFlexible hours$184k - $287.5k
...next-gen distributed storage services for HPC workloads, optimizing both performance... ...infrastructure environments, to automate operational monitoring and alerting, and to enable... ...degree in Computer Science, Electrical Engineering or related field or equivalent experience...SeniorFull time- Terracon is seeking an experienced leader to manage and direct activities of an engineering consulting office, overseeing profit and loss, staff, and client relations. The role drives business development, budgets, safety, and risk management across multiple service lines...SeniorWork at office
- ...supercomputers across industry, academia, and national labs.The AMD HPC & Sovereign AI applications team seeks a strong, experienced,... ..., bringing together developers, customers, and product engineers to deliver programs on-time. You delight in winning business by...Senior
- Join our Austin lab team to operate and maintain sophisticated test fixtures for a world-renowned technology company. As a Senior Hardware Test Automation Engineer, you'll deploy hardware test systems, execute daily testing protocols, and troubleshoot complex electrical...Senior
$184k - $287.5k
...next‑gen distributed storage services for HPC workloads, optimizing both performance... ...infrastructure environments, including operational monitoring, alerting, and self‑service consumption... ...degree in Computer Science, Electrical Engineering or related field, or equivalent...Senior- ...Quality Assurance Engineer Clearly communicate and document quality plans for scrum teams to review. Develop and maintain high... ...and code in Swift/Obj-C with colleagues of different skills and seniority. Validate bug fixes and recommend product improvements to...Senior
$174k - $225k
...assumptions and reimagining how work gets done. Engineers define intent, author precise... ...member — right now at MyWellatDell.comThermal Senior Principal EngineerThe ISG Thermal Engineering... ...AI, Cloud, High Performance Computing (HPC), Edge computing devices and enterprise...Senior- Senior ServiceNow Developer (Modernization, Automation & AI)Be a part of a team that’s ensuring Dell Technologies' product integrity and customer satisfaction. Our IT Software Engineer team turns business requirements into technology solutions by designing, coding and...Senior
$106.7k - $133.4k
POSITION SUMMARY:The Senior Lab Automation Engineer is an experienced engineer with expertise in the development and implementation of complex laboratory... ...with external contractors.Experience in programming and operating liquid handlers. (Tecan, Hamilton, etc)Experience working...SeniorFor contractorsWork at officeLocal areaImmediate startWorldwide$165k - $210k
...solve complex challenges, and stay at the forefront of data engineering and AI advancements. Remote first with casual, award-winning... ...background with most of your experience in infrastructure and operations (managing enterprise data platforms). Responsibilities Leading...SeniorCasual workRemote work$109.6k - $150.7k
Thermal Senior EngineerThe ISG Thermal Engineering team leads and delivers the development of innovative and compliant thermal design solutions, as well... ...interfaces for AI, Cloud, High Performance Computing (HPC), Edge computing devices and enterprise networking, server...Senior- ...involves delivering performance and feature enhancements for GPU products, leading a team, and designing software libraries for AI and HPC applications. Applicants should have over 10 years of professional experience in software development, a strong background in GPU...
- Title: Senior Building Automation Systems (BAS) EngineerLocation: Austin, TXSalary: $13... ...in 1987, this organization is a global engineering firm specializing in building automation... ...SCADA systems to ensure reliable, 24/7 operations across complex environments. As we continue...Senior
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Senior HPC and LSF Operations Engineer. Be the first to apply!
- production network engineer Austin, TX
- network operations center engineer Austin, TX
- production operations engineer Austin, TX
- operations quality engineer Austin, TX
- application operations engineer Austin, TX
- operations engineer Austin, TX
- security operations center engineer Austin, TX
- data center operations engineer Austin, TX
- post production engineer Austin, TX
- remote operation drilling engineer Austin, TX

