Systems Engineer - HPC & GPU Infrastructure
FiveInsights
Job Description
Job Description
Job Description
Our partner, Vexterra Group, is looking for a mid-career Systems Engineer (Mid-Career) – HPC & GPU Infrastructure with a deep understanding of operating systems, hardware, Kubernetes, and NVIDIA GPU products. As a Systems Engineer (Mid-Career) – HPC & GPU Infrastructure, you will play a pivotal role in designing, developing, and optimizing GPU clusters for the IC community customers. This is a 100% on-site position. All work must be performed at the customer site in Bethesda at the Intelligence Community Campus.
Primary Responsibilities
1. HPC and GPU environment engineering: Contribute to the installation and maintenance of GPU
and HPC hardware on-prem and in the cloud, providing insights into hardware performance to
ensure efficient interaction with software components.
2. Performance Optimization: Analyze HPC/GPU cluster performance, identify bottlenecks, and
develop strategies to enhance performance across various applications in Linux, addressing both
hardware and software considerations. Regularly monitor and improve performance.
3. HPC/GPU tooling: Install and configure HPC/GPU job scheduling and workload management
platforms such as Slurm, PBS , Apache Airflow, Kubernetes
4. Power Efficiency: Work on power management techniques to optimize GPU power
consumption, ensuring efficient operation on both mobile and desktop Linux platforms.
Continuously assess and enhance power efficiency strategies.
5. Testing and Validation: Design and execute tests to validate GPU performance and functionality
on Linux, including stress testing, benchmarking, and debugging to ensure robust operation.
Maintain and expand the testing suite.
6. Documentation: Maintain comprehensive technical documentation, including architectural
specifications, code documentation, and Linux-specific best practices for GPU development.
Keep documentation up to date with changes and improvements.
7. Industry Insight: Stay updated on the latest trends, innovations, and competitive landscapes
within the GPU industry, contributing to research efforts and proposing Linux-specific
approaches to GPU design and optimization. Share regular updates and insights with the team.
QualificationsBasic Qualifications
- Bachelor's or higher degree in Computer Science, Electrical Engineering, or a related field.
- Additional years of experience may be considered in lieu of a degree.
- 4+ years of relevant systems engineering experience
- Expertise in operating system integration for Linux.
- Strong understanding of computer hardware architecture, particularly as it relates to Linux systems.
- Knowledge of parallel computing, graphics algorithms, and real-time rendering in Linux environments.
- Excellent problem-solving skills and the ability to collaborate within a team.
- Strong communication skills for conveying technical information in a Linux context.
- Proficiency with scripting languages such as Python or BASH.
- Proficiency with automation tools such Ansible, Puppet, Salt, Terraform, etc.
- Candidate must, at a minimum, meet DoD 8570.11- IAT Level II certification requirements (currently Security+ CE, CCNA-Security, GICSP, GSEC, or SSCP along with an appropriate computing environment (CE) certification). An IAT Level III certification would also be acceptable (CASP+, CCNP Security, CISA, CISSP, GCED, GCIH, CCSP).
- TS/SCI clearance with Polygraph required or a TS/SCI and willingness to obtain a Polygraph prior to starting.
Preferred Qualifications
- Knowledge of GPU virtualization, cloud computing, and emerging Linux-based technologies in the field.
- Experience with container technologies (Docker, Kubernetes)
- Experience with Prometheus/Grafana for monitoring
- Knowledge of distributed resource scheduling systems
- Understanding data center networking hardware and cabling concepts.
- Understanding of networking technologies such as DHCP, DNS, TCP/IP, VLANs, HSRP, and SNMP.
- Knowledge of data center networking security principles Firewall ACLs, IPS/IDS, and Policy Based Routing.
All your information will be kept confidential according to EEO guidelines.
- Xcelerate Solution is looking for a mid-career Systems Engineer (Mid-Career) - HPC & GPU Infrastructure with a deep understanding of operating systems, hardware, Kubernetes, and NVIDIA GPU products. As a Systems Engineer (Mid-Career) - HPC & GPU Infrastructure, you will...Suggested
- Xcelerate Solutions in Bethesda, MD, is seeking a Systems Engineer - HPC & GPU Infrastructure (Mid-Career) to design, develop, and optimize GPU clusters for IC customers. This 100% on-site role requires deep knowledge of Linux, HPC hardware, Kubernetes, and NVIDIA GPUs...Suggested
- Leidos Inc in Bethesda, Maryland is seeking a Junior Systems Engineer for HPC and GPU Infrastructure. This role involves designing and optimizing GPU clusters for the IC community customers, requiring a strong background in Linux and hardware integration. The ideal candidate...Suggested
$87.1k - $157.45k
Leidos, located in Bethesda, MD, seeks a Mid-Career Systems Engineer specializing in HPC & GPU Infrastructure. This on-site position involves designing and optimizing GPU clusters for the Intelligence Community. The ideal candidate will have a Bachelor’s degree in Computer...Suggested- Xcelerate Solutions seeks a mid-career Systems Engineer to design, develop, and optimize GPU clusters for IC customers. The role emphasizes HPC, Linux OS integration, and GPU hardware performance across on-premise and cloud environments. Work is on-site at the Intelligence...Suggested
- MAXISIQ, Inc. seeks a mid-career Systems Engineer to design, develop, and optimize GPU clusters for IC customers. This 100% on-site role is based at Bethesda's... ...TS/SCI with polygraph consideration. You will install HPC/GPU hardware, tune Linux performance, and implement job...
$207k - $275k
...CoreWeave combines superior infrastructure performance with deep... ...As a Staff Software Engineer, you will define and drive... ...technical vision for GPU performance validation... ...implementation of scalable systems for validating... ...performance optimization. HPC, distributed computing,...Permanent employmentFull timeTemporary workCasual workWork at officeFlexible hours- Position Summary Support enterprise AI mission systems by designing, developing, and optimizing GPU clusters, with deep focus on operating systems, hardware... ..., and required features. Work closely with AI/ML engineers to integrate GPUs with Linux-based systems....
- RPMGlobal seeks a senior engineer to design, deploy, and optimize GPU clusters for enterprise AI in a secure government environment in Bethesda, MD. You will work with AI/ML engineers to integrate GPUs with Linux, tune drivers, and ensure compliance with DoD security standards...
- Base-2 Solutions seeks a Senior GPU Systems Engineer to design, deploy, and optimize GPU clusters in a secure enterprise environment in Bethesda, MD. You will collaborate across teams to implement Linux-based GPU solutions for demanding AI workloads. You will leverage...
- We Are: The Global AI Infrastructure team is at the center of... ...powers AI platforms, GPU-accelerated workloads,... ...computing solutions, aligning system architecture and... ...along with LLM inference engines (TensorRT-LLM), production... ...+ GPU clusters for AI, HPC, and agentic AI...Work experience placementLive inWork at officeLocal area
- VT Group (VTG) is seeking a High-Performance Computing Engineer in McLean, VA. The role focuses on designing,... ...ideal candidate brings deep expertise in large-scale HPC environments, cluster management, GPU computing, and high-speed networking, with an emphasis...
$200k - $240k
...Suitland, MD, US Date Posted: 2026-09-23 Category: Engineering and Sciences Subcategory: Systems Engineer Schedule: Full-Time Shift: Day Job... ...Varen, an SAIC company is seeking an experienced Infrastructure Systems Engineer to join our team! This position...Full timeWork experience placementRemote workShift work- Systems Engineer - Communications and Infrastructure Jr. Level Olympus Systems Engineer - Communications and Infrastructure - Jr. Level R&K Enterprise Solutions, Inc. (R&K) is a certified Service-Disabled Veteran Owned Small Business (SDVOSB) with core capabilities in...Permanent employmentFull timeContract workApprenticeship
$140k - $160k
Amentum is seeking a highly experienced Senior Systems/Infrastructure Engineer with advanced technical expertise supporting complex enterprise environments to join our team and support our McLean, VA customer. We are looking for team members who are passionate about making...Hourly payContract workFor contractorsWork at officeLocal area- ...Job Description Job Description Classified Infrastructure Systems Engineer Location: Falls Church, VA Enabled Intelligence Enabled Intelligence, Inc. provides extremely accurate, precise and secure data labeling and AI solutions to help our government and...Work at office
$100k - $135k
...We are seeking a Cloud Systems Engineer to support and operate large-scale AI and high-performance computing (HPC) environments. This role will be responsible for... ..., and lifecycle management of GPU-accelerated compute infrastructure that powers critical AI, machine learning...Casual workWork at officeImmediate startWorldwide$152k - $241.5k
Sr Software Engineer - Distributed Systems Engineer, EDA InfrastructureNVIDIA is hiring engineers to build and scale the infrastructure that supports our Electronic Design Automation (EDA) workloads... ...that manage large fleets of GPU-based and CPU-based compute systems...Full timeRemote work$165k - $265k
...life on Mars. SR. SITE RELIABILITY ENGINEER (STARSHIELD) At SpaceX we're leveraging... ..., test, and operate all parts of the system - receivers that allow users to connect... ...engineer focused on Starshield's software and GPU infrastructure, you will design, operate and scale the...Temporary workImmediate startWeekend work- ...Systems Engineer Location Bethesda, MD Job Code 2579 of Openings 1 Apply Now ( DCCA is a veteran-owned IT business specializing... ...enterprises, helping them to feel confident in their IT infrastructure. With DCCA, these organizations can be confident in the...For contractorsFlexible hours
- ...Career Title: Systems Engineer Department: Information Technology Salary: Competitive salary commensurate with experience... ...its Microsoft 365, Azure, Entra, endpoint, and on-premises infrastructure environments. This role provides high-quality technical...Work experience placementWork at officeLocal areaFlexible hours
$75k - $90k
...Ntiva Systems Engineer OpportunityAre you looking for limitless career opportunities with a company that values growth, innovation, and... ...maintaining, upgrading, inspecting and managing client IT infrastructure. This may extend on occasion to ad hoc, priority, or emergency...Temporary workWork experience placementWork at officeRemote workWork visaMonday to Friday$75k - $90k
...grow with us! How you’ll make an Impact As a Systems Engineer, you will be the accountable owner for IT operations governance... ...maintaining, upgrading, inspecting and managing client IT infrastructure. This may extend on occasion to ad hoc, priority, or...Temporary workWork experience placementWork at officeRemote workWork visaMonday to Friday$112.3k - $202.5k
The Senior Systems Administrator/Engineer - Identity and Infrastructure designs, automates, secures, and supports enterprise identity and hybrid infrastructure services across Azure, AWS, and on-premises environments. The role provides senior-level expertise in Active...Full timeTemporary workFor contractorsWork experience placementFor subcontractorLocal areaImmediate start$190k - $200k
...continues to grow. We are actively hiring ElasticSearch Systems Engineer with TS/SCI clearance and polygraph to support a program... ..., security and administration Maintain appropriate infrastructure to maintain performance and data integrity Keep...Contract workWork experience placement- We are seeking a Security Systems Infrastructure Engineer to join our federal team supporting our Electronic Security Systems (ESS) projects. This is a hybrid role with the ability to come in to our Rockville, MD office and other project sites in the Washington D.C. Metro...Work experience placementWork at officeLocal areaWorldwide
- ...Group, is looking for a highly skilled platform engineer with deep expertise in operating systems, hardware, GPU, and high-speed networking. In this role, you... ...performance, efficiency, and feature requirements. Infrastructure as Code: Develop and manage Infrastructure as...Full time
- ...exceptional benefits? RER Solutions, Inc., could be your new home. RER Solutions, Inc. is accepting resumes for a Systems Administrator/Infrastructure Engineer to manage, maintain, and enhance DOE owned systems and enterprise server environments. The Systems...
- ...Job Description: Vexterra is looking to fill a Windows Systems Engineer and Administrator position within the Analysis Solutions Division... ...for preventing reoccurrence and analyzes existing infrastructure for tuning/performance enhancements. The individual will provide...Shift work
- ...Solutions is seeking a Senior GPU Platform Engineer to design, develop, and... ...high-performance computing (HPC) to support mission-critical... ...to other operating systems. Operating System Integration... ...Vetting & Analysis, Critical Infrastructure Protection, Digital Solutions...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Systems Engineer - HPC & GPU Infrastructure. Be the first to apply!
- healthcare systems engineer Bethesda, MD
- mission system engineer Bethesda, MD
- senior linux systems engineer Bethesda, MD
- senior staff systems engineer Bethesda, MD
- system engineer remote Bethesda, MD
- systems engineer Bethesda, MD
- software system engineer Bethesda, MD
- computer system validation engineer Bethesda, MD
- system engineer contract Bethesda, MD
- system performance engineer Bethesda, MD



