Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Systems Engineer - HPC & GPU Infrastructure

FiveInsights

Job Description

Job Description

Job Description

Our partner, Vexterra Group, is looking for a mid-career Systems Engineer (Mid-Career) – HPC & GPU Infrastructure with a deep understanding of operating systems, hardware, Kubernetes, and NVIDIA GPU products. As a Systems Engineer (Mid-Career) – HPC & GPU Infrastructure, you will play a pivotal role in designing, developing, and optimizing GPU clusters for the IC community customers. This is a 100% on-site position. All work must be performed at the customer site in Bethesda at the Intelligence Community Campus.

Primary Responsibilities

1. HPC and GPU environment engineering: Contribute to the installation and maintenance of GPU

and HPC hardware on-prem and in the cloud, providing insights into hardware performance to

ensure efficient interaction with software components.

2. Performance Optimization: Analyze HPC/GPU cluster performance, identify bottlenecks, and

develop strategies to enhance performance across various applications in Linux, addressing both

hardware and software considerations. Regularly monitor and improve performance.

3. HPC/GPU tooling: Install and configure HPC/GPU job scheduling and workload management

platforms such as Slurm, PBS , Apache Airflow, Kubernetes

4. Power Efficiency: Work on power management techniques to optimize GPU power

consumption, ensuring efficient operation on both mobile and desktop Linux platforms.

Continuously assess and enhance power efficiency strategies.

5. Testing and Validation: Design and execute tests to validate GPU performance and functionality

on Linux, including stress testing, benchmarking, and debugging to ensure robust operation.

Maintain and expand the testing suite.

6. Documentation: Maintain comprehensive technical documentation, including architectural

specifications, code documentation, and Linux-specific best practices for GPU development.

Keep documentation up to date with changes and improvements.

7. Industry Insight: Stay updated on the latest trends, innovations, and competitive landscapes

within the GPU industry, contributing to research efforts and proposing Linux-specific

approaches to GPU design and optimization. Share regular updates and insights with the team.

Qualifications

Basic Qualifications

  • Bachelor's or higher degree in Computer Science, Electrical Engineering, or a related field.
  • Additional years of experience may be considered in lieu of a degree.
  • 4+ years of relevant systems engineering experience
  • Expertise in operating system integration for Linux.
  • Strong understanding of computer hardware architecture, particularly as it relates to Linux systems.
  • Knowledge of parallel computing, graphics algorithms, and real-time rendering in Linux environments.
  • Excellent problem-solving skills and the ability to collaborate within a team.
  • Strong communication skills for conveying technical information in a Linux context.
  • Proficiency with scripting languages such as Python or BASH.
  • Proficiency with automation tools such Ansible, Puppet, Salt, Terraform, etc.
  • Candidate must, at a minimum, meet DoD 8570.11- IAT Level II certification requirements (currently Security+ CE, CCNA-Security, GICSP, GSEC, or SSCP along with an appropriate computing environment (CE) certification). An IAT Level III certification would also be acceptable (CASP+, CCNP Security, CISA, CISSP, GCED, GCIH, CCSP).
  • TS/SCI clearance with Polygraph required or a TS/SCI and willingness to obtain a Polygraph prior to starting.

Preferred Qualifications

  • Knowledge of GPU virtualization, cloud computing, and emerging Linux-based technologies in the field.
  • Experience with container technologies (Docker, Kubernetes)
  • Experience with Prometheus/Grafana for monitoring
  • Knowledge of distributed resource scheduling systems
  • Understanding data center networking hardware and cabling concepts.
  • Understanding of networking technologies such as DHCP, DNS, TCP/IP, VLANs, HSRP, and SNMP.
  • Knowledge of data center networking security principles Firewall ACLs, IPS/IDS, and Policy Based Routing.

Additional Information

All your information will be kept confidential according to EEO guidelines.

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Systems Engineer - HPC & GPU Infrastructure in Bethesda, MD vacancy
  • Xcelerate Solution is looking for a mid-career Systems Engineer (Mid-Career) - HPC & GPU Infrastructure with a deep understanding of operating systems, hardware, Kubernetes, and NVIDIA GPU products. As a Systems Engineer (Mid-Career) - HPC & GPU Infrastructure, you will... 
    Suggested

    Xcelerate Solutions

    Bethesda, MD
    2 days ago
  • Xcelerate Solutions in Bethesda, MD, is seeking a Systems Engineer - HPC & GPU Infrastructure (Mid-Career) to design, develop, and optimize GPU clusters for IC customers. This 100% on-site role requires deep knowledge of Linux, HPC hardware, Kubernetes, and NVIDIA GPUs... 
    Suggested

    VMD Corp

    Bethesda, MD
    2 days ago
  • Leidos Inc in Bethesda, Maryland is seeking a Junior Systems Engineer for HPC and GPU Infrastructure. This role involves designing and optimizing GPU clusters for the IC community customers, requiring a strong background in Linux and hardware integration. The ideal candidate... 
    Suggested

    Leidos Inc

    Bethesda, MD
    4 days ago
  • $87.1k - $157.45k

    Leidos, located in Bethesda, MD, seeks a Mid-Career Systems Engineer specializing in HPC & GPU Infrastructure. This on-site position involves designing and optimizing GPU clusters for the Intelligence Community. The ideal candidate will have a Bachelor’s degree in Computer... 
    Suggested

    Leidos

    Bethesda, MD
    1 day ago
  • Xcelerate Solutions seeks a mid-career Systems Engineer to design, develop, and optimize GPU clusters for IC customers. The role emphasizes HPC, Linux OS integration, and GPU hardware performance across on-premise and cloud environments. Work is on-site at the Intelligence... 
    Suggested

    Xcelerate Solutions

    Bethesda, MD
    2 days ago
  • MAXISIQ, Inc. seeks a mid-career Systems Engineer to design, develop, and optimize GPU clusters for IC customers. This 100% on-site role is based at Bethesda's...  ...TS/SCI with polygraph consideration. You will install HPC/GPU hardware, tune Linux performance, and implement job... 

    MAXISIQ, Inc.

    Bethesda, MD
    14 hours ago
  • $207k - $275k

     ...CoreWeave combines superior infrastructure performance with deep...  ...As a Staff Software Engineer, you will define and drive...  ...technical vision for GPU performance validation...  ...implementation of scalable systems for validating...  ...performance optimization. HPC, distributed computing,... 
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Flexible hours

    Coreweave

    Washington DC
    more than 2 months ago
  • Position Summary Support enterprise AI mission systems by designing, developing, and optimizing GPU clusters, with deep focus on operating systems, hardware...  ..., and required features. Work closely with AI/ML engineers to integrate GPUs with Linux-based systems.... 

    Base-2 Solutions

    Bethesda, MD
    3 days ago
  • RPMGlobal seeks a senior engineer to design, deploy, and optimize GPU clusters for enterprise AI in a secure government environment in Bethesda, MD. You will work with AI/ML engineers to integrate GPUs with Linux, tune drivers, and ensure compliance with DoD security standards... 

    RPMGlobal

    Bethesda, MD
    2 days ago
  • Base-2 Solutions seeks a Senior GPU Systems Engineer to design, deploy, and optimize GPU clusters in a secure enterprise environment in Bethesda, MD. You will collaborate across teams to implement Linux-based GPU solutions for demanding AI workloads. You will leverage... 

    Base-2 Solutions

    Bethesda, MD
    1 day ago
  • We Are: The Global AI Infrastructure team is at the center of...  ...powers AI platforms, GPU-accelerated workloads,...  ...computing solutions, aligning system architecture and...  ...along with LLM inference engines (TensorRT-LLM), production...  ...+ GPU clusters for AI, HPC, and agentic AI... 
    Work experience placement
    Live in
    Work at office
    Local area

    Accenture

    Arlington, VA
    2 days ago
  • VT Group (VTG) is seeking a High-Performance Computing Engineer in McLean, VA. The role focuses on designing,...  ...ideal candidate brings deep expertise in large-scale HPC environments, cluster management, GPU computing, and high-speed networking, with an emphasis... 

    VT Group (VTG)

    Mc Lean, VA
    1 day ago
  • $200k - $240k

     ...Suitland, MD, US Date Posted: 2026-09-23 Category: Engineering and Sciences Subcategory: Systems Engineer Schedule: Full-Time Shift: Day Job...  ...Varen, an SAIC company is seeking an experienced Infrastructure Systems Engineer to join our team! This position... 
    Full time
    Work experience placement
    Remote work
    Shift work

    SAIC

    Silver Hill, MD
    3 days ago
  • Systems Engineer - Communications and Infrastructure Jr. Level Olympus Systems Engineer - Communications and Infrastructure - Jr. Level R&K Enterprise Solutions, Inc. (R&K) is a certified Service-Disabled Veteran Owned Small Business (SDVOSB) with core capabilities in... 
    Permanent employment
    Full time
    Contract work
    Apprenticeship

    R&K Enterprise Solutions

    Arlington, VA
    4 hours ago
  • $140k - $160k

    Amentum is seeking a highly experienced Senior Systems/Infrastructure Engineer with advanced technical expertise supporting complex enterprise environments to join our team and support our McLean, VA customer. We are looking for team members who are passionate about making... 
    Hourly pay
    Contract work
    For contractors
    Work at office
    Local area

    Amentum

    Mc Lean, VA
    1 day ago
  •  ...Job Description Job Description Classified Infrastructure Systems Engineer Location: Falls Church, VA Enabled Intelligence Enabled Intelligence, Inc. provides extremely accurate, precise and secure data labeling and AI solutions to help our government and... 
    Work at office

    Enabled Intelligence

    Falls Church, VA
    11 days ago
  • $100k - $135k

     ...We are seeking a Cloud Systems Engineer to support and operate large-scale AI and high-performance computing (HPC) environments. This role will be responsible for...  ..., and lifecycle management of GPU-accelerated compute infrastructure that powers critical AI, machine learning... 
    Casual work
    Work at office
    Immediate start
    Worldwide

    Alarm.com

    McLean, VA
    1 day ago
  • $152k - $241.5k

    Sr Software Engineer - Distributed Systems Engineer, EDA InfrastructureNVIDIA is hiring engineers to build and scale the infrastructure that supports our Electronic Design Automation (EDA) workloads...  ...that manage large fleets of GPU-based and CPU-based compute systems... 
    Full time
    Remote work

    Nvidia

    Washington DC
    1 day ago
  • $165k - $265k

     ...life on Mars. SR. SITE RELIABILITY ENGINEER (STARSHIELD) At SpaceX we're leveraging...  ..., test, and operate all parts of the system - receivers that allow users to connect...  ...engineer focused on Starshield's software and GPU infrastructure, you will design, operate and scale the... 
    Temporary work
    Immediate start
    Weekend work

    SpaceX

    Washington DC
    14 hours ago
  •  ...Systems Engineer Location Bethesda, MD Job Code 2579 of Openings 1 Apply Now ( DCCA is a veteran-owned IT business specializing...  ...enterprises, helping them to feel confident in their IT infrastructure. With DCCA, these organizations can be confident in the... 
    For contractors
    Flexible hours

    DCCA

    Bethesda, MD
    2 days ago
  •  ...Career Title: Systems Engineer Department: Information Technology Salary: Competitive salary commensurate with experience...  ...its Microsoft 365, Azure, Entra, endpoint, and on-premises infrastructure environments. This role provides high-quality technical... 
    Work experience placement
    Work at office
    Local area
    Flexible hours

    American Nurses Association ANA

    Silver Spring, MD
    3 days ago
  • $75k - $90k

     ...Ntiva Systems Engineer OpportunityAre you looking for limitless career opportunities with a company that values growth, innovation, and...  ...maintaining, upgrading, inspecting and managing client IT infrastructure. This may extend on occasion to ad hoc, priority, or emergency... 
    Temporary work
    Work experience placement
    Work at office
    Remote work
    Work visa
    Monday to Friday

    Ntiva

    Bethesda, MD
    14 hours ago
  • $75k - $90k

     ...grow with us! How you’ll make an Impact As a Systems Engineer, you will be the accountable owner for IT operations governance...  ...maintaining, upgrading, inspecting and managing client IT infrastructure. This may extend on occasion to ad hoc, priority, or... 
    Temporary work
    Work experience placement
    Work at office
    Remote work
    Work visa
    Monday to Friday

    Ntiva

    Bethesda, MD
    2 days ago
  • $112.3k - $202.5k

    The Senior Systems Administrator/Engineer - Identity and Infrastructure designs, automates, secures, and supports enterprise identity and hybrid infrastructure services across Azure, AWS, and on-premises environments. The role provides senior-level expertise in Active... 
    Full time
    Temporary work
    For contractors
    Work experience placement
    For subcontractor
    Local area
    Immediate start

    Financial Industry Regulatory Authority , Inc.

    Rockville, MD
    4 days ago
  • $190k - $200k

     ...continues to grow.   We are actively hiring ElasticSearch Systems Engineer with TS/SCI clearance and polygraph to support a program...  ..., security and administration Maintain appropriate infrastructure to maintain performance and data integrity Keep... 
    Contract work
    Work experience placement

    Acclaim Technical Services

    Bethesda, MD
    a month ago
  • We are seeking a Security Systems Infrastructure Engineer to join our federal team supporting our Electronic Security Systems (ESS) projects. This is a hybrid role with the ability to come in to our Rockville, MD office and other project sites in the Washington D.C. Metro... 
    Work experience placement
    Work at office
    Local area
    Worldwide

    Johnson Controls

    Washington DC
    25 days ago
  •  ...Group, is looking for a highly skilled platform engineer with deep expertise in operating systems, hardware, GPU, and high-speed networking. In this role, you...  ...performance, efficiency, and feature requirements.  Infrastructure as Code: Develop and manage Infrastructure as... 
    Full time

    MAXISIQ, Inc.

    Bethesda, MD
    10 days ago
  •  ...exceptional benefits? RER Solutions, Inc., could be your new home. RER Solutions, Inc. is accepting resumes for a Systems Administrator/Infrastructure Engineer to manage, maintain, and enhance DOE owned systems and enterprise server environments. The Systems... 

    RER Solutions, Inc.

    Washington DC
    a month ago
  •  ...Job Description: Vexterra is looking to fill a Windows Systems Engineer and Administrator position within the Analysis Solutions Division...  ...for preventing reoccurrence and analyzes existing infrastructure for tuning/performance enhancements.  The individual will provide... 
    Shift work

    Vexterra Group

    Bethesda, MD
    a month ago
  •  ...Solutions is seeking a Senior GPU Platform Engineer to design, develop, and...  ...high-performance computing (HPC) to support mission-critical...  ...to other operating systems.  Operating System Integration...  ...Vetting & Analysis, Critical Infrastructure Protection, Digital Solutions... 

    Xcelerate Solutions

    Bethesda, MD
    a month ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Systems Engineer - HPC & GPU Infrastructure. Be the first to apply!