SME Platform Engineer
$191.25k - $258.75kFull-time
Gdit
Responsibilities for this Position
Location: USA VA ArlingtonFull Part/Time: Full time
Job Req: RQ226853 Type of Requisition:
Regular Clearance Level Must Currently Possess:
Top Secret/SCI Clearance Level Must Be Able to Obtain:
Top Secret SCI + Polygraph Public Trust/Other Required:
None Job Family:
IT Infrastructure and Operations Job Qualifications: Skills:
CI/CD, Cloud Infrastructure, Cluster Administration, Kubernetes, Linux Server Administration
Certifications:
None
Experience:
15 + years of related experience
US Citizenship Required:
Yes Job Description: YOUR IMPACT Own your opportunity to work with the largest government agency in the nation. Make an impact by advancing the Department of War's mission to keep our country safe and secure. OUR COMPANY Iron EagleX (IEX), a wholly owned subsidiary of General Dynamics Information Technology (GDIT), delivers agile IT and Intelligence solutions. Combining small-team flexibility with global scale, IEX leverages emerging technologies to provide innovative, user-focused solutions that empower organizations and end users to operate smarter, faster, and more securely in dynamic environments. JOB DESCRIPTION Iron EagleX is seeking a SME Platform Engineer to support our Engineering team in Crystal City, VA. This role will lead the design, implementation, and management of our secure, on-premises cloud infrastructure. In this role, you will be the driving force behind our advanced computing environments, ensuring the seamless orchestration of containerized applications and large-scale data science platforms. You will work at the intersection of infrastructure, security, and machine learning, managing robust compute clusters and providing foundational support for AI model training and deployment. The ideal candidate has deep expertise in Kubernetes ecosystem tools, GitOps methodologies, and strict compliance standards. MEANINGFUL WORK AND PERSONAL IMPACT As a SME Platform Engineer, your work will directly empower our data science and engineering teams to push the boundaries of machine learning and data analytics. By building and maintaining resilient GPU and Ray clusters, you will accelerate the fine-tuning and deployment of advanced models. Your commitment to security and compliance will ensure our critical systems remain protected against vulnerabilities, providing a safe, compliant, and highly performant foundation for the organization's most impactful technical initiatives. You will not just be managing infrastructure; you will be enabling innovation. JOB DUTIES (INCLUDE BUT ARE NOT LIMITED TO)
- Infrastructure & Orchestration: Architect, deploy, and manage on-premises cloud infrastructure using RKE2 and maintain storage solutions like Longhorn and Object storage.
- Platform Enablement: Host and maintain robust data science environments, including software such as POSIT Workbench/Connect and Hive Metastore.
- AI/ML Infrastructure: Manage and scale robust GPU clusters, Ray Clusters for fine-tuning machine learning models, and VLLM Routers for efficient model inference.
- CI/CD & Automation: Build, maintain, and optimize CI/CD pipelines using Git, Helm charts, and ArgoCD for reliable software delivery.
- Security & Compliance: Ensure continuous FIPS compliance across the environment. Actively manage and mitigate critical and high-level vulnerabilities.
- Identity & Access: Implement and maintain robust authentication and authorization mechanisms using Keycloak and Open Policy Agent (OPA).
- System Administration: Pull and manage container images from secure registries such as Harbor, Docker Hub, or Containeryard. Manage all core capabilities and troubleshoot issues effectively via the command-line console.
- Demonstrated experience designing, deploying, administering, and troubleshooting production Kubernetes environments; hands-on experience with RKE2 or similar.
- Strong Linux systems administration skills, including the ability to manage, diagnose, and troubleshoot infrastructure and platform services through the command line (CLI).
- Experience implementing authentication and authorization solutions using technologies such as Keycloak, Open Policy Agent (OPA), OIDC, RBAC, or comparable identity and access management frameworks.
- Hands-on experience with Git-based development and deployment workflows, including Helm charts, CI/CD pipelines, and GitOps practices.
- Experience with Argo CD or similar tools for declarative, GitOps-based continuous delivery.
- Experience managing Kubernetes storage solutions, including distributed block storage and object storage; experience with Longhorn or comparable technologies preferred.
- Experience hosting and administering data science or analytics platforms; experience with Posit Workbench, Posit Connect, Hive Metastore, or similar technologies.
- Experience configuring and operating systems in accordance with FIPS or comparable security and compliance requirements.
- Demonstrated experience identifying, prioritizing, and remediating critical and high-severity system and application vulnerabilities.
- Experience administering GPU-enabled compute environments supporting AI/ML, high-performance computing, or other compute-intensive workloads.
- Experience supporting distributed AI/ML workloads using Ray or comparable distributed computing frameworks, including model training and fine-tuning use cases.
- Experience deploying or supporting large language model inference and serving technologies; experience with vLLM and related routing capabilities preferred.
- Experience pulling, managing, securing, and troubleshooting container images using private or public registries such as Harbor, Docker Hub, Container Yard, or equivalent container registry platforms.
- Hands-on experience with Infrastructure as Code (IaC) and configuration management tools such as Terraform, Ansible, or comparable technologies.
- Proficiency in scripting or programming languages such as Python, Go, or Bash to support infrastructure automation, platform operations, and troubleshooting.
- Advanced knowledge of Linux system administration, networking concepts, and protocols within complex or highly available infrastructure environments.
- Experience implementing and maintaining monitoring, logging, and observability solutions using tools such as Prometheus, Grafana, or comparable platforms.
- Familiarity with MLOps practices and the machine learning lifecycle, including model development, deployment, monitoring, versioning, and operational support.
- Experience automating infrastructure provisioning, configuration, deployment, and operational workflows in secure or regulated environments.
- Familiarity with performance tuning, capacity planning, and resource optimization for Kubernetes, GPU, or other compute-intensive environments.
- Clearance: Current TS/SCI Clearance with current or willingness to obtain CI polygraph
- Experience:15+ years of related experience
- Education: Bachelor's degree in Computer Science, Software Engineering, or a related field (or equivalent experience)
- Role requirements: Work is onsite in Crystal City, VA with optional CONUS travel
- Due to US Government Contract Requirements, only US Citizens are eligible for this role
At GDIT, the mission is our purpose, and our people are at the center of everything we do.
- Growth: AI-powered career tool that identifies career steps and learning opportunities
- Support: An internal mobility team focused on helping you achieve your career goals
- Rewards: Comprehensive benefits and wellness packages, 401K with company match, competitive pay and paid time off
- Community: Award-winning culture of innovation and a military-friendly workplace
Explore a career at GDIT and you'll find endless opportunities to grow alongside colleagues who share your passion for the mission and delivering results. #iexjobs #iexpriority Equal Opportunity Employer / Individuals with Disabilities / Protected Veterans The likely salary range for this position is $191,250 - $258,750. This is not, however, a guarantee of compensation or salary. Rather, salary will be set based on experience, geographic location and possibly contractual requirements and could fall outside of this range. Scheduled Weekly Hours:
40 Travel Required:
10-25% Telecommuting Options:
Onsite Work Location:
USA VA Arlington Additional Work Locations: Total Rewards at GDIT:
Our benefits package for all US-based employees includes a variety of medical plan options, some with Health Savings Accounts, dental plan options, a vision plan, and a 401(k) plan offering the ability to contribute both pre and post-tax dollars up to the IRS annual limits and receive a company match. To encourage work/life balance, GDIT offers employees full flex work weeks where possible and a variety of paid time off plans, including vacation, sick and personal time, holidays, paid parental, military, bereavement and jury duty leave. To ensure our employees are able to protect their income, other offerings such as short and long-term disability benefits, life, accidental death and dismemberment, personal accident, critical illness and business travel and accident insurance are provided or available. We regularly review our Total Rewards package to ensure our offerings are competitive and reflect what our employees have told us they value most. Our Identity Verification Process:
As part of the hiring process, we will ask you to complete an identity verification process that leverages advanced biometrics and artificial intelligence to ensure authenticity and protect against identity fraud. You are expected to be on camera during virtual interviews. We reserve the right to take your picture to verify your identity and prevent fraud. By proceeding, you authorize the collection, processing, and use of your biometric data for identity verification and security purposes. About Our Work:
We are GDIT. A global technology and professional services company that delivers technology solutions and mission services to every major agency across the U.S. government, defense and intelligence community. Our 26,000 experts extract the power of technology to create immediate value and deliver solutions at the edge of innovation. We operate across 50+ countries worldwide, offering leading mission-ready capabilities in AI, cloud, cyber and software development. Join our Talent Community to stay up to date on our career opportunities and events at
gdit.com/tc . Equal Opportunity Employer / Individuals with Disabilities / Protected Veterans
PI286619423
YOUR IMPACT
Own your opportunity to work with the largest government agency in the nation. Make an impact by advancing the Department of War's mission to keep our country safe and secure.
OUR COMPANY
Iron EagleX (IEX), a wholly owned subsidiary of General Dynamics Information Technology (GDIT), delivers agile IT and Intelligence solutions. Combining small-team flexibility with global scale, IEX leverages emerging technologies to provide innovative, user-focused solutions that empower organizations and end users to operate smarter, faster, and more securely in dynamic environments.
JOB DESCRIPTION
Iron EagleX is seeking a SME Platform Engineer to support our Engineering team in Crystal City, VA. This role will lead the design, implementation, and management of our secure, on-premises cloud infrastructure. In this role, you will be the driving force behind our advanced computing environments, ensuring the seamless orchestration of containerized applications and large-scale data science platforms. You will work at the intersection of infrastructure, security, and machine learning, managing robust compute clusters and providing foundational support for AI model training and deployment. The ideal candidate has deep expertise in Kubernetes ecosystem tools, GitOps methodologies, and strict compliance standards.
MEANINGFUL WORK AND PERSONAL IMPACT
As a SME Platform Engineer, your work will directly empower our data science and engineering teams to push the boundaries of machine learning and data analytics. By building and maintaining resilient GPU and Ray clusters, you will accelerate the fine-tuning and deployment of advanced models. Your commitment to security and compliance will ensure our critical systems remain protected against vulnerabilities, providing a safe, compliant, and highly performant foundation for the organization's most impactful technical initiatives. You will not just be managing infrastructure; you will be enabling innovation.
JOB DUTIES (INCLUDE BUT ARE NOT LIMITED TO)
- Infrastructure & Orchestration: Architect, deploy, and manage on-premises cloud infrastructure using RKE2 and maintain storage solutions like Longhorn and Object storage.
- Platform Enablement: Host and maintain robust data science environments, including software such as POSIT Workbench/Connect and Hive Metastore.
- AI/ML Infrastructure: Manage and scale robust GPU clusters, Ray Clusters for fine-tuning machine learning models, and VLLM Routers for efficient model inference.
- CI/CD & Automation: Build, maintain, and optimize CI/CD pipelines using Git, Helm charts, and ArgoCD for reliable software delivery.
- Security & Compliance: Ensure continuous FIPS compliance across the environment. Actively manage and mitigate critical and high-level vulnerabilities.
- Identity & Access: Implement and maintain robust authentication and authorization mechanisms using Keycloak and Open Policy Agent (OPA).
- System Administration: Pull and manage container images from secure registries such as Harbor, Docker Hub, or Containeryard. Manage all core capabilities and troubleshoot issues effectively via the command-line console.
REQUIRED SKILLS
- Demonstrated experience designing, deploying, administering, and troubleshooting production Kubernetes environments; hands-on experience with RKE2 or similar.
- Strong Linux systems administration skills, including the ability to manage, diagnose, and troubleshoot infrastructure and platform services through the command line (CLI).
- Experience implementing authentication and authorization solutions using technologies such as Keycloak, Open Policy Agent (OPA), OIDC, RBAC, or comparable identity and access management frameworks.
- Hands-on experience with Git-based development and deployment workflows, including Helm charts, CI/CD pipelines, and GitOps practices.
- Experience with Argo CD or similar tools for declarative, GitOps-based continuous delivery.
- Experience managing Kubernetes storage solutions, including distributed block storage and object storage; experience with Longhorn or comparable technologies preferred.
- Experience hosting and administering data science or analytics platforms; experience with Posit Workbench, Posit Connect, Hive Metastore, or similar technologies.
- Experience configuring and operating systems in accordance with FIPS or comparable security and compliance requirements.
- Demonstrated experience identifying, prioritizing, and remediating critical and high-severity system and application vulnerabilities.
- Experience administering GPU-enabled compute environments supporting AI/ML, high-performance computing, or other compute-intensive workloads.
- Experience supporting distributed AI/ML workloads using Ray or comparable distributed computing frameworks, including model training and fine-tuning use cases.
- Experience deploying or supporting large language model inference and serving technologies; experience with vLLM and related routing capabilities preferred.
- Experience pulling, managing, securing, and troubleshooting container images using private or public registries such as Harbor, Docker Hub, Container Yard, or equivalent container registry platforms.
DESIRED SKILLS
- Hands-on experience with Infrastructure as Code (IaC) and configuration management tools such as Terraform, Ansible, or comparable technologies.
- Proficiency in scripting or programming languages such as Python, Go, or Bash to support infrastructure automation, platform operations, and troubleshooting.
- Advanced knowledge of Linux system administration, networking concepts, and protocols within complex or highly available infrastructure environments.
- Experience implementing and maintaining monitoring, logging, and observability solutions using tools such as Prometheus, Grafana, or comparable platforms.
- Familiarity with MLOps practices and the machine learning lifecycle, including model development, deployment, monitoring, versioning, and operational support.
- Experience automating infrastructure provisioning, configuration, deployment, and operational workflows in secure or regulated environments.
- Familiarity with performance tuning, capacity planning, and resource optimization for Kubernetes, GPU, or other compute-intensive environments.
WHAT YOU'LL NEED TO SUCCEED
- Clearance: Current TS/SCI Clearance with current or willingness to obtain CI polygraph
- Experience:15+ years of related experience
- Education: Bachelor's degree in Computer Science, Software Engineering, or a related field (or equivalent experience)
- Role requirements: Work is onsite in Crystal City, VA with optional CONUS travel
- Due to US Government Contract Requirements, only US Citizens are eligible for this role
GDIT IS YOUR PLACE
At GDIT, the mission is our purpose, and our people are at the center of everything we do.
- Growth: AI-powered career tool that identifies career steps and learning opportunities
- Support: An internal mobility team focused on helping you achieve your career goals
- Rewards: Comprehensive benefits and wellness packages, 401K with company match, competitive pay and paid time off
- Community: Award-winning culture of innovation and a military-friendly workplace
OWN YOUR OPPORTUNITY
Explore a career at GDIT and you'll find endless opportunities to grow alongside colleagues who share your passion for the mission and delivering results.
Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the SME Platform Engineer in Arlington, VA vacancy
$200k - $275k
...country. About You You are a senior cyber infrastructure engineer, security platform engineer, or cyber systems architect with hands‑on experience... ...Infrastructure Engineer & Architect (Security Platforms SME) to support cybersecurity platform engineering, security infrastructure...SuggestedFor contractorsRelocation package$135k - $172k
GovCIO is currently hiring a highly experienced SME Systems Engineer specializing in Identity Management (IdM) and Automation across all software products, with a primary focus on the Power Platform product area. This technical role supports critical Identity, Credential...SuggestedCurrently hiring- ...consulting services. We are in search of a highly motivated candidate to join our talented Team. Job Title: Palo Cloud Firewall Engineer SME. Location: McLean, VA. Job Description:Need Palo Alto Cloud Firewall Engineer Subject Matter Expert (SME) for the high-level...Suggested
$154.05k - $278.48k
Leidos has an exciting opportunity for a DevOps Engineer (SME) in our Intel Security Sector's Analysis Solutions Business Are). Our talented team is at the forefront in Security Engineering, Computer Network Operations (CNO), Mission Software, Analytical Methods and Modeling...SuggestedFull timeWork experience placementImmediate startFlexible hours$62k - $141k
DevOps Platform EngineerThe Opportunity:Are you looking for an opportunity to make a difference? What if you could find a position that is tailor made for your mix of development, engineering, and communication skills? Efficient software development teams make the most...SuggestedFull timeContract workPart timeWork at officeLocal areaRemote work- Position: Mainframe Platform Engineer Location: McLean, VA - 100% ONSITE Duration: 8+ months In-Person Interview Responsibilities Technical... ...involving RACF security. Work as client side technical SME for Mainframe infrastructure subsystems to hold Managed Service...Contract workFlexible hours
$176k - $282k
ResponsibilitiesPeraton is seeking an Information Systems Security Engineer - Subject Matter Expert (SME)/Cloud-based to support its Federal Strategic Cyber programs.Location: National Capital Region (NCR):In this role, you will: Lead, mentor, and supervise a team of contractor...Contract workTemporary workFor contractorsWork at officeShift work$140k - $165k
Job DescriptionEverforth ECS is seeking a Senior Databricks Platform Engineer to work in our Arlington, VA office (Hybrid).We are seeking a highly skilled Senior Databricks Platform Engineer to design, implement, and maintain our enterprise-level Databricks platform supporting...Work at office$101.15k - $192k
...trusted government contractor with decades of experience delivering cutting-edge solutions, is seeking a driven and innovative Platform Engineer to help build and extend IronGate, an ETL solution focused on moving data from commercial environments into secure...Full timeContract workFor contractorsWork at office$184k - $285k
As Sr. Kubernetes Platform Engineer III, you'll provide platform engineering expertise for containerized environments supporting mission-critical workloads, designing and operating Kubernetes platforms while contributing immediately within a disciplined Agile delivery model...Full timeLocal areaImmediate start$148.5k - $223.9k
...Salesforce.We're looking for an experienced and hands-on software engineer to join our team to build and scale the next generation of... ...infrastructure automation tools, frameworks, workflows, and validation platforms, applying AI tools, that help Salesforce scale infrastructure....Full time$176k - $276k
...organization. We deploy, integrate, and operate the Kubernetes-based platform and shared services used to provision, monitor, and operate... ...across environments.We are looking for a hands-on senior engineer to own the lifecycle and automation of the Kubernetes platform...Full timeRemote workWeekend work- Job Family:Data Engineering & Architecture ConsultingTravel Required:NoneClearance Required:Active SecretWhat You Will Do:Focus on engineering, building, and maintaining shared platform capabilities, data pipelines, applications, workflows, automations, and governance tools...Full timeRemote workFlexible hours
$65k - $170k
We are seeking a highly skilled AWS Platform Engineer to build and safeguard our AWS cloud ecosystem. You will be responsible for the end-to-end lifecycle of AWS tenant accounts, ensuring a secure, cost-optimized, and observable environment. You will bridge the gap between...Full timeTemporary workWork at office$108.01k - $183.61k
This role is contingent upon a contract award. ICF is seeking a Cloud Platform Engineer to deploy, configure, and maintain cloud infrastructure for a federal technology program. Reporting to the Cloud / Enterprise Architecture Lead, this role is responsible for building...Full timeContract workWork experience placementWork at officeRemote work$75.2k - $158.1k
Job Title: Hybrid Cloud Platform EngineerJob Category: Information TechnologyTime Type: Full timeMinimum Clearance Required to Start:... ...:Required:Bachelor's Degree in Computer Programming, Science, Engineering or a related technical discipline.5-8 years Developer experienceMeet...Contract workWork experience placementLocal areaFlexible hoursShift work- ...away from the Federal Bureau of Investigation's Criminal Justice Information Services Division's Headquarters. Founded in 2007 by an Engineer-by-trade, Fusion Technology dedicates our valuable resources to providing comprehensive IT services and solutions to mission-...Temporary workWork at office
- ...requires an active U.S. Government Security Clearance at the TS/SCI level with required polygraph.We are seeking a Geospatial Platform Engineer to support geospatial, imagery, AI/ML, and data-driven application development, deployment, and operations. This role will focus...Full timeRemote work
- ...Bowhead is seeking a Chief Cloud Architect & Lead Infrastructure SME to join our team in Arlington, VA, providing expert knowledge of AWS GovCloud and secure cloud infrastructure for the NAUT contract. You will design, implement, and manage complex cloud solutions,...Contract work
- DC Government is seeking an Information Technology Specialist (Devp Ops Engineer) to design and maintain the agency’s software delivery infrastructure, Kubernetes platform, CI/CD pipelines, and monitoring systems supporting DC Health Link. The role covers Kubernetes manifests...
$86.9k - $198k
Cyber Platforms Software Engineer, SeniorThe Opportunity: We build software for large, fast-changing cyber problems where speed, scale, and complexity demand better tools. As a Senior Software Engineer, you’ll work across production systems used by operators, analysts,...Full timeContract workPart timeWork at officeLocal areaRemote work$117.2k - $176.7k
...future of AI, and you are the future of Salesforce.The ExperienceSalesforce is seeking a Software Engineer (MTS) to design and build compliance automation on the Salesforce Platform within Product Security. This role is ideal for a Salesforce Platform Developer who wants to...Full time$174k - $226k
...transforming the financing experience and joining our team?About the roleWe're hiring a hands-on engineering leader to own both the people and the delivery of our platform engineering team. This is a formal people manager role — you are accountable for your team's performance...Work at officeRemote workFlexible hours$180k - $205k
...DescriptionEverforth ECS is seeking a Sr Cloud DevOps Engineer to work in our Arlington, VA office /... ...Content Management as a Service (WCMaaS) platform. This includes managing all aspects of... ...integration points. Serve as a technical SME for cloud services, providers, and...Work at officeRemote work$114.6k - $252.1k
Job Title: AI Cloud Platform EngineerJob Category: Information TechnologyTime Type: Full timeMinimum Clearance Required to Start: NoneEmployee... ...: None* * *The Opportunity:We are seeking an AI and Cloud Engineer to design, build, and scale our next-generation intelligent...Contract workWork experience placementFlexible hours$69.4k - $158k
Salesforce Platform EngineerThe Opportunity:As a full stack developer, you can resolve a problem with a complete end-to-end solution... ...lifecycle Bachelors degree, or 5+ years of experience in software engineering in lieu of degree Nice If You Have:Experience supporting Sales...Full timeContract workPart timeWork at officeLocal areaRemote work- IRS CAS (Sr. Full Stack Developer/Platform Engineer)Job Overview:Customer: Internal Revenue Services (IRS) Location: Remote, with occasional onsite support for meetings, deployments, production activities, or knowledge transfer as required.Clearance: No clearance required...Contract workFor contractorsFor subcontractorRemote workWorldwideRelocation package
- ...Platform Engineer Job Locations US-VA-Tysons Job ID 2026-13778 # of Openings 1 Benefit Type Salaried High Fringe/Full-Time Overview We are looking for a Platform Engineer to design, build, and operate the foundational...Full timeContract workLocal area
- ...About the job AWS Cloud Platform Engineer AWS Cloud Platform Engineer Location: Hybrid - Reston, VA (Onsite 1x/week; local DC/... ...initiatives within a large-scale AWS environment. This is a hands-on SME-level role requiring deep expertise in EKS-based provisioning,...Temporary workLocal area
$92.5k - $209.5k
...services in a distributed, multi-tenant cloud environment. OCI gives engineers the opportunity to work on systems that directly power... ...build resilient production services, and improve the internal platforms that enable OCI teams to build, validate, secure, and release...Temporary workFixed term contractFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to SME Platform Engineer. Be the first to apply!
Related searches
- platform engineer Arlington, VA
- platform developer Arlington, VA
- senior platform engineer Arlington, VA
- digital platform specialist Arlington, VA
- power platform Arlington, VA
- director of digital platform Arlington, VA
- platform product manager Arlington, VA
- platform manager Arlington, VA
- platform engineer
- platform engineering manager

