Distributed Systems ML Infrastructure Engineer
$145k - $250kOpenTeams
Who We Are
Every organization runs on intelligence: years of accumulated knowledge, decisions, and context. As AI takes on more of that work, companies face a choice: rent that intelligence from vendors who keep the data, the context, and the results, or own it.
OpenTeams exists to make ownership possible. Founded by Travis Oliphant, creator of NumPy and SciPy, and built by people with deep roots across the open-source ecosystem, including NumPy, SciPy, PyTorch, and Jupyter, we help enterprises and governments build AI they control, govern, and evolve themselves. If that sounds like your kind of work, we'd like to meet you.Distributed Systems ML Infrastructure Engineer
Location: Washington, DC; Denver, CO; or Colorado Springs, CO preferred (hybrid). Highly qualified candidates outside these locations may also be considered for unclassified work.
Work Authorization: U.S. citizenship required
Clearance: An active TS/SCI clearance with CI polygraph is strongly preferred. Candidates without an active clearance may be considered for unclassified work but must be eligible to obtain and maintain a U.S. security clearance.
Salary Range: $145,000–$250,000 USD, dependent on experience level and location
About the Role
We're looking for a Distributed Systems and ML Infrastructure Engineer to build the core services of a containerized, API-first AI platform. This is a role for someone who wants to build the thing itself, not integrate someone else's.
You design and implement the services the platform runs on — workflow orchestration, data ingestion, results management, model serving, policy enforcement, usage accounting, audit logging. Those services have to hold up across cloud, dedicated, isolated, and limited-connectivity deployments, which means portability and operability are design constraints from the first commit rather than problems handed to someone downstream.
Development happens primarily on unrestricted infrastructure with an open-source toolchain. Engineers with the right access also carry releases into controlled production environments, integrate data sources there, and validate the platform in place — so there's a path to seeing your work through to where it actually runs.
This position is contingent upon contract award. Travel of up to 15% may be required, primarily to Government facilities and between company locations. Unclassified work may be performed remotely, while classified promotion and validation activities require onsite work in an accredited facility and the appropriate security clearance.
Key Responsibilities
- Design and implement platform services for workflow orchestration, data ingest, and results management, exposed through documented APIs with no proprietary front end
- Implement and operate model gateway and serving services that route invocations to approved managed model services with policy enforcement, usage accounting, and audit logging
- Maintain a documented provider abstraction so the platform runs on AWS-native managed services where appropriate while remaining deployable across other cloud and dedicated environments
- Deploy platform releases into classified host environments, perform data source integration, and execute validation procedures on a recurring promotion cadence
- Verify environment parity after each promotion
- Size and validate the platform against documented workload models, and verify capacity and performance by load test
- Constrain platform dependencies to services confirmed available in the target environments, and gate any development-only dependency behind feature flags
- Reproduce high-side defects on the low side through sanitized feedback paths and fix them where the full toolchain is available
Required Skills & Experience
- U.S. citizenship and eligibility to obtain and maintain a U.S. security clearance
- 6+ years of experience in distributed systems, platform engineering, infrastructure engineering, or a related software engineering role
- Production experience operating Kubernetes and containerized workloads on a major cloud platform
- Experience with managed Kubernetes services such as Amazon EKS or an equivalent platform
- Experience supporting machine learning workloads in production, such as model serving, GPU scheduling, or large-scale data and evaluation pipelines
- Experience designing, building, or operating distributed services that support reliability, scalability, and performance requirements
- Proficiency in Python, Go, or a comparable programming language
- Experience with infrastructure-as-code and deployment tools such as Terraform, Helm, or equivalent technologies
- Experience designing API-first services and implementing documented interface specifications
- Experience testing platform capacity and performance against expected workload requirements
- Ability to document technical interfaces, deployment procedures, architectural decisions, and validation results
- Bachelor’s degree in computer science, engineering, or a related field, or equivalent practical experience
Nice to Have
- Active TS/SCI clearance with CI polygraph
- Hands-on experience deploying or operating software in classified, air-gapped, or limited-connectivity environments
- Experience with classified cloud environments, including AWS Secret or Top Secret regions
- Experience building or operating model gateways, LLM routing layers, inference services, or inference-brokering platforms
- Familiarity with software promotion into classified environments, cross-domain transfer processes, and release packaging
- Experience with agentic workflow frameworks or the orchestration of multi-step AI pipelines
- Experience supporting GPU-accelerated workloads and distributed model inference
- Experience designing vendor-agnostic platforms that operate across multiple cloud or dedicated environments
- Experience supporting rapid prototyping programs or defense innovation initiatives
What We Offer
- Medical, Dental & Vision – 100% paid for employees, 75% for dependents
- 401(k) Match – Up to 5% with full vesting after 2 years
- Unlimited PTO – With a required minimum of 15 days off annually
- Fully Remote Setup – Includes up to $3,000 equipment reimbursement
- Continuous Education – Includes up to $500 reimbursement
- Disability & Life Insurance – 100% employer-paid
- HSA & FSA Options – With monthly HSA contributions from OpenTeams
Grow With Us
At OpenTeams, growth isn’t just about the company—it’s about you.
We believe the best careers are built at the edge of your potential. That is where new tools, ideas, and technologies change the world. Here, you’ll work alongside pioneers of AI, solving problems that matter: making AI more transparent, more ethical, and more empowering. As your skills grow, our career framework provides a pathway and recognition of that increased impact.Opportunities aren’t limited by geography. You’ll collaborate with global experts, contribute to open source projects that power the world’s technology, and stretch your skills daily. That global perspective and diversity makes our solution more universal and robust. We are committed to continuing to celebrate diversity on our team.
Supported people are successful people. We offer 100% employer paid medical premiums for employees and self-managed PTO with a minimum time off requirement, so that our teams are able to do their best work.
We invest in curiosity, creativity, and ownership. That means you’ll be trusted to boldly innovate, supported to learn fast, and celebrated for successful collaboration.Commitment to diversity, equity, inclusion, and belonging
OpenTeams understands that valuing diverse creative practices and forms of knowledge is crucial to and enriches the company’s core mission. We encourage applications from everyone, including members of all equity-seeking communities, such as (but certainly not limited to) women, racialized and Indigenous persons, disabled people, persons of all sexual orientations, gender identities and expressions.
We are an equal opportunity employer - all qualified applicants will receive equal consideration for recruitment, interviews, employment, training, compensation, promotion, and related activities. We do not discriminate based on race, religion, gender, gender identity, gender expression, color, national origin, pregnancy, ancestry, domestic partner status, disability, sexual orientation, age, genetic predisposition, medical condition, marital status, citizenship status, military or veteran status, or any other basis covered by applicable laws. OpenTeams will not tolerate discrimination or harassment based on these characteristics or any other unlawful behavior, conduct, or purpose.
$97.5k - $209.5k
...experience designing and building distributed systems and large-scale... ...development environment AI/ML experience: Proven track... ...technical guidance to junior engineers and peers Drive adoption... ...brings together the data, infrastructure, applications, and expertise...SuggestedTemporary workFlexible hours$135k - $175k
...real operational value. The Senior AI/ML Platform Engineer will help build and operate the... ...platform engineering role focused on the systems, patterns, environments, controls, and... ...production support. Implement CI/CD, infrastructure automation, environment management, secrets...SuggestedFull timeWork experience placementWork at officeLocal areaRemote workWorldwideRelocationFlexible hours- ...and implement automation of technical systems infrastructure solutions to support development and... ...management for large-scale fault tolerant and distributed virtual machine and container farms supporting various advanced IP video engineering teams.Engineer automated deployment...SuggestedWork at officeImmediate start
- ...Lead Systems Engineer (Rust) - AI PlatformWhat if your mastery of Rust could directly shape the infrastructure powering the next generation of AI? We're... ...systemsFamiliarity with AI/ML workflows, model... ...infrastructureExperience building distributed systems or developer...SuggestedHourly payOngoing contractContract workRemote workFlexible hours
$130k - $170k
...seeking a Senior Object Storage Software Engineer to design, implement, and optimize the object storage data path and core distributed services of our scale-out object storage platform... ...across the cluster. This is a deep systems role for engineers passionate about low-...SuggestedWork at officeRemote work- Principal Senior Systems Engineer - HPE Networking (Mid-Telco Accounts)This role has been designated as ‘Remote/Teleworker’, which means... ..., broadband providers, utilities, and other network infrastructure customers.As the trusted technical advisor, you will partner...Full timeWork experience placementWork at officeLocal areaImmediate startRemote workWork from home
$95k
...global XFNs to ensure seamless network operations. Support infrastructure initiatives and contribute to the design and implementation... ...cable and connector types, optical components, and transport systems. Demonstrated experience in preparing RCA reports for network...Work at officeWorldwide$175k - $230k
...Intelligence (AI) and Machine Learning (ML) Engineer - Subject Matter Expert (SME) Location... ...enterprise and mission-focused AI/ML systems, platforms, and capabilities. Develop,... ...methodologies, data architectures, APIs, distributed systems, and modern application...Full time- ...Machine Learning Engineer (Llama AI Platform) Location: Remote (... ...automation, and connected business systems. We are building a next-... ...performance, latency, and infrastructure utilization. Monitor... ...tuning open-source LLMs. ML Engineering and MLOps practices...Full timeRemote work
- Systems Engineering Manager - HPE Networking (Mid-Telco Accounts)This role has been designated as ‘Remote/Teleworker’, which means you will... ..., and other critical network operators to modernize infrastructure, simplify operations, and accelerate innovation.We're looking...Full timeWork experience placementLocal areaImmediate startRemote workWork from home
$144.7k - $261.3k
...34 Job Description The Senior ML Validation Research Engineer will lead applied machine learning... ...in robotics and autonomous driving systems. This role centers on simulation-based... ...predictors, uncertainty and Out-of-Distribution detection methods for autonomy ML systems...Full timeLocal areaRemote workWork from homeFlexible hours$142.4k - $180k
Title:Big Data Systems EngineerBelong, Connect, Grow, with KBR... ...NSS) team provides high-end engineering and advanced technology... ...configuration of their big data infrastructure. Must have strong... ...Testing Capability (JMETC) and distributed testing and training.Experience...Full timeContract workTemporary workLocal areaRemote workWork from homeRelocation package$132.54k - $165.71k
...The Systems Software Developer 5 will serve as a technical leader and innovator in our organization. As a Principal Software Developer... ...app development, or backend development. • Mentor and coach engineers at all levels, fostering technical growth and innovation. • Serve...Full timeWork experience placement- ...the only vertically integrated AI infrastructure company built from the ground up, we... ...infrastructure.We are seeking a Principal Systems Integration Engineer to join our R&D team as the senior... ...spanning power generation and distribution, cooling technologies (including...Temporary work
- ...Network Engineer / Systems Administrator (SA) LOCATION Aurora, CO 80014 CLEARANCE TS/SCI Full Poly (Please note this position... ...play a pivotal role in maintaining and enhancing our IT infrastructure. In this role, you will be responsible for designing, implementing...Temporary workFor contractorsImmediate startFlexible hours
- ...Senior Platform Infrastructure EngineerVantor is forging the new frontier... ...Platform Infrastructure Engineer to design, deploy, and operate... ...troubleshooting production systems, and supporting customer-facing... ...defining centralized vs distributed platform engineering operating...Work at office
- ...resolve complex Tier 3/4 issues across commercial video and internet platforms. You will work with hardware, software, and RF/QAM systems and collaborate with internal teams and field leadership to ensure seamless deployment and ongoing stability. The role emphasizes diagnostic...
$156k - $195k
...AI at work. We’re looking for a People Systems Developer to design and deploy AI-native... ...Are3-5 years of experience in software engineering, systems engineering, or advanced... ...complexity at scale. It brings applications, infrastructure, data, models, and security into one...Work at office$120k - $150k
...and technically deep Sr. Microsoft Azure Engineer to join the Microsoft Operations team... ...role. You will own the Azure platform, infrastructure, identity governance, security posture,... ...Science, Information Technology, Information Systems, or a directly related field; OR...Full timeWork at officeLocal area$172.5k - $260.1k
...technical leader to serve as a Lead Software Engineer within the DET IT Finance Engineering... ...the intersection of enterprise finance systems, integration engineering, AI-powered... ...and operate cloud-native integration infrastructure on AWS and/or GCP, modernizing legacy patterns...Full time$228k - $253k
...Principal Machine Learning Engineer to join our Core Data... ...complex, high-impact ML initiatives spanning... ...models, data systems, and large-scale ML platforms... ...in larger technology infrastructure.Be the leading contributor... ...working with distributed big-data tools and event...Full timeWork at officeImmediate startRemote workRelocationRelocation packageFlexible hours- Aurora, CO LOCATION Aurora, CO 80014 Network Engineer / Systems Administrator (SA) LOCATION Aurora, CO 80014 CLEARANCE TS/SCI Full Poly (... ...and play a pivotal role in maintaining and enhancing our IT infrastructure. In this role, you will be responsible for designing,...Temporary workFor contractorsImmediate startFlexible hours
- ...MISSIONAs a Staff Mission Network Engineer, you will be a critical... ...and cloud network infrastructure supporting classified U.S. Government... ...environments across geographically distributed sites Deploy and configure... ...for cutting-edge space systems with national security impact...Permanent employmentWork at office3 days per week
- ...Description & Qualifications Are you a Network Engineer looking for a place to make an impact... ...in support of our Public Safety Systems (PSS) program with the United States Navy... ...deployments, including controller-based and distributed architectures Conduct site surveys,...Full timeContract workPart timeLocal areaRemote workFlexible hours
- ...Responsible for data center engineering by working on architecture,... ...center Spine-leaf network infrastructure utilizing a multi-vendor strategy... ..., including monitoring systems, inventory systems, and... ...use only and is not to be distributed publicly, or to any third party...Local area
- ...Network Engineer 3 (TS/SCI with Poly) Location: San Diego, CA and Aurora, CO (2 openings... ...remediate equipment issues to support system build-out and sell-off. Support the existing... ...connectivity and support large Client distribution Collaborate with the government...Day shift
- ...Job Description Job Description Software Engineer (Systems): Designs, develops, troubleshoots and debugs software programs for enhancements... .... Develops software and tools in support of design, infrastructure and technology platforms, including operating systems, compilers...
$103.6k - $155.4k
...employees have incredible opportunities to work on revolutionary systems that impact people's lives around the world today, and for... ...interaction and presentation skills.Experience mentoring junior engineers or leading small project teams.Primary Level Salary Range: $103...Full timeRelocation packageShift work$69.1k - $141.5k
Job Title: Systems Integrator - MidJob Category: EngineeringTime Type: Full timeMinimum... ...full stack — applications, data, and infrastructure — in a mission-critical setting. Grow alongside... ...senior integrators and Government engineers.Responsibilities:Deploy and integrate...Contract workWork experience placementFlexible hours$82.1k - $172.4k
Job Title: Senior Systems IntegratorJob Category: Information TechnologyTime Type: Full... ...owners, operations teams, and program engineers.Responsibilities:Lead integration and deployment... ...applications, data, networks, and infrastructure. Coordinate integration activities,...Contract workWork experience placementFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Distributed Systems ML Infrastructure Engineer. Be the first to apply!
- senior staff systems engineer Denver, CO
- application system engineer Denver, CO
- system engineer contract Denver, CO
- operations support system engineer Denver, CO
- sr systems engineer Denver, CO
- system performance engineer Denver, CO
- systems engineering technician Denver, CO
- software system engineer Denver, CO
- mission system engineer Denver, CO
- systems engineer Denver, CO



