Data & ML Engineer
RedCell Corporation
Data & ML Engineer
Remote, USA
About Us
Red Cell Partners is an incubation firm building and investing in rapidly scalable technology-led companies that are bringing revolutionary advancements to market in three distinct practice areas: healthcare, cyber, and national security. United by a shared sense of duty and deep belief in the power of innovation, Red Cell is developing powerful tools and solutions to address our Nation's most pressing problems.
ABOUT DEFCON AI
RESILIENCE IN THE FACE OF DISRUPTION. DEFCON AI is an insights company that leverages artificial intelligence, mathematical optimization, data analytics, and software engineering for resilient optimization of complex systems. In today's dynamically changing world, DEFCON AI's technology aligns outcomes with operational goals, better decision making, and empowers customers to anticipate assess, and mitigate the impacts of disruptions.
About the Role
As a Data & ML Engineer you will build the data and model layer behind an AI-enabled decision-support system operating inside an accredited environment. That work covers ingestion from many source systems, resolution of incoming records against a shared data model, relevance scoring, and generation of explanations a user can act on and defend.
Three characteristics make this a substantial technical challenge. The incoming data is predominantly low-signal, which means a model can report strong overall accuracy while failing on the cases that matter most. Every output must remain traceable to the underlying sources, because a person downstream is accountable for the result. Record matching is probabilistic rather than exact, so false matches and missed matches both carry meaningful cost.
You will not be starting from an empty repository. We operate an established platform for source custody, extraction, and retrieval, and its architect is a member of this team, so existing design decisions are documented and accessible. Your work will focus on new capability rather than maintenance: record matching, calibrated scoring, and grounded generation, hardened for the target environment. We build with current tooling and expect the same, including the use of AI assistance in our own engineering practice.
This is a fully remote role with occasional travel (up to 25%) to DEFCON AI HQ, customer sites, and vendor facilities as required.
Key Responsibilities
The technical work falls into four areas. Deep expertise in all four is not expected, so please indicate where your depth lies when you apply. The engineering standards that follow apply to everyone on the team.
Data Modeling and Record Matching
- Design and maintain the graph of entities, records, and the typed relationships between them
- Implement probabilistic matching, including blocking, candidate generation, pairwise scoring, clustering, and threshold policy
- Build deduplication and known-record suppression
- Establish provenance so that every node and edge traces to the source that asserted it
- Produce interface and data-flow design documentation detailed enough to serve as an implementation reference for other engineers
Scoring and Calibration
- Develop relevance and priority models over large, imperfect record sets
- Own calibration and threshold design, establishing what a score means rather than only how it ranks
- Design abstention policy that routes uncertain and high-risk cases to a person rather than returning a confident answer
- Perform feature engineering, establish baselines before introducing complex models, and conduct error analysis that accounts for the differing cost of false positives and false negatives
Retrieval and Generation
- Implement embeddings, vector storage, and retrieval across a large provenance-tracked evidence base
- Integrate language models through an approved managed service, and maintain a self-hosted or open-weight alternative within the same boundary
- Design prompts and output schemas
- Bind generated text to cited source records, and treat "insufficient evidence" as a valid system response rather than forcing a conclusion
- Own model packaging, serving, versioning, and rollback
Pipelines and Source Handling
- Build secure ingestion, transformation, validation, and publishing across structured, semi-structured, and unstructured sources
- Implement quality checks, schema validation, lineage capture, and audit logging
- Establish source drift detection so that degradation is surfaced rather than carried into the analysis
- Generate statistically representative synthetic data so that development can proceed ahead of live data access
Engineering Standards
- Work to the data model and standards set by the Data Lead, who approves designs and owns them through customer review
- Document assumptions, caveats, transformation logic, and known limitations, since deliverables are formally reviewed
- Instrument telemetry so that measurement does not require manual reconstruction
- Maintain the audit trail covering recommendations, human overrides, and model versions
- Submit model and pipeline changes through a gated release process rather than deploying in place
Required Qualifications
- 5+ years of experience in data engineering, data architecture, applied machine learning, ML engineering, or production analytics engineering
- Strong Python and SQL , with demonstrated experience working with large, imperfect operational data
- Experience delivering systems for sustained operational use rather than exploratory analysis alone
- Routine use of AI-assisted development, with informed judgment about where it adds value and where its output requires verification
- Ability to explain a technical decision to a stakeholder who must defend that decision without understanding its internals
- US Citizenship Required
- Active US Secret clearance. The work is performed in a controlled government cloud environment and requires a favorable investigation and CAC eligibility from the start
- Elevated personnel security requirements apply to portions of this work and are discussed during screening
- Willingness to travel up to 25% to customer sites, DEFCON AI HQ, and vendor facilities as required
Preferred Qualifications
- Clearance: active Top Secret
- Matching: direct experience applying probabilistic matching to inconsistent identity data, including names, dates, addresses, and identifiers, and familiarity with the failure modes of each. Record linkage, master data management, or identity management. Graph data modeling. PostgreSQL and pgvector or comparable. Graph algorithms applied in production
- Modeling: model calibration and threshold design. Cost-sensitive learning where error types carry unequal consequences. scikit-learn, XGBoost, PyTorch
- Retrieval and generation: retrieval-augmented generation in production. Prompt and output-schema design. Establishing that generated output remains grounded in its sources, and testing to confirm it. Self-hosted or open-weight model operation. Fine-tuning, adapters, or custom embeddings
- Pipelines: AWS Glue, Airflow, dbt, Spark, Kafka, or NiFi. Unstructured and semi-structured document ingestion. Synthetic or representative test data generation
- Environment: federal DevSecOps, RMF, ATO, or DoW cloud environments. Hardened base images. Experience advancing a pipeline from development through accreditation and deployment
- Domain: sensitive federal or defense data, and work performed under privacy or comparable handling constraints
- Responsible AI: documentation, model cards, fairness testing, and model monitoring. NIST AI RMF or comparable practice
What Success Looks Like
- A data model that the rest of the team builds on without needing to redesign it
- Matching decisions that can be explained and defended to a non-technical reviewer
- Models whose miss rate is characterized, not only their overall accuracy
- Generated explanations that assert no more than the sources support, with the citation path intact
- Pipelines that surface problems early and trace them to a specific source
- Consistent development progress, including during periods when live data is not yet available
What We Offer:
- A fully remote, results-based environment
- Competitive salary, bonus, and equity package
- 100% employer paid, comprehensive health insurance including medical, dental, and vision for you and your family
- Unlimited PTO, with your manager's approval
- Flexible work environment where you manage your work day
- 14 weeks of fully
$68.64k - $74.88k
...Job TitleIntern – Data AI/ML Engineering – Plymouth, MN – Summer 2027 Job Description Intern – Data AI/ML Engineering – Plymouth, MN – Summer 2027 Are you interested in an Internship opportunity with Philips? We welcome individuals who are currently pursuing an undergraduate...SuggestedHourly payFull timeSummer workInternshipSummer internshipWork at officeImmediate startWork visaRelocation package3 days per week- ...technical lead responsible for taking AI/ML projects from development to production in... ...)Expert Python for production-grade data and backend engineeringStrong SQL and data... ...auditability)Strong observability and reliability engineering skills (monitoring, alerting, incident...SuggestedFull timeWork at office
- Data & ML Architect - New Resources ConsultingSummary: We’re looking for a Data & ML Architect to design and deliver modern data solutions that support AI initiatives, including large language model (LLM) projects. This role will work closely with both business and technical...SuggestedRemote workFlexible hours
- ...CClient: TechMContact: Bhargav Chukka, Gulam Mohammed , Jhansi Bugatha, Vamsi SattaruCompany: SRI Tech SolutionsJob Summary (Senior ML Data Engineer) - Lead the design and maintenance of scalable data pipelines using Databricks, Spark, and Airflow. - Develop and manage...SuggestedHourly pay
$227.33k - $312.58k
We’re looking for a Staff ML Data Engineer to join Procore’s AI & Frontier Models organization. In this role, you’ll be responsible for designing and building the data systems that power frontier‑scale machine learning research and applied AI products, with a particular...SuggestedFull timeWork at officeLocal areaImmediate start3 days per week- ...Data Scientist / ML Engineer KSA Integration is a Service-Disabled Veteran-Owned Small Business (SDVOSB) that provides business and management solutions through three core capabilities: data analytics, comprehensive veterans support, and business process improvement...Full timeFor contractorsWork experience placementSummer workFlexible hours
- ...Data Scientist / ML Engineer We're looking for a Data Scientist / ML Engineer to help Themis turn data into intelligence that makes governance, risk, and compliance faster, smarter, and more reliable for our customers. In this role you'll work across the full lifecycle...Remote workFlexible hoursShift work
$65 - $67 per hour
...Position Title: ML Data Engineer Location: Bethesda, MD - 5 Days onsite - No Relocation Assignment Type: Contract to Hire Pay Rate: $65.00 - $67.00 Work Schedule: Onsite, Monday - Friday Benefits: This position is eligible for medical, dental, vision,...Contract workLocal areaRelocationMonday to Friday- ...Business consulting services. We are in search of a highly motivated candidate to join our talented Team. Job Title: Senior Data Engineer / AI ML Engineer with Python, AI/ML & LLMs. Location: Reston, VA ( Hybrid ) Job Summary:We are seeking a highly skilled and...
- ...the transformation of the eyewear and eyecare industry. Discover more by following us on LinkedIn! GENERAL FUNCTION The Data & ML Engineer is a self-sufficient engineering professional responsible for designing, building, and operating scalable, secure, and...Minimum wageFull timeLocal area
- ...WiproContact: Meghana GorusuCompany: SRI Tech SolutionsRole:Data Scientist + AI/ML EngineerLocation: Chicago, IL (Day 1 Onsite) - work in the... ...Retrieval-Augmented Generation (RAG), Agentic RAG, prompt engineering, context engineering, vector databases, embedding/chunking...Work at office
- Role Description As a Data & ML Engineer you will build the data and model layer behind an AI-enabled decision-support system operating inside an accredited environment. Your work will cover: ~Ingestion from many source systems ~Resolution of incoming records against...Full timeRemote workFlexible hours
$184.5k
...surface, ship it, and see the impact on real travelers within weeks — with decisions driven by data and AI tooling used in practice every day.As a Senior AI / ML Data Engineer, you will build the data foundations that power Layla’s AI-native travel experiences — the...Full time- ...Role Feature Engineer / ML Data Engineer Location Cleveland, OH | Pittsburgh, PA | Dallas, TX (Onsite) Fulltime Job Description Must Have Technical/Functional Skills: 10+ years in Data Engineering, Feature Engineering, or ML Engineering...Full time
- Role Description As a Data & ML Engineer you will build the data and model layer behind an AI-enabled decision-support system operating inside an accredited environment. Your work will cover: ~Ingestion from many source systems ~Resolution of incoming records against...Full timeRemote workFlexible hours
- A technology consulting firm in Georgia is seeking a Data Scientist / Machine Learning Engineer to support advanced analytics initiatives. The candidate will develop and deploy machine learning models, process large datasets, and provide insights for mission-critical operations...
$157.95k - $259.48k
...been leading the way in sports performance software, science, and data, in a world where 1% can literally mean the difference between... ...role or a pipeline maintenance job. It is the highest-leverage engineering position in Phase 1 of a platform that will define what...$135k - $170k
Career Opportunity United States Hybrid Data & Analytics ML Data Engineer ML Data Engineer contributes to Willsfly’s enterprise-scale software and product systems initiatives, working across architecture, engineering, delivery, and operational quality to support durable...Full timeH1b- Link Logistics Real Estate in New York is seeking a hybrid ML/Backend Engineer to own the intelligence layer of Link's Analytics Engine. You will design data pipelines, build knowledge graphs, and create backend services that surface insights to investment teams in real...
$125k - $140k
Job Description We are looking for a highly skilled and passionate Data Engineer with strong Python/PySpark experience and understanding of AL/ML models to implement AL/ML models in OnPrem and Cloud environment. In this role, you will be responsible for implementing AL...- Willsfly Technologies Inc is seeking an experienced ML Data Engineer to contribute to enterprise-scale software and product systems across architecture, engineering, delivery, and operational quality. The role sits within the Data & Analytics department, focused on data...Remote job
- ...Consumer Electronics in 2019, 2022, and 2023. We are looking to bring on an experienced Data Engineer to join Eight Sleep’s machine learning team. This role involves monitoring production ML systems, building tools and pipelines for data analysis, dataset curation, training...Full timeSleeping nightsFlexible hoursNight shift
$19 - $65 per hour
...huge impact and drive the future of autonomy, Plus is looking for talented individuals to join its fast-growing teams. As the ML Data Engineer intern, you will work with our ML Data Infra team to implement a peer-to-peer Distributed Data Orchestration layer to serve as...Hourly payInternship- Vizcom Technologies, Inc. is seeking an ML Data Engineer in San Francisco, an in-person, full-time role. You will build data systems turning design signals into training-ready datasets and bridge from product signals to research outcomes. You’ll collaborate with researchers...Full time
$193.93k - $291.15k
A technology company based in California is seeking a Perception ML Data Engineer to enhance autonomy performance through machine learning and data curation. The role requires over 4 years of software engineering experience, proficiency in Python, and familiarity with...- Google is seeking a Data Scientist specialized in machine learning engineering to tackle cross-functional tax-related challenges. You will collaborate with data scientists, engineers, and product managers to deploy ML pipelines and deliver end-to-end data solutions that...
- A leading autonomous technology company in Mountain View is seeking a Senior Software Engineer specializing in Perception ML Data. This role involves bridging machine learning and autonomy infrastructure, addressing complex data challenges, and developing innovative systems...
- Data Scientist / ML Engineer Year Of Experience : 7+ years Location: Overland Park KS/ Frisco TX ( 5 days onsite from day 1) Visa Type :- (US Citizen only ) (Female candidate only required ) Employment Type :- W2 Duration :- Long Term Job Description :- 7plus years of...
- ...technology firm in Los Angeles is seeking an AI/Machine Learning Engineer with 3-5 years of experience. The ideal candidate will have... ...deploying models on cloud platforms like GCP. A strong foundation in data analysis and machine learning theory is essential. This position...
- Bright Vision Technologies is seeking a Machine Learning Data Engineer to build and operate large-scale data systems powering AI training and evaluation pipelines. The role emphasizes ingestion, transformation, quality assurance, lineage, and high-throughput delivery of...Remote job
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data & ML Engineer. Be the first to apply!
- data processing engineer United States
- bi data engineer United States
- data engineer analytics United States
- principal data engineer United States
- senior cloud data engineer United States
- data center engineer United States
- staff data engineer United States
- director data engineering United States
- data engineer graduate United States
- big data devops engineer United States


