Data & ML Engineer
$150k - $200kRed Cell Partners
About Us
Red Cell Partners is an incubation firm building and investing in rapidly scalable technology-led companies that are bringing revolutionary advancements to market in three distinct practice areas: healthcare, cyber, and national security. United by a shared sense of duty and deep belief in the power of innovation, Red Cell is developing powerful tools and solutions to address our Nation’s most pressing problems.
ABOUT DEFCON AI
RESILIENCE IN THE FACE OF DISRUPTION. DEFCON AI is an insights company that leverages artificial intelligence, mathematical optimization, data analytics, and software engineering for resilient optimization of complex systems.
In today’s dynamically changing world, DEFCON AI’s technology aligns outcomes with operational goals, better decision making, and empowers customers to anticipate assess, and mitigate the impacts of disruptions.
About the Role
As a Data & ML Engineer you will build the data and model layer behind an AI-enabled decision-support system operating inside an accredited environment. That work covers ingestion from many source systems, resolution of incoming records against a shared data model, relevance scoring, and generation of explanations a user can act on and defend.
Three characteristics make this a substantial technical challenge. The incoming data is predominantly low-signal, which means a model can report strong overall accuracy while failing on the cases that matter most. Every output must remain traceable to the underlying sources, because a person downstream is accountable for the result. Record matching is probabilistic rather than exact, so false matches and missed matches both carry meaningful cost.
You will not be starting from an empty repository. We operate an established platform for source custody, extraction, and retrieval, and its architect is a member of this team, so existing design decisions are documented and accessible. Your work will focus on new capability rather than maintenance: record matching, calibrated scoring, and grounded generation, hardened for the target environment. We build with current tooling and expect the same, including the use of AI assistance in our own engineering practice.
This is a fully remote role with occasional travel (up to 25%) to DEFCON AI HQ, customer sites, and vendor facilities as required.
Key Responsibilities
The technical work falls into four areas. Deep expertise in all four is not expected, so please indicate where your depth lies when you apply. The engineering standards that follow apply to everyone on the team.
Data Modeling and Record Matching
- Design and maintain the graph of entities, records, and the typed relationships between them
- Implement probabilistic matching, including blocking, candidate generation, pairwise scoring, clustering, and threshold policy
- Build deduplication and known-record suppression
- Establish provenance so that every node and edge traces to the source that asserted it
- Produce interface and data-flow design documentation detailed enough to serve as an implementation reference for other engineers
Scoring and Calibration
- Develop relevance and priority models over large, imperfect record sets
- Own calibration and threshold design, establishing what a score means rather than only how it ranks
- Design abstention policy that routes uncertain and high-risk cases to a person rather than returning a confident answer
- Perform feature engineering, establish baselines before introducing complex models, and conduct error analysis that accounts for the differing cost of false positives and false negatives
Retrieval and Generation
- Implement embeddings, vector storage, and retrieval across a large provenance-tracked evidence base
- Integrate language models through an approved managed service, and maintain a self-hosted or open-weight alternative within the same boundary
- Design prompts and output schemas
- Bind generated text to cited source records, and treat 'insufficient evidence' as a valid system response rather than forcing a conclusion
- Own model packaging, serving, versioning, and rollback
Pipelines and Source Handling
- Build secure ingestion, transformation, validation, and publishing across structured, semi-structured, and unstructured sources
- Implement quality checks, schema validation, lineage capture, and audit logging
- Establish source drift detection so that degradation is surfaced rather than carried into the analysis
- Generate statistically representative synthetic data so that development can proceed ahead of live data access
Engineering Standards
- Work to the data model and standards set by the Data Lead, who approves designs and owns them through customer review
- Document assumptions, caveats, transformation logic, and known limitations, since deliverables are formally reviewed
- Instrument telemetry so that measurement does not require manual reconstruction
- Maintain the audit trail covering recommendations, human overrides, and model versions
- Submit model and pipeline changes through a gated release process rather than deploying in place
Required Qualifications
- 5+ years of experience in data engineering, data architecture, applied machine learning, ML engineering, or production analytics engineering
- Strong Python and SQL , with demonstrated experience working with large, imperfect operational data
- Experience delivering systems for sustained operational use rather than exploratory analysis alone
- Routine use of AI-assisted development, with informed judgment about where it adds value and where its output requires verification
- Ability to explain a technical decision to a stakeholder who must defend that decision without understanding its internals
- US Citizenship Required
- Active US Secret clearance. The work is performed in a controlled government cloud environment and requires a favorable investigation and CAC eligibility from the start
- Elevated personnel security requirements apply to portions of this work and are discussed during screening
- Willingness to travel up to 25% to customer sites, DEFCON AI HQ, and vendor facilities as required
Preferred Qualifications
- Clearance: active Top Secret
- Matching: direct experience applying probabilistic matching to inconsistent identity data, including names, dates, addresses, and identifiers, and familiarity with the failure modes of each. Record linkage, master data management, or identity management. Graph data modeling. PostgreSQL and pgvector or comparable. Graph algorithms applied in production
- Modeling: model calibration and threshold design. Cost-sensitive learning where error types carry unequal consequences. scikit-learn, XGBoost, PyTorch
- Retrieval and generation: retrieval-augmented generation in production. Prompt and output-schema design. Establishing that generated output remains grounded in its sources, and testing to confirm it. Self-hosted or open-weight model operation. Fine-tuning, adapters, or custom embeddings
- Pipelines: AWS Glue, Airflow, dbt, Spark, Kafka, or NiFi. Unstructured and semi-structured document ingestion. Synthetic or representative test data generation
- Environment: federal DevSecOps, RMF, ATO, or DoW cloud environments. Hardened base images. Experience advancing a pipeline from development through accreditation and deployment
- Domain: sensitive federal or defense data, and work performed under privacy or comparable handling constraints
- Responsible AI: documentation, model cards, fairness testing, and model monitoring. NIST AI RMF or comparable practice
What Success Looks Like
- A data model that the rest of the team builds on without needing to redesign it
- Matching decisions that can be explained and defended to a non-technical reviewer
- Models whose miss rate is characterized, not only their overall accuracy
- Generated explanations that assert no more than the sources support, with the citation path intact
- Pipelines that surface problems early and trace them to a specific source
- Consistent development progress, including during periods when live data is not yet available
What We Offer:
- A fully remote, results-based environment
- Competitive salary, bonus, and equity package
- 100% employer paid, comprehensive health insurance including medical, dental, and vision for you and your family
- Unlimited PTO, with your manager’s approval
- Flexible work environment where you manage your work day
- 14 weeks of fully-paid parental leave
Salary Range: $150,000-$200,000. This represents the typical salary range for this position based on experience, skills, and other factors.
Our Red Cell Partners Benefits:
For full-time roles
- Career track opportunity with potential for rapid advancement with strong performance as the firm grows
- 100% employer paid, comprehensive health care including medical, dental, and vision for you and your family.
- Paid maternity and paternity for 14 weeks at employees' normal pay.
- Unlimited PTO, with management approval.
- Opportunities for professional development and continued learning.
- Optional 401K, FSA, and equity incentives available.
- Mental health benefits are available through Tara Mind .
- Cost effective GLP-1 solutions available through Crux .
We’re an Equal Opportunity Employer: You’ll receive consideration for employment without regard to race, sex, color, religion, sexual orientation, gender identity, national origin, protected veteran status, or on the basis of disability.
Applicant Data Disclosure
By submitting an application, you acknowledge that Red Cell Partners, LLC ('Red Cell') uses third-party service providers to facilitate its recruitment and hiring processes. These providers include applicant tracking systems, candidate verification platforms, and fraud detection tools (collectively, 'Hiring Platforms'). Your application materials, including your résumé, cover letter, work samples, responses to application questions, and any other information you submit, may be transmitted to and processed by these Hiring Platforms for the following purposes:
- Managing and administering your application throughout the hiring process;
- Verifying the accuracy and authenticity of application materials, including by cross-referencing information you provide against publicly available sources and proprietary databases;
- Identifying indicators of potentially fraudulent, fabricated, or materially misleading application content, including but not limited to discrepancies between submitted materials and publicly available professional profiles, geographic anomalies, and fabricated work histories.
Applications that are flagged through this process as containing indicators of fraud or material misrepresentation may be declined from further consideration. If you have questions about the status of your application or the evaluation process, please contact talent @redcellpartners.com .
Red Cell requires its Hiring Platform providers to process your information solely for the purposes described above and in accordance with applicable law. Your information will be retained only for as long as necessary to fulfill these purposes and any applicable legal obligations, after which it will be deleted in accordance with Red Cell's data retention policies.
For more information about how your data is used, please refer to our Privacy Policy and Applicant Privacy Notice .
$116.1k - $170.28k
...About the Role: We are seeking a highly skilled and experienced Senior Big Data Engineer to join our dynamic team. The ideal candidate will have a strong background in developing batch processing systems, with extensive experience in the Apache Hadoop ecosystem (Map...SuggestedRemote jobFull time- ...Job brief Join our San Francisco office as an ML Engineer focused on Data Engineering. Visa sponsorship available for global talent ~3 days ago A machine learning engineer with an emphasis on data engineering is needed by this organization to manage...SuggestedFull timeH1bWork at officeImmediate startVisa sponsorship
$137.8k - $206.6k
...MassachusettsShift: Department: HC-NA-IDA Analytics Engineering & BI ArchitectureRecruiter: Troy... ...the organization.Your RoleAt EMD Serono, data, analytics, and AI are core to how we... ...documentation, and governance standards for data and ML solutions.Apply data privacy, security,...SuggestedWork experience placementWork at officeLocal areaImmediate startWork from home2 days per week3 days per week- Data & ML Architect - New Resources ConsultingSummary: We’re looking for a Data & ML Architect to design and deliver modern data solutions that support AI initiatives, including large language model (LLM) projects. This role will work closely with both business and technical...SuggestedRemote workFlexible hours
$100k - $150k
...ML Data Engineer – Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-...SuggestedFull timeH1bLocal areaImmediate startRemote workVisa sponsorship$145k - $165k
...healthcare fintech company, Wellfit is investing heavily in AI, data, and modern technology to better understand our business,... ...value. About the Role: We are seeking a Senior Data & AI/ML Engineer to help Wellfit unlock the value of its data through applied AI...Full timeImmediate start$150k - $200k
Role Description As a Data & ML Engineer you will build the data and model layer behind an AI-enabled decision-support system operating inside an accredited environment. Your work will cover: ~Ingestion from many source systems ~Resolution of incoming records against...Full timeRemote workFlexible hours$150k - $200k
Role Description As a Data & ML Engineer you will build the data and model layer behind an AI-enabled decision-support system operating inside an accredited environment. Your work will cover: ~Ingestion from many source systems ~Resolution of incoming records against...Full timeRemote workFlexible hours- ...performance simulation, and modern software engineering to accelerate the design, validation, and... ...This role builds and operates the data and machine-learning infrastructure the platform... ...aren’t left waiting on it. Build the ML data pipelines for training, fine-tuning,...
- EVA ищет младшего специалиста по данным и машинному обучению для подготовки данных, разработки и поддержки ML‑моделей. Ваша работа будет связана с аналитикой данных, экспериментами и внедрением решений, работа удаленно. Указывается владение Python и SQL, базовые знания...Remote job
- Willsfly Technologies Inc is seeking an experienced ML Data Engineer to contribute to enterprise-scale software and product systems across architecture, engineering, delivery, and operational quality. The role sits within the Data & Analytics department, focused on data...Remote job
- ...Labs is seeking a Member of Technical Staff to build and scale data pipelines behind large video generation models. You will... ...to model-ready samples. Role sits at the intersection of data engineering and ML research, driving quality improvements by turning #J-18808-Ljbffr...Remote job
- A leading defense contractor is seeking a data scientist/machine learning engineer to build analytics solutions impacting critical decisions. This hybrid... ...implementing CI/CD practices, and developing production-grade ML applications. Candidates with relevant degrees and...For contractorsRemote work
- ...brand presence. Job Description What You’ll Do Data Engineering & Pipeline Development (Primary) — 60% Architect, build, and... ...scalable, and high-performing pipelines that enable downstream ML, analytics, and operational applications ML/AI Platform...Full timeWork at officeRemote workVisa sponsorshipFlexible hoursShift work
- ...industry. Job Duties Own production AI/ML systems including monitoring, performance... ..., cloud-native AI platforms Ensure data governance, security, and responsible AI... ...and Data Modeling ETL / Data Pipeline engineering AI/ML tools and frameworks Cloud AI...Full timePart timeSecond jobWork from home
$120k - $150k
...capabilities are our top-tier program and project management, data analytics, and audit services, the backbone of which is our integrated... ...results. We are looking for a dynamic Data Scientist/ML Engineer to join our team. The Data Scientist/ML Engineer will work directly...Full timeTemporary workWork experience placement- ...Job Description CapTech Machine Learning Engineers are responsible for designing and implementing data-driven solutions for our clients, with a specific focus... ...language processing, risk scoring). Productionizing ML systems with a focus on optimization and...Full timeWork at officeRemote workVisa sponsorshipWork visaFlexible hours
- ...Seeking Founding Data Scientists and Machine Learning Engineers Imagine Multiplying Your Impact You've unlocked major wins in your career - you've shipped... ..., and more. You'll help extend those domains: building ML and AI models to detect and surface product...Full time
- ...\ Duration: 6 months \ \ Work Authorization: \\ GC, USC, All valid EADs except H1B, OPT, CPT \ Must Have: \ \ ~4–6+ years of data science or machine learning experience \ ~ NLP classification for customer messages or call transcripts \ ~ Intent, topic, sentiment...Full timeH1bRemote work
- ...develop, and operationalize software that performs large-scale cyber data analytics. You will leverage your expertise to mentor team... .... The role requires a strong background in machine learning engineering, data engineering, and cybersecurity. A flexible work culture allows...Remote workFlexible hours
- Role Description The AI/ML Engineer builds, deploys, and supports production forecasting and AI capabilities on top of the company’s governed... ...This role is accountable for turning trusted enterprise data into production AI and forecasting capabilities that are reliable...Full timeTemporary work
$106k - $176.6k
We are seeking an experienced Machine Learning Data Engineer to develop, operationalize, and continuously improve production bioinformatics... ..., proteomics, functional genetics, and genetics) as inputs for ML model training or inference.Developing and evolving an omics data...Permanent employmentFull timeH1bLocal areaVisa sponsorshipWork visaRelocation package$166k - $244k
...and scalable.About the RoleWe are looking for a Machine Learning Data Engineer (DataOps) to build and unify the data infrastructure that... ..., quality control frameworks, and dataset versioning practices.ML Data Lifecycle Understanding: Hands-on experience structuring datasets...Full timeRemote work- QAVION GROUP sucht einen erfahrenen AI/ML/Data Engineer, der 7+ Jahre Berufserfahrung mitbringt und komplexe Modelle von Design bis Deployment realisiert. Sie arbeiten remote mit flexibler Zeiteinteilung, entwickeln NLP-Lösungen, LLMs und RAG-Architekturen, bauen semantische...Remote job
- ...safe and reliable infrastructure. Visit coreandmain.com to learn more. Job Summary - St. Louis, MO onsite. The AI/ML Intern will support the Data Engineering team in developing and applying machine learning and artificial intelligence solutions that enhance business...For contractorsInternshipLocal area
- ...financial platform in Palo Alto is seeking a Machine Learning Engineer to design scalable data ingestion pipelines and monitor data quality. The role... ...frameworks. Successful candidates will work closely with ML teams to enhance model performance. This position offers attractive...Remote job
- itD is seeking a Data Engineer (AI/ML) III to drive the development of machine learning solutions and scalable data processing pipelines that enhance AR display systems. The role requires expertise in ML, computer vision, image and signal processing, and large-scale data...Remote work
$500 per month
...across their business, turning real-time data into actionable intelligence that helps them... ...for an innovative Machine Learning Engineer who is excited by the challenge of transforming... ...working with a modern cloud-native data and ML platform built on technologies such as...Full timeWeekend work$50.04 - $58.84 per hour
...About the Team: As a member of the Charging Data Modeling team, you will develop models and algorithms that bridge the engineering, service, deployment, and operation of Tesla’... ...posts, papers, etc.) Experience applying ML or optimization approaches to real-world infrastructure...Full timeTemporary workPart timeInternshipWorldwideFlexible hours- ...Description THE OPPORTUNITY AKUVO is seeking a hands-on Senior Data & Machine Learning Engineer to build and own the production lifecycle of our... ...Machine Learning, Databricks, MLflow, or comparable cloud-based ML platforms; experience building reproducible training and...Local areaRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data & ML Engineer. Be the first to apply!





