Data & ML Engineer
$150k - $200kFull-time
DEFCON AI
ABOUT DEFCON AI
RESILIENCE IN THE FACE OF DISRUPTION. DEFCON AI is an insights company that leverages artificial intelligence, mathematical optimization, data analytics, and software engineering for resilient optimization of complex systems. In today’s dynamically changing world, DEFCON AI’s technology aligns outcomes with operational goals, better decision making, and empowers customers to anticipate assess, and mitigate the impacts of disruptions. About the Role As a Data & ML Engineer you will build the data and model layer behind an AI-enabled decision-support system operating inside an accredited environment. That work covers ingestion from many source systems, resolution of incoming records against a shared data model, relevance scoring, and generation of explanations a user can act on and defend. Three characteristics make this a substantial technical challenge. The incoming data is predominantly low-signal, which means a model can report strong overall accuracy while failing on the cases that matter most. Every output must remain traceable to the underlying sources, because a person downstream is accountable for the result. Record matching is probabilistic rather than exact, so false matches and missed matches both carry meaningful cost. You will not be starting from an empty repository. We operate an established platform for source custody, extraction, and retrieval, and its architect is a member of this team, so existing design decisions are documented and accessible. Your work will focus on new capability rather than maintenance: record matching, calibrated scoring, and grounded generation, hardened for the target environment. We build with current tooling and expect the same, including the use of AI assistance in our own engineering practice. This is a fully remote role with occasional travel (up to 25%) to DEFCON AI HQ, customer sites, and vendor facilities as required. Key Responsibilities The technical work falls into four areas. Deep expertise in all four is not expected, so please indicate where your depth lies when you apply. The engineering standards that follow apply to everyone on the team. Data Modeling and Record Matching Design and maintain the graph of entities, records, and the typed relationships between them Implement probabilistic matching, including blocking, candidate generation, pairwise scoring, clustering, and threshold policy Build deduplication and known-record suppression Establish provenance so that every node and edge traces to the source that asserted it Produce interface and data-flow design documentation detailed enough to serve as an implementation reference for other engineers Scoring and Calibration Develop relevance and priority models over large, imperfect record sets Own calibration and threshold design, establishing what a score means rather than only how it ranks Design abstention policy that routes uncertain and high-risk cases to a person rather than returning a confident answer Perform feature engineering, establish baselines before introducing complex models, and conduct error analysis that accounts for the differing cost of false positives and false negatives Retrieval and Generation Implement embeddings, vector storage, and retrieval across a large provenance-tracked evidence base Integrate language models through an approved managed service, and maintain a self-hosted or open-weight alternative within the same boundary Design prompts and output schemas Bind generated text to cited source records, and treat "insufficient evidence" as a valid system response rather than forcing a conclusion Own model packaging, serving, versioning, and rollback Pipelines and Source Handling Build secure ingestion, transformation, validation, and publishing across structured, semi-structured, and unstructured sources Implement quality checks, schema validation, lineage capture, and audit logging Establish source drift detection so that degradation is surfaced rather than carried into the analysis Generate statistically representative synthetic data so that development can proceed ahead of live data access Engineering Standards Work to the data model and standards set by the Data Lead, who approves designs and owns them through customer review Document assumptions, caveats, transformation logic, and known limitations, since deliverables are formally reviewed Instrument telemetry so that measurement does not require manual reconstruction Maintain the audit trail covering recommendations, human overrides, and model versions Submit model and pipeline changes through a gated release process rather than deploying in place Required Qualifications 5+ years of experience in data engineering, data architecture, applied machine learning, ML engineering, or production analytics engineering Strong Python and SQL, with demonstrated experience working with large, imperfect operational data Experience delivering systems for sustained operational use rather than exploratory analysis alone Routine use of AI-assisted development, with informed judgment about where it adds value and where its output requires verification Ability to explain a technical decision to a stakeholder who must defend that decision without understanding its internals US Citizenship Required Active US Secret clearance. The work is performed in a controlled government cloud environment and requires a favorable investigation and CAC eligibility from the start Elevated personnel security requirements apply to portions of this work and are discussed during screening Willingness to travel up to 25% to customer sites, DEFCON AI HQ, and vendor facilities as required Preferred Qualifications Clearance: active Top Secret Matching: direct experience applying probabilistic matching to inconsistent identity data, including names, dates, addresses, and identifiers, and familiarity with the failure modes of each. Record linkage, master data management, or identity management. Graph data modeling. PostgreSQL and pgvector or comparable. Graph algorithms applied in production Modeling: model calibration and threshold design. Cost-sensitive learning where error types carry unequal consequences. scikit-learn, XGBoost, PyTorch Retrieval and generation: retrieval-augmented generation in production. Prompt and output-schema design. Establishing that generated output remains grounded in its sources, and testing to confirm it. Self-hosted or open-weight model operation. Fine-tuning, adapters, or custom embeddings Pipelines: AWS Glue, Airflow, dbt, Spark, Kafka, or NiFi. Unstructured and semi-structured document ingestion. Synthetic or representative test data generation Environment: federal DevSecOps, RMF, ATO, or DoW cloud environments. Hardened base images. Experience advancing a pipeline from development through accreditation and deployment Domain: sensitive federal or defense data, and work performed under privacy or comparable handling constraints Responsible AI: documentation, model cards, fairness testing, and model monitoring. NIST AI RMF or comparable practice What Success Looks Like A data model that the rest of the team builds on without needing to redesign it Matching decisions that can be explained and defended to a non-technical reviewer Models whose miss rate is characterized, not only their overall accuracy Generated explanations that assert no more than the sources support, with the citation path intact Pipelines that surface problems early and trace them to a specific source Consistent development progress, including during periods when live data is not yet available What We Offer: A fully remote, results-based environment Competitive salary, bonus, and equity package 100% employer paid, comprehensive health insurance including medical, dental, and vision for you and your family Unlimited PTO, with your manager’s approval Flexible work environment where you manage your work day 14 weeks of fully-paid parental leave Salary Range: $150,000-$200,000. This represents the typical salary range for this position based on experience, skills, and other factors. We’re an Equal Opportunity Employer: You’ll receive consideration for employment without regard to race, sex, color, religion, sexual orientation, gender identity, national origin, protected veteran status, or on the basis of disability. Applicant Data Disclosure By submitting an application, you acknowledge that Defcon AI uses third-party service providers to facilitate its recruitment and hiring processes. These providers include applicant tracking systems, candidate verification platforms, and fraud detection tools (collectively, "Hiring Platforms"). Your application materials, including your résumé, cover letter, work samples, responses to application questions, and any other information you submit, may be transmitted to and processed by these Hiring Platforms for the following purposes: Managing and administering your application throughout the hiring process; Verifying the accuracy and authenticity of application materials, including by cross-referencing information you provide against publicly available sources and proprietary databases; Identifying indicators of potentially fraudulent, fabricated, or materially misleading application content, including but not limited to discrepancies between submitted materials and publicly available professional profiles, geographic anomalies, and fabricated work histories. Applications that are flagged through this process as containing indicators of fraud or material misrepresentation may be declined from further consideration. If you have questions about the status of your application or the evaluation process, please contact View email address on click.appcast.io. Defcon AI requires its Hiring Platform providers to process your information solely for the purposes described above and in accordance with applicable law. Your information will be retained only for as long as necessary to fulfill these purposes and any applicable legal obligations, after which it will be deleted in accordance with Defcon AI's data retention policies. For more information about how your data is used, please refer to our Privacy Policy and Applicant Privacy Notice.Vacancy posted 14 hours ago
Similar jobs that could be interesting for youBased on the Data & ML Engineer in United States vacancy
- ...technical lead responsible for taking AI/ML projects from development to production in... ...)Expert Python for production-grade data and backend engineeringStrong SQL and data... ...auditability)Strong observability and reliability engineering skills (monitoring, alerting, incident...SuggestedFull timeWork at office
$137.8k - $206.6k
...MassachusettsShift: Department: HC-NA-IDA Analytics Engineering & BI ArchitectureRecruiter: Troy... ...the organization.Your RoleAt EMD Serono, data, analytics, and AI are core to how we... ...documentation, and governance standards for data and ML solutions.Apply data privacy, security,...SuggestedWork experience placementWork at officeLocal areaImmediate startWork from home2 days per week3 days per week- ...involved. If you want to make an impact on a global scale, come make a difference at Fiserv.Job TitleSenior Data & ML EngineerAbout Your role:As a Senior Data & ML Engineer, you will play a lead technical role in building and operationalizing the data engineering, ETL, and...SuggestedFull timeContract workFor contractorsWork at officeLocal areaMonday to Friday
- ...CClient: TechMContact: Bhargav Chukka, Gulam Mohammed , Jhansi Bugatha, Vamsi SattaruCompany: SRI Tech SolutionsJob Summary (Senior ML Data Engineer) - Lead the design and maintenance of scalable data pipelines using Databricks, Spark, and Airflow. - Develop and manage...SuggestedHourly pay
- ...and management solutions through three core capabilities: (1) data analytics, (2) comprehensive veterans support, and (3) business... ...Position Overview: KSA Integration is seeking a Data Scientist / ML Engineer to support the Marine Corps Logistics Plans (LP) Division’s...SuggestedFull timeFor contractorsWork experience placementSummer workFlexible hours
$116.1k - $170.28k
...About the Role: We are seeking a highly skilled and experienced Senior Big Data Engineer to join our dynamic team. The ideal candidate will have a strong background in developing batch processing systems, with extensive experience in the Apache Hadoop ecosystem (Map...Remote jobFull time- ...Job brief Join our San Francisco office as an ML Engineer focused on Data Engineering. Visa sponsorship available for global talent ~3 days ago A machine learning engineer with an emphasis on data engineering is needed by this organization to manage...Full timeH1bWork at officeImmediate startVisa sponsorship
$122.13k - $183.2k
JOB DESCRIPTIONJob Description: Data Infrastructure & ML Engineer (Hybrid Role)Role SummaryWe are seeking a Senior Data Infrastructure & Machine Learning Engineer to design and implement scalable data systems and pipelines that support advanced analytics and machine learning...Full timeLocal area$150k - $200k
Role Description As a Data & ML Engineer you will build the data and model layer behind an AI-enabled decision-support system operating inside an accredited environment. Your work will cover: ~Ingestion from many source systems ~Resolution of incoming records against...Full timeRemote workFlexible hours$100k - $150k
Role Description We are seeking an ML Data Engineer to build and operate the large-scale data systems that power modern AI training and evaluation pipelines. The role combines deep data engineering expertise with a strong understanding of AI workloads, focusing on: ~...Full timeLocal areaImmediate start$150k - $200k
Role Description As a Data & ML Engineer you will build the data and model layer behind an AI-enabled decision-support system operating inside an accredited environment. Your work will cover: ~Ingestion from many source systems ~Resolution of incoming records against...Full timeRemote workFlexible hours- ...the transformation of the eyewear and eyecare industry. Discover more by following us on LinkedIn! GENERAL FUNCTION The Data & ML Engineer is a self-sufficient engineering professional responsible for designing, building, and operating scalable, secure, and...Minimum wageFull timeLocal area
$65 - $67 per hour
...Position Title: ML Data Engineer Location: Bethesda, MD - 5 Days onsite - No Relocation Assignment Type: Contract to Hire Pay Rate: $65.00 - $67.00 Work Schedule: Onsite, Monday - Friday Benefits: This position is eligible for medical, dental, vision...Contract workLocal areaRelocationMonday to Friday- ...Ecommerce Retail Business based our of Fort Lauderdale, who are looking to hire a critical technical role for the business. This Data & ML Engineer would be responsible for designing, building, and scaling the companies next-generation data platform and ML ecosystem. You...Work at office
- ...Company Culture Responsibilities Data Platform, Pipelines, & Quality (Primary... ...serve analytics foundations. Applied ML Ownership - Smart Exit Cart-Empty... ...machine learning pipelines from feature engineering to deployment and monitoring. Build and...Contract workLocal areaFlexible hours
- ## Data Science / ML Data EngineerUnited States · Full-time · Senior#### About The Position**Department:** Sales and Delivery Team - Empower... ...**: 6+ years**Basic Qualification:** Master/Bachelor of Engineering or Equivalent**Travel Requirements:**Not required**Website**...Full time
- ...Data Scientist Exusia, a cutting-edge digital transformation company, seeks a Data Scientist to join our global delivery team's Data Engineering & Analytics practice in the United States. The Data Scientist role will be part of our Data Engineering & Analytics practice...
- ...Data Scientist / ML Engineer We're looking for a Data Scientist / ML Engineer to help Themis turn data into intelligence that makes governance, risk, and compliance faster, smarter, and more reliable for our customers. In this role you'll work across the full lifecycle...Remote workFlexible hoursShift work
- ...Joining the Data Lab team, the full-time Data Scientist / ML Engineer will design and develop operational AI and machine learning solutions, working remotely while collaborating on innovative projects that include data extraction, modeling, and deployment. Key responsibilities...Full timeRemote work
- ...Business consulting services. We are in search of a highly motivated candidate to join our talented Team. Job Title: Senior Data Engineer / AI ML Engineer with Python, AI/ML & LLMs. Location: Reston, VA ( Hybrid ) Job Summary:We are seeking a highly skilled and...
- ...WiproContact: Meghana GorusuCompany: SRI Tech SolutionsRole:Data Scientist + AI/ML EngineerLocation: Chicago, IL (Day 1 Onsite) - work in the... ...Retrieval-Augmented Generation (RAG), Agentic RAG, prompt engineering, context engineering, vector databases, embedding/chunking...Work at office
- Data Scientist / ML Engineer Year Of Experience : 7+ years Location: Overland Park KS/ Frisco TX ( 5 days onsite from day 1) Visa Type :- (US Citizen only ) (Female candidate only required ) Employment Type :- W2 Duration :- Long Term Job Description :- 7plus years of...
- ...Consumer Electronics in 2019, 2022, and 2023. We are looking to bring on an experienced Data Engineer to join Eight Sleep’s machine learning team. This role involves monitoring production ML systems, building tools and pipelines for data analysis, dataset curation, training...Full timeSleeping nightsFlexible hoursNight shift
$145k - $165k
...healthcare fintech company, Wellfit is investing heavily in AI, data, and modern technology to better understand our business,... ...value. About the Role: We are seeking a Senior Data & AI/ML Engineer to help Wellfit unlock the value of its data through applied AI...Full timeImmediate start$135k - $170k
Career Opportunity United States Hybrid Data & Analytics ML Data Engineer ML Data Engineer contributes to Willsfly’s enterprise-scale software and product systems initiatives, working across architecture, engineering, delivery, and operational quality to support durable...Full timeH1b- Role Description The AI/ML Engineer builds, deploys, and supports production forecasting and AI capabilities on top of the company’s governed... ...This role is accountable for turning trusted enterprise data into production AI and forecasting capabilities that are reliable...Full timeTemporary work
- ...xperience Required: Minimum 6 to 8 years of experience as data engineer in AI ML. Skills Needed: Snowflake and Python/Scala/Java SQL, No SQL database, Hadoop, Spark ETL Tools, Informatica Tableau AI & ML Qualifications:...
- ...DeNOVO Solutions is seeking a Mid-Level Data Scientist Engineer in Aurora, CO. This role involves leading AI/ML solution design and development for complex mission environments, requiring 5-10 years of relevant experience and a bachelor's degree in a quantitative field...
- ...Job Title : AI ML Data Engineer Location : Minnetonka, MN (ONSITE) FULLTIME ONLY Job Description Must Have Technical/Functional Skill • We are seeking a senior, hands on AI/ML Engineer contractor with expert level Python skills and...Full timeFor contractors
- ...create a brighter future and make a meaningful difference.As a Lead Data Engineer at JPMorganChase within the Commercial & Investment Bank, you... ...fault toleranceCollaborate closely with data scientists and ML engineers to ensure training datasets for different models, and...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data & ML Engineer. Be the first to apply!
Related searches
- software data engineer United States
- entry level big data engineer United States
- big data developer United States
- senior data quality engineer United States
- sr data engineer United States
- junior big data engineer United States
- big data cloud engineer United States
- junior data engineer United States
- data platform engineer United States
- associate data engineer United States






