Research Engineer — Post-Training & Small Language Models (SLMs), Healthcare AI
$110.7k - $379.2kDeloitte
Position Summary Research Engineer — Post-Training & Small Language Models (SLMs), Healthcare AI Three hundred fifty million Americans rely on a healthcare system whose decision-making has become slow, costly, and adversarial — care delayed by prior authorization and paperwork, claims that misfire, clinical decisions made without the right information at the right moment, and patients who struggle to navigate or afford the care they need. Deloitte has a new AI-first effort,, backed by $1B in committed investment, building the reasoning models and agentic systems to rebuild how that system decides — across payers, providers, and life sciences, and for the patients they serve — so that care is faster, fairer, and far less wasteful. This is not AI applied at the margins. It is a ground-up rebuild of the decision-making machinery behind American healthcare, at national scale. This is resourced to do real post-training at scale — committed investment in GPU compute and training infrastructure, not toy fine-tunes. As a Research Engineer on our post-training team, you will design, train, evaluate, and align the models that reason about healthcare — working across the full post-training lifecycle to shape model behavior for clinical and operational decisioning across the industry. Healthcare decisioning is one of the cleanest verifiable-reward domains outside math and code: the problems are hard. We ground that reward in real signals — clinical policy and criteria, adjudicated outcomes, and clinical-expert judgment — so correctness is checkable rather than asserted. You will own the post-training stack for our clinical reasoning models end to end — from data and reward design through trained, evaluated models that ship. This is not a prompt-engineering role. We are looking for people who understand not just how to use LLMs, but how to improve and shape model behavior through advanced post-training. You do not need a healthcare background. We pair every engineer with clinical and domain experts and teach you the domain — you bring the modeling depth. We hire on demonstrated depth, not years — the level you join at is determined through our interview process, based on the depth and judgment you demonstrate, not your years in a title. Work you’ll do Post-training & alignment
- Design and execute post-training pipelines: supervised fine-tuning (SFT), preference optimization, and reinforcement learning / alignment workflows.
- Build and optimize training using techniques such as SFT, RLHF, PPO, DPO, GRPO, RLAIF, and Constitutional AI, and understand how each affects reasoning quality, safety, latency, cost, and reliability.
- Train reasoning models for healthcare decisioning using verifiable-reward RL — designing reward signals and verifiers grounded in clinical guidelines, policy and criteria, and adjudicated outcomes.
- Develop reward models and preference datasets to improve reasoning quality, factuality, safety, policy adherence, and task performance.
- Curate, clean, synthesize, and evaluate large-scale instruction, preference, and domain-specific datasets, with rigorous filtering, deduplication, and quality control.
- Build verification and reward pipelines from our proprietary clinical, claims, and operational data and from clinical-expert labeling — turning guidelines, policy, and adjudicated outcomes into checkable reward signals at scale.
- Implement efficient fine-tuning strategies including LoRA, QLoRA, PEFT, and adapter-based approaches; build scalable distributed training using DeepSpeed, FSDP, Megatron-LM, Ray, or equivalent.
- Optimize inference performance — latency, throughput, quantization, and deployment efficiency — for production, including frameworks such as vLLM, TensorRT-LLM, or TGI.
- Design evaluation frameworks covering reasoning, hallucination detection, factuality, instruction following, structured outputs, and domain-specific metrics.
- Build healthcare-grade evaluation — held-out clinical benchmarks, deployment regression gates, calibration and uncertainty, factuality against ground truth, and bias/fairness evaluation across patient populations and subgroups — co-designed with clinical experts.
- Apply PHI/HIPAA-aware data handling and produce model documentation suitable for regulated clinical use.
- Perform red teaming and adversarial testing to identify alignment failures, unsafe behaviors, jailbreak vulnerabilities, and regression risks; collaborate with agentic and application teams to improve tool use, grounding, and long-horizon reasoning.
- Bachelor’s degree in Computer Science, Machine Learning, Artificial Intelligence, Applied Mathematics, Computational Linguistics, or a related field.
- Demonstrated depth training and post-training large transformer-based language models in production or research — this is your craft, not coursework or a one-off fine-tune. Genuine depth including SFT and at least one preference-optimization or RL method, evidenced by shipped models, releases, or research.
- Hands-on experience with reasoning-model training and/or verifiable-reward (RLVR) workflows.
- Strong understanding of modern post-training techniques: SFT, RLHF, PPO, DPO, GRPO, RLAIF, and preference optimization workflows.
- Experience with open-weight foundation models such as Llama, Qwen, Mistral, DeepSeek, or equivalent architectures.
- Strong expertise in PyTorch and modern deep-learning tooling; experience with distributed training frameworks such as DeepSpeed, FSDP, Megatron-LM, or Ray.
- Experience implementing efficient fine-tuning techniques such as LoRA, QLoRA, PEFT, and quantization-aware workflows.
- Deep understanding of transformer architectures, tokenization, attention mechanisms, decoding strategies, and model scaling trade-offs.
- Strong grasp of LLM evaluation methodologies, benchmarking, reward modeling, and alignment trade-offs; experience with large-scale and synthetic datasets, filtering, deduplication, and quality-control pipelines.
- Strong Python engineering skills and production-grade software practices; ability to work through ambiguous, highly complex technical problems in fast-moving environments.
- Ability to travel 0–50%, on average, based on the work you do and the clients and industries/sectors you serve.
- Limited immigration sponsorship may be available.
- Experience building or optimizing reasoning models, agentic models, or tool-using LLM systems.
- Familiarity with inference optimization frameworks such as vLLM, TensorRT-LLM, TGI, or Ollama.
- Experience with multimodal models, speech models, or domain-specific foundation models; experience using large-scale GPU clusters and distributed compute.
- Contributions to open-source AI projects, research publications, benchmark development, or model releases.
- Familiarity with safety, governance, and responsible-AI practices; experience in regulated or high-stakes industries such as healthcare, finance, insurance, or public sector.\
- Applied Research Associates, Inc. has an exciting and challenging... ...a Nuclear Weapons Effects Modeling & Simulation Engineer/Scientist who will support... ...software in high level languages (e.g. C++, Python, Java,... ...gives employees the tools, training, and opportunities to take...TrainingLanguageFull timeWork experience placementWork at office3 days per week
- Position Overview Healthcare Provider AI Decision Science Consultant. Key... ...customize DS / ML / AI / Analytics models and solutions for specific... ...and governance by training users on new processes, maintaining... ...Proficiency in programming languages such as Python, R, or SQL....TrainingLanguageWork at officeLocal area
$50 - $60 per hour
DataAnnotation in North Carolina is seeking Registered Nurses to assist in training AI models for healthcare applications. The role involves evaluating AI chatbot outputs and ensuring medical accuracy. With a flexible schedule and the ability to work from home, compensation...TrainingRemote jobHourly payWork from homeFlexible hours$50 - $60 per hour
...Overview Join to apply for the Healthcare Expert role at DataAnnotation. We are looking... ...Healthcare Expert to join our team to train AI models. You will measure the progress of these... ...Therapists, Occupational Therapists, Speech-Language Pathologists, Respiratory Therapists,...LanguageHourly payFull timeContract workPart timeRemote work- ...schedule Opportunity for advancement Training & development About Us Global Language System (GLS), a division of Global... ..., Service-Disabled Veteran-Owned Small Business (SDVOSB) providing high-... ...solutions across government, healthcare, and commercial sectors. Position...TrainingLanguageHourly payCasual workWork from homeFlexible hours
- ...Analytical and Spatial Intelligence Research Team (ASIRT) at Applied... ...artificial intelligence (AI), machine learning, and computer... ...GenAI, Agentic AI Systems, Large Language Models (LLMs), Vision Language... ...Scientist - Artificial Intelligence Engineer Position Highlights: Research...LanguageFull timeWork experience placement
$40 per hour
...for medical experts to join our team to train AI models. You will measure the progress of these... ...role you will need to be an expert in healthcare. We are interested in a wide range of expertise... ...Occupational Therapists Speech-Language Pathologists Respiratory Therapists...LanguageHourly payFull timeContract workPart timeRemote work$118.3k - $219.8k
...as a final approver of models before reaching customers... ...Managing and leading small-to-medium sized teams,... ...development Requirements Large language models (prompting, fine... ...intelligenceAgentic AI architecture, multi-... ...in mentoring or training others, and acting as a...TrainingLanguageFull timeLocal area- ...Competitive salary Flexible schedule Training & development About Us: Global... ...certified Service-Disabled Veteran-Owned Small Business (SDVOSB). Our mission is to... ...deliver accurate, culturally competent language services across healthcare, government, legal, education, and...TrainingLanguageHourly payContract workFor contractorsFlexible hours
$240.5k - $270.5k
...timeposted on: Posted 2 Days... ...built to power AI-enabled precision... ...accelerate research and improve... ..., and healthcare, Verily transforms... ...insights, models, and actions... ...leaders and engineers by promoting... ...programming language (e.g., C/C++... ...or training. Your recruiter...TrainingLanguageFull timeWork experience placement- ...Analytical and Spatial Intelligence Research Team (ASIRT) at Applied... ...artificial intelligence (AI), machine learning, and computer... ...GenAI, Agentic AI Systems, Large Language Models (LLMs), Vision Language... ...Scientist – Artificial Intelligence Engineer Position Highlights Research,...LanguageWork experience placement
$104.9k - $174.7k
...the development and training of junior members and... ...of advanced AI and machine learning models. The ideal candidate... ...data scientists and engineers to design, develop,... ...and governanceLeading small- to medium-sized teamsSetting... ...directly with large language models and...TrainingLanguageFull timeLocal area$43k - $56.2k
...completion of MA school/training program or a Certified... ...exam prior to foreign language communication... ...employment at the time of posting. The pay range may be... ...personal wellness and smart healthcare decisions for you and... ...more. Our unique care model focuses on...TrainingLanguageFull timeTemporary workApprenticeship- ...talented Senior Software Engineer to support the... ...that incorporate AI and ML... ...of the following languages (Scala / Python /... ...with large language models and their applications... ...development, including training, validation, and evaluation... ...and in a small, fast-paced R&D team...TrainingLanguageFull timeRemote work
$86.4k
...integration, and use of healthcare data and... ...support of scholarly research and addressing healthcare... ..., NLP, and AI/ML algorithms. Includes... ...expertise and training. Other duties as... ..., hospital, post-acute ~3 years of... ...research purposes Language (Other than English...TrainingLanguageFor contractorsWork at officeLocal area$43k - $56.2k
...completion of MA school/training program or a Certified... ...exam prior to foreign language communication... ...employment at the time of posting. The pay range may be... ...personal wellness and smart healthcare decisions for you and... ...more. Our unique care model focuses on...TrainingLanguageFull timeTemporary workApprenticeship$50 - $60 per hour
DataAnnotation is seeking an Appellate Attorney to train AI models by evaluating their legal outputs and solving complex legal challenges. This role allows for flexible remote work, enabling you to choose projects based on your own schedule and preferences. A J.D. is mandatory...TrainingRemote jobHourly payFlexible hours- ...EZLabs, the AI innovation division... ...AI Product Engineers to join our growing... ...intelligent healthcare technology platforms... ...solutions. Research emerging AI... ...modern development languages, along with... ...large language models (LLMs), prompt... ...Corporate Laptop. ~ Training opportunities....TrainingLanguageRemote jobFull time
$40 per hour
A healthcare technology firm is seeking medical experts to evaluate AI chatbots' performance and ensure their medical accuracy. This position allows for flexible scheduling and project selection, making it suitable for both full-time and part-time professionals. Candidates...Hourly payFull timePart timeRemote workFlexible hours- ...NC border. Job Posting TitleCare Manager... ...Occasional in-person training and travel will be... ...large cities to small towns, Community Care... ...community-based healthcare delivery systems.... ...program Care Management model, including... ...diversity of cultures, language barriers, health...TrainingLanguageLocal areaRemote workWork from homeShift work
- ...Technical Skills ~ Python, Model Training/Testing Primary... ...Working on an awesome AI product for document information... ...an international top-notch engineering team with full commitment on... ...in the following programming languages: Python 3. Good English skills...TrainingLanguageH1bImmediate start
$25k
...Transitions of Care and post-hospitalization needs,... ...personal wellness and smart healthcare decisions for you and... ...more. Our unique care model focuses on personalized... ...and selection for training, including apprenticeship... .... We also provide free language interpreter services. See...TrainingLanguageFull timeTemporary workApprenticeshipWork at officeMonday to Friday$113.1k - $188.5k
...Manager Data Engineering page is... ...timeposted on: Posted Todayjob requisition... ...in applying AI and advanced... ...our legal research tools and AI... ....* Data Modeling & Storage:... ...manipulation language including optimization... ...to lead small to medium-... ..., hiring, training, appraising...TrainingLanguageFor contractorsWork at officeLocal areaImmediate startRemote workWorldwideFlexible hours$107.66k - $161.7k
...explore and build with a wide variety of AI language models (bots), including o3, o4-mini, Claude... .... About the Team and Role: Our small engineering team works on challenging problems... ...from prototyping, data pipelines and training, to realtime LLM application at scale...TrainingLanguageRemote jobFull timeWork experience placementInternship$73 per hour
...ethically shape the future of AI. What We Do The Mindrift... ...version intended to train the agent to succeed... ...Get Started Apply to this post and get the chance to contribute... ...prompts to refining model responses, you'd be... ...tasks in simple, structured language. Ability to analyze and...TrainingLanguagePart timeFreelanceRemote workFlexible hours$207k - $301k
...Software Engineer Manager, Google Meet Infrastructure... ..., natural language processing, distributed... ...for Meet. AI will change the future... ...will work with model builders (Google... ...relevant education or training. US: $207000 -... ...in the job posting. To all recruitment...TrainingLanguage- ...We're looking for a Data Engineer to develop and enhance application... ...components that support ML/AI models and data ingestion pipelines.... ...data for machine learning model training and inference Implement and... ...Bitbucket Familiarity with Natural Language Processing (NLP) techniques...TrainingLanguage
$207k - $301k
...code developed by other engineers and provide feedback to... ...intelligence, natural language processing, distributed... ...infrastructure for Meet.AI will change the future... ...future. You will work with model builders (Google... ...relevant education or training. US: $207000 - $301000...TrainingLanguage- ...is seeking a Bilingual Annotator (Japanese/English) to enhance AI chatbots. Responsibilities include evaluating and creating conversational data for AI models. Ideal candidates are fluent in both languages, detail-oriented, and possess a strong command of grammar and style...LanguageHourly payContract workRemote workFlexible hours
$174k - $252k
...one or more programming languages.3 years of experience... ...infrastructure (e.g., model deployment, model evaluation... .... Google's software engineers develop the next-generation... ...software solutions.AI will change the future... ...relevant education or training. US: $174000 - $252000...TrainingLanguage
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Engineer — Post-Training & Small Language Models (SLMs), Healthcare AI. Be the first to apply!




