ML Engineer, LLM's
Jobleads-US
Job title: ML Engineer / Researcher, Large Language Models
WORK AUTHORIZATION
Must be authorized to work in the U.S.; we are unable to sponsor visas for this role.
About Knowtex
Knowtex is building the future of voice AI operating systems for clinicians, transforming how healthcare documentation happens at the point of care. We are growing fast across both commercial health systems and federal healthcare, and our ambient documentation platform is scaling to thousands of clinicians across hundreds of specialties.
We are at an inflection point where advances in language models and clinical AI can fundamentally change how clinicians interact with technology. That gives them more time to focus on what matters most: their patients.
Position Overview
We are hiring an ML Engineer / Researcher focused on Large Language Models to help build the next generation of Knowtex's AI stack.
You will develop and optimize the models behind our clinical documentation and structured clinical reasoning, improving quality, cost, latency, and control. Your work will draw on Knowtex's large proprietary clinical dataset, built from real-world encounters across hundreds of specialties.
This is a research-heavy role with a direct path to production. You will design experiments, build datasets and evaluation systems, and fine-tune and post-train models. You will also work closely with engineering and clinical teams to deploy successful approaches at scale. The role plays a central part in defining Knowtex's long-term LLM strategy.
Key Responsibilities
Clinical Documentation & Reasoning
- Develop and optimize models that generate high-quality clinical documentation, including SOAP notes and specialty-specific note formats
- Build models for downstream clinical tasks such as medication extraction, orders, ICD-10 coding, E&M coding, patient visit summaries, and other structured clinical artifacts
- Apply structured generation, tool use, and agentic approaches to produce reliable, controllable clinical outputs
Model Strategy & Training
- Evaluate open-weight and proprietary models, and determine where fine-tuning, distillation, structured generation, or task-specific models can beat general-purpose API-based approaches
- Fine-tune and post-train open-weight LLMs on Knowtex's proprietary clinical datasets, using SFT, distillation, preference optimization, and reinforcement learning
- Research ways to reduce inference cost and latency while maintaining or improving clinical quality
- Serve and optimize open-weight models in production at scale
Evaluation
- Build rigorous evaluation frameworks for clinical accuracy, hallucinations, completeness, formatting, and clinician preferences
- Build datasets, benchmarks, and evaluation infrastructure that make model improvements measurable and reproducible
- Design experiments that clearly show whether an approach improves real‑world clinical outcomes
Research to Production
- Move quickly from idea → dataset → experiment → evaluation → production
- Take successful research beyond prototypes and help deploy models into production
- Balance model quality with latency, inference cost, reliability, and scalability
- Collaborate closely with clinicians, speech and applied ML engineers, and platform engineers
Required Qualifications
- 4+ years of experience in machine learning research or ML engineering, with deep expertise in large language models
- Hands‑on experience fine-tuning or post‑training open‑weight LLMs (e.g., SFT, distillation, preference optimization, RL)
- Strong Python and PyTorch skills
- Deep understanding of modern transformer architectures and LLM training techniques
- Strong experimental methodology and the ability to independently design and execute research projects
- Experience working with large‑scale datasets and distributed training environments
- Strong understanding of LLM evaluation and benchmarking
- Ability to translate research results into production systems
- Bachelor's, Master's, or PhD in Computer Science, Machine Learning, or a related technical field, or equivalent research experience
Preferred Qualifications
- Experience building LLM evaluation systems, including LLM‑as‑judge, human preference, and task‑specific benchmarks
- Experience serving and optimizing open‑weight models at scale (e.g., vLLM, TensorRT‑LLM, quantization, speculative decoding)
- Experience with structured generation, tool use, or agentic systems
- Experience in healthcare AI or clinical NLP
- Familiarity with clinical documentation workflows and medical terminology
- Knowledge of coding systems such as ICD-10, CPT, E&M, or SNOMED
- Publications at leading ML or NLP conferences
- Experience deploying ML systems in HIPAA‑compliant or regulated environments
- Experience in fast‑moving startups where researchers own projects from experimentation through production
Technical Environment
- AWS
- Python, PyTorch
- Open‑weight and frontier language models
- Large‑scale clinical text and transcript datasets
- Distributed model training and inference
- GPU‑based model serving and optimization
- Real‑time clinical AI pipelines
- Structured clinical evaluation and benchmarking infrastructure
Compensation & Benefits
- Competitive salary
- Meaningful equity compensation
- Unlimited PTO
- Premium health, dental, and vision coverage
- 401(k) plan
- Work model: Hybrid, in person (M-W in our SF office)
- ...What you’ll do: Lead and grow a team of experienced ML, backend, and data engineers responsible for building innovative ads measurement products... ...Champion the effective use of AI-assisted development and LLM-powered tools to accelerate engineering, prototyping, experimentation...Suggested
- ...ML Engineer San Francisco, California, United States Or refer someone Job Openings ML Engineer About the Job Our client is a rapidly... ...of the AI stack, improve tooling, and drive innovation in LLM and audio ML applications. Work directly with customers to identify...SuggestedFull time
- ...in San Francisco (Mission Bay) is seeking an experienced backend engineer to design and scale AI-powered services. You will build APIs... ...implement CI/CD pipelines. The role emphasizes hands-on work with LLM-based applications, multi-step workflows, and prompt engineering...Suggested
$200k - $280k
...Engineering San Francisco Full-time $200,000 - $280,000 About the Role Join our ML Infrastructure team to build the systems that train, deploy, and serve our AI models at... ...Nice to Have ~ Experience with LLM fine-tuning and deployment ~ Background in...SuggestedFull timeWork at office$300k
...term relationship. We're doubling down on ML as the future of Grindr, and in these... ...leadership across teams, collaborating with engineering, data science and product teams to turn bold... ...Expertise in building and maintaining LLM workflows for nuanced, human-driven tasks...SuggestedCasual workWork at officeImmediate startWorldwideFlexible hours$189.6k - $237k
Scale’s ML platform (RLXF) team builds our internal distributed framework for large language... ...and automatic training and evaluation of LLM's, as well as evaluation of data quality.... ...distributed ML systemsStrong software engineering skills, proficient in frameworks and...Full time- ...WLEN WorldJobs LLP in San Francisco is seeking an experienced ML Engineer to develop multimodal AI solutions at enterprise scale. You will collaborate with product teams to deploy LLM-powered applications, integrate AI services, and build robust MLOps pipelines across...
- .... About the Role We are looking for a visionary Senior ML Engineer who will bridge the gap between high-level architecture and hands... ...phase) ~2+ years of ML, specifically training or fine-tuning LLM models, embeddings; building clustering models; utilizing...Shift work
- Title: ML Engineer Location: San Francisco, CA (Onsite) Direct HireCompany Mission Our client’s mission is to scale medical knowledge and create reliable AI for the health benefit of every person. The company is developing a new class of neurosymbolic AI systems, combining...
$145.5k - $249.5k
...Optum AI is UnitedHealth Group's enterprise AI team. We are AI/ML scientists and engineers with deep expertise in AI/ML engineering for health care.... ...of our current projects involve cutting edge ML, NLP and LLM techniques. Generative AI methods for working with structured...Minimum wageFull timeWork experience placementWork at officeLocal areaRemote work$195k - $255k
...eliminates administrative waste.The RoleAs an AI Engineer at Candid Health, you’ll be empowered to... ...set of differentiated, high impact ML/AI use cases throughout our platform and across... ...deterministic (ML) vs. non-deterministic (LLM) approaches to deliver the best ROI,...$210k - $240k
...You'll build the ML behind Firecrawl — the models and the systems that serve them. That... ...extending that work across extraction quality and LLM-driven features. You'll also own how we... ...production — whether your title says ML engineer or data scientist — this is for you....Full timeTemporary workFor contractorsRemote workVisa sponsorshipFlexible hours$230k
...authorization, governance, and trust become the real engineering challenge. Arcade is the MCP runtime... ...database team at Redis, shipped 100+ LLM applications, and is a contributor to... .... You'll own it, decide what the ML stack at Arcade looks like, and set the patterns...Work at officeShift work- ...Operational Design Domain. You will build the ML systems that carry us from L1 to L3 and... ...will be Shepherd's first Machine Learning Engineer, embedded in the Fully Autonomous... ...Experience with agentic frameworks or multi-step LLM orchestration (LangChain, LangGraph, or custom...Work at office
- ...something that matters to the world. Role Scope Build ML and LLM systems that run inside the company's operations: forecasting... ...company systems instead of just advising. Partner with data engineering and product pods to put predictions in the tools people...
- ...our Glassdoor page! Machine Learning Engineer @ Clay Clay's ambition is to build a... ...every time someone uses it. This means data, ML, and AI are at the heart of everything we... ...Experience designing eval frameworks for LLM or ML systems Familiarity with modern...
$118k - $176k
...Total Visits, March 2025) Day to Day The Machine Learning Engineer I role partners closely with business partners across various... ...business problems across Indeed. Work spans classical ML through LLM systems. You improve search and retrieval quality using real...Work experience placementLocal area- ...Machine Learning Engineer Location: Onsite in San Francisco Compensation: Competitive Salary + Equity UniversalAGI is building... ...to hear from you. About the Role UniversalAGI is hiring an ML Engineer to help ship ML outcomes by owning the execution layer:...Work at officeFlexible hours1 day per week
$150k - $300k
...Founding ML Engineer Location: San Francisco, CA Company Stage: Early-Stage (YC-backed, Profitable, High-Growth) Office Type: Onsite Salary... ...distributed training on GPU clusters Experience scaling LLM inference pipelines in production Research publications or open...Visa sponsorship- ...environments. What to expect This role is for a machine learning engineer who wants to work on the models that give LeLamp its... ...for training and deploying machine learning models Integrate ML systems with robotic hardware and embedded systems Improve robot...Immediate start
- ...with training, fine tuning, and evaluating ML models used in production systems... ...Experimenting with novel techniques to improve LLM accuracy Build data pipelines, evaluate... ...to shape the product direction and engineering strategy Bonus points if you:...Work at officeLocal area
- ...We're hiring a Founding Machine Learning Engineer to build the AI systems that power our platform... .... What matters is that you're a strong ML engineer who ships. This role is in‑... ...What You'll Work On Build and improve the LLM‑powered pipelines that turn natural...Work at office
- ...team is looking for a highly technical Machine Learning Engineer to work at the intersection of LLM systems, AI agents, inference and performance... ...have trained foundation models or come from a traditional ML research background. We're particularly interested in strong...
- ...Machine Learning Engineer Bucket Robotics is hiring a Machine Learning Engineer to push the frontier of CAD-native computer vision for manufacturing. You'll work on the core ML systems that turn 3D geometry and synthetic data into reliable, production-grade vision...Shift work
$200k - $240k
...Machine Learning Engineer - Stealth Startup, San Francisco Employment type: Full-time... ...you'll build, ship, and run the models and ML services behind our marketplace. You'll own... ...serving, and monitoring Build applied AI and LLM-powered product features where they add...Full time$163.42k - $285.98k
...billion ideas saved, Pinterest Machine Learning engineers build personalized experiences to help... ...anywhere else.Within the Monetization ML Engineering team, we try to connect the dots... ...testing, and refactoringFamiliarity with LLM-powered productivity tools for documentation...Local areaRelocation package- ...is looking for a Senior Machine Learning Engineer to join our Search & Intelligence organization... ...agents. You will own projects across the ML lifecycle, from problem definition and... ...regression testing, human evaluation, or LLM-as-a-judge.Experience evaluating agent planning...Work at officeLocal area
- ...Responsibilities Senior Machine Learning Engineer — Agentic Search & Query... ...What you’ll doDevelop machine learning and LLM capabilities for query intelligence, including... ...that strengthen the team’s engineering and ML practices.Your backgroundOn the first day,...Work at officeLocal area
$138.91k - $285.98k
...Our initiatives span a diverse range of AI/ML fields, including fundamental computer... ...collaborative group of approximately six engineers and a product prototyping team, to create... ...experience working with finetuning open-source LLM models and improve their visual perception...Work at officeLocal areaRemote workRelocationRelocation package$131.4k - $236k
...we enter our next phase of growth, we’re investing deeply in AI/ML, LLMs, and Industrial IoT to transform how frontline teams operate... ...technical direction for predictive maintenance, anomaly detection, and LLM-powered intelligence across MaintainX products.Architect end-to-...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to ML Engineer, LLM's. Be the first to apply!
- computer vision machine learning engineer San Francisco, CA
- junior machine learning research engineer San Francisco, CA
- machine learning software engineer San Francisco, CA
- ai ml engineer San Francisco, CA
- senior ml engineer San Francisco, CA
- machine learning ai engineer San Francisco, CA
- machine learning engineer San Francisco, CA
- data engineer machine learning San Francisco, CA
- artificial intelligence - machine learning intern San Francisco, CA
- machine learning intern San Francisco, CA



