Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Project Lead, AI Model Training

$150k - $240k
Full-time

NewtonX

About NewtonX

NewtonX is a B2B insights company trusted by the world's most innovative companies to make high-stakes decisions with confidence. We combine a verified network of business professionals with AI-powered research tools to deliver research intelligence faster, more precise, and more defensible than traditional methods.

Our clients include Google, Microsoft, TikTok, DoorDash, Stripe, and Coinbase. Our research has been cited by Fortune, Forbes, TechCrunch, Adweek, and the Wall Street Journal.

NewtonX has raised $47M from investors including Two Sigma Ventures, Third Prime, XFund, and Citi Ventures.

About the Role

NewtonX is turning the world’s largest verified expert network into the evaluation, post-training, and audit data that enterprises and AI labs need to deploy AI in production. We build high-quality, expert-generated datasets that train and evaluate frontier AI systems — spanning domain eval sprints, custom data solutions, model behavior audits, agentic RL environments, and domain QA-as-a-service.

This business is taking off, and we are expanding our Delivery team to meet it. We are hiring a Project Lead to own client engagements end to end, from scoping through delivery, on our data operations platform. You will be the client’s point of contact and the person accountable for the operational execution and the quality outcome across a portfolio of concurrent projects — many of them complex agentic-workflow evaluations.

Our commercial team wins the business. From the moment it is won, it is yours: you scope it, configure it, run it, defend its quality, and deliver it. When a client asks how it is going, you are the person who answers. When the spec turns out to be wrong at scale, you are the person who fixes it.

You will be joining a small, early Delivery team, working closely with our ML Lead, annotator delivery team, commercial team, and product and engineering. Because the team is still young, you will help build and scale the delivery function itself — the playbooks, quality standards, and processes that everyone who follows you will run on — and you will have direct exposure to NewtonX’s C-level. This is a ground-floor seat in a business we expect to grow extremely quickly, with the career path that comes with it.

Location: Remote (US or Canada) or Hybrid (New York City)

In this role you will

1. Scope and set up every engagement

  • Translate the client’s need into a bounded statement of work. Turn a training or evaluation objective into concrete task design, a defined quality bar (defect-rate and inter-annotator agreement targets), an annotator profile, and — for agentic projects — an environment and tool strategy.

  • Configure the project on the platform. Stand up the multi-layer quality pipeline, task forms, grading rubrics, annotator enablement, and routing.

  • Author the first-draft guidelines and gold set with the ML Lead, and pressure-test them before production starts. Most delivery failures are scoping failures; your job is to catch them here.

2. Run production

  • Own quality, throughput, and cost. Track how each project is performing against its targets, diagnose what is drifting, and act before it reaches the client.

  • Manage the annotator pool on your projects. Assignment, layer access, qualification, and performance — working with the annotator delivery team to get the right experts on the right tasks.

  • Keep annotators aligned. Run project channels and calibration sessions, handle exceptions and escalations, and adjust configuration as the work reveals what the spec missed.

3. Own the client relationship

  • Be the client’s point of contact throughout — scoping calls, mid-project updates, quality escalations, and delivery.

  • Manage expectations against scope, and diagnose acceptance issues candidly when they arise. Clients trust the person who tells them early.

  • Grow the account from the inside. Surface expansion opportunities and new use cases to the commercial team, and act as the technical subject matter expert on calls when new work is being scoped.

4. Deliver and close out

  • Assemble and deliver batches on schedule, and own the handoff to the client.

  • Run client acceptance testing against agreed criteria, capture sign-off, and trigger the commercial milestone.

  • Run the post-mortem and feed what you learned back into guidelines, tooling, and the next scope.

5. Build the machine as you run it

  • Turn each engagement into repeatable process. Codify what works into playbooks, templates, and quality standards that the next Project Lead can run from day one.

  • Be the voice of delivery to product and engineering. You will be running on a platform that is still being built; what you flag gets built.

Who you are

  • 3+ years of experience in data labeling, RLHF/SFT, or AI evaluation delivery — you have shipped real annotation or evaluation projects at a data labeling company (Scale, Surge, Mercor, Snorkel, Handshake, Centific, Labelbox, SuperAnnotate, or similar) or inside an AI lab. This is a firm requirement; we are not looking for a generalist project manager.

  • An operator who goes into the weeds. You are comfortable owning quality metrics, debugging a stalled queue, and rewriting a rubric that is not holding up in production.

  • Client-facing and credible. You can run a scoping call with a technical buyer, deliver bad news without losing the room, and translate operational mechanics into terms a client cares about.

  • Analytically rigorous. Data analysis skills using SQL or Python are a strong plus.

  • A builder. Processes here are young and some are still manual. You want to stand things up rather than inherit a finished machine.

  • Exceptional at holding many threads. You will run several concurrent projects with different clients, quality bars, and annotator pools.

  • Adaptable and resilient. Specs change, clients change their minds, and no two weeks look the same. You roll with it and find the new angle.

  • Discreet. This role handles annotator personal data and client-confidential information under strict data-security and confidentiality standards; all work stays within approved systems.

  • Valid US or Canadian work authorization required. Visa sponsorship is not available for this role.

What we offer

  • Competitive compensation: $150-240K total on target compensation + participation in our company stock buying program

  • Massive impact: the chance to help build a new business unit from the ground up with direct C-level influence at an extremely fast-growing late-stage startup

  • Fast-track career growth toward operational and commercial leadership

  • Excellent medical, dental, and vision insurance

  • 401k match with immediate vesting

  • Health savings/flexible savings account, and pre-tax commuter benefits

  • Paid time off: vacation, holidays, sick, and parental leave

  • A diverse, collaborative, and positive culture where we invest in and celebrate each other’s success (happy hours, team projects, and retreats)

NewtonX is proud to be an equal opportunity workplace. We do not discriminate based upon race, religion, color, national origin, sex, sexual orientation, gender identity/expression, age, status as a protected veteran, status as an individual with a disability, or any other applicable legally protected characteristics.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Project Lead, AI Model Training in Remote vacancy
  • $150k - $240k

     ...position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Project Lead, AI Model Training based in United States. This is a high-impact opportunity to lead complex AI training and evaluation engagements... 
    Training
    Remote job
    Full time
    Immediate start
    Flexible hours

    jobgether

    United States
    14 hours ago
  • $164.78k - $314.96k

     ...business needs.The OpportunityAs the AI Model Governance & Monitoring Lead, you will lead the development and...  ....Proven experience leading projects or programs in which you applied data...  ...support (i.e., H-1B, TN, STEM OPT Training Plans, etc.).Compensation: USAA has... 
    Training
    Full time
    H1b
    Work at office
    Remote work
    Home office
    Relocation package
    Flexible hours

    USAA - United Services Automobile Association

    San Antonio, TX
    6 hours ago
  • Who are we?Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems.We’re training and deploying frontier models for enterprises who are building... 
    Training
    Work at office
    Local area
    Remote work
    Home office

    Cohere

    New York, NY
    6 hours ago
  • $165.2k - $223.6k

     ...unparalleled ML inference and training performance.The...  ...running a wide range of models and supporting novel architecture...  ...of what's possible in AI acceleration.As part of...  ...role will help lead the efforts in building...  ...and strive to assign projects that help our team members... 
    Training
    Work experience placement
    Internship
    Local area
    Flexible hours

    Amazon

    Cupertino, CA
    1 day ago
  • $182k - $242k

     ...CoreWeave, the AI Hyperscaler™, acquired Weights & Biases to create...  ...together CoreWeave’s industry-leading cloud infrastructure with the...  ...standard for how AI is built, trained, and scaled. The...  ...From experiment tracking and model optimization to high-performance... 
    Training
    Permanent employment
    Full time
    Temporary work
    Casual work
    Work at office
    Remote work
    Flexible hours

    Weights & Biases

    New York, NY
    2 days ago
  • $40 - $80 per hour

     ...Apply military writing and reporting expertise to help train next-generation AI systems using realistic, standards-aligned federal and military communications. This remote contractor role focuses on creating, reviewing, and evaluating materials that reflect authentic... 
    Training
    Hourly pay
    For contractors
    Remote work

    SaidGig

    Remote
    17 days ago
  • $60 - $80 per hour

     ...Network to contribute domain expertise to AI research. Experts in this network work...  ...remotely on contract engagements that train and evaluate AI models in physical sciences, design realistic...  ...Operate independently while meeting project milestones and deliverable... 
    Training
    Hourly pay
    Contract work
    Immediate start
    Remote work

    SaidGig

    United States
    29 days ago
  • $60 - $80 per hour

     ...expertise to help develop, test, and improve AI systems used for HR tasks. This is an...  ...the network makes you eligible for projects as they become available, matched on...  ...Responsibilities Participate in training and evaluating AI models focused on HR and administration... 
    Training
    Hourly pay
    Contract work
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    more than 2 months ago
  • $50 - $60 per hour

     ...Overview Support a high-priority technical expert panel for a frontier AI lab by interviewing shortlisted engineers and delivering clear, evidence-based assessments of their fit for an AI model training and evaluation engagement. Key Responsibilities Conduct... 
    Training
    Hourly pay
    Remote work
    10 hours per week
    Weekday work

    SaidGig

    Remote
    3 days ago
  • $70 - $80 per hour

     ...Role Overview Apply advanced drug safety expertise to help improve next-generation AI systems through high-quality evaluations, safety-report analysis, and structured feedback. This remote contractor opportunity is designed for professionals with experience authoring... 
    Training
    Hourly pay
    For contractors
    Remote work

    SaidGig

    Remote
    25 days ago
  • $70 - $126 per hour

     ...Role Overview Help train next generation AI systems by applying practical technical expertise to agentic AI workflows. You will work with autonomous...  ...tools, contributing real world input that improves how models learn, reason, and perform. Key Responsibilities Use... 
    Training
    Hourly pay
    For contractors
    Remote work

    SaidGig

    Canada
    a month ago
  • $60 - $100 per hour

     ...Overview Provide management consulting expertise to AI labs and companies by training and evaluating AI models, designing realistic consulting tasks and...  ...and model outputs Work independently on remote projects and communicate findings clearly to project teams... 
    Training
    Hourly pay
    Contract work
    Immediate start
    Remote work

    SaidGig

    United States
    more than 2 months ago
  • $70 - $90 per hour

     ...Role Overview Help evaluate Neuron Kernel Interface development tasks that support the training and evaluation of advanced AI models. You will assess kernel quality, numerical correctness, CUDA-to-NKI migration fidelity, and whether implementations are well suited to... 
    Training
    Hourly pay
    Remote work

    SaidGig

    Remote
    3 days ago
  • $11 - $19 per hour

     ...Evaluate AI-generated music and lyrics in Telugu and English, helping assess outputs...  ...music performance, theory, or composition training is preferred. Credited or published songwriting...  ...engagement with an immediate start. Project expected to last up to 6 months.... 
    Training
    Hourly pay
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    16 days ago
  •  ...Evaluate AI-generated music and lyrics across a broad range of genres, applying your Bengali music expertise to help assess quality...  ...suitable for critical listening. Preferred Qualifications Formal training in music performance, theory, or composition. Credited or... 
    Training
    Hourly pay
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    16 days ago
  • $15 per hour

     ...your Punjabi music expertise to evaluate AI-generated music and lyrics across a broad...  ...listening. Preferred Qualifications Formal training in music performance, theory, or...  ...changes to a per-task structure during the project, the effective hourly rate will remain... 
    Training
    Hourly pay
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    16 days ago
  • $35 - $62 per hour

     ...your Japanese music expertise to evaluate AI-generated music and lyrics across a wide...  ...listening. Preferred Qualifications Formal training in music performance, theory, or...  ...contractor engagement. Immediate start, with a project duration of up to six months. Flexible... 
    Training
    Hourly pay
    For contractors
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    16 days ago
  • $20 - $60 per hour

     ...Help train next-generation AI systems by creating rigorous, real-world evaluations that test how well advanced models learn, reason, and perform. This remote contract opportunity is open...  ...to reviewer feedback while following project guidelines and quality standards.... 
    Training
    Hourly pay
    Contract work
    For contractors
    Remote work

    SaidGig

    Indiana
    5 days ago
  • $60 - $150 per hour

     ...experienced legal professionals with AI labs and companies for short term contract...  ...legal subject matter expertise to train and evaluate AI models, create realistic tasks and deliverables...  ...work. Typical time commitment per project: 15 to 30 hours per week, projects vary... 
    Training
    Hourly pay
    Contract work
    Temporary work
    Immediate start
    Remote work

    SaidGig

    United States
    more than 2 months ago
  •  ...Top STEM PhD in Math or Physics to craft and review challenging math or physics problems. The role involves supporting training for large AI models while ensuring problem quality and relevance. Qualified candidates must possess a Master's or PhD from a top university,... 
    Training
    Remote job
    Hourly pay
    Contract work

    Crossing Hurdles

    New York, NY
    5 days ago
  • $60 - $80 per hour

     ...Apply your biochemistry expertise to help train next-generation AI systems for life sciences. You will...  ...-world scientific input that helps models learn, reason, and perform. No previous...  ...research, biotechnology, or health science projects, is preferred. Interest in... 
    Training
    Hourly pay
    Contract work
    Remote work

    SaidGig

    Remote
    a month ago
  • $60 - $120 per hour

     ...Apply statistical expertise to help train next-generation AI systems through high-quality, real-world...  ...communicate key findings and support model development. Annotate, label, or enrich...  .... Collaborate asynchronously with project stakeholders to clarify requirements,... 
    Training
    Hourly pay
    Contract work
    Remote work

    SaidGig

    Indiana
    16 days ago
  • $80 - $130 per hour

     ...materials science expertise to inform and train next-generation AI systems by analyzing experimental data...  ...high-quality inputs that improve how models learn and reason. No prior AI...  ...consistency, accuracy, and adherence to project specifications. Collaborate remotely... 
    Training
    Hourly pay
    For contractors
    Remote work

    SaidGig

    Indiana
    a month ago
  • $28 - $60 per hour

     ...Evaluate AI-generated music and lyrics across a wide range of genres, applying your knowledge of the Dutch music scene and strong...  ...appropriate for critical listening. Preferred Qualifications Formal training in music performance, theory, or composition. Credited or... 
    Training
    Hourly pay
    For contractors
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    5 days ago
  •  ...Role Overview Work with a leading AI lab to evaluate outputs from generative music models in German and English. This role focuses on listening, scoring, and annotating...  ...musical quality. Score the quality of musical training data used by models. Create and assess lyrical... 
    Training
    Hourly pay
    Part time
    Immediate start
    Remote work
    10 hours per week

    SaidGig

    United States
    a month ago
  • $70 - $90 per hour

     ...Review GPU and accelerator kernel development tasks that support the training and evaluation of advanced AI models. This role focuses on assessing task quality, numerical correctness, completeness, fair performance benchmarking, appropriate scope, and whether kernels compile... 
    Training
    Hourly pay
    Remote work

    SaidGig

    Remote
    3 days ago
  • $20 - $60 per hour

     ...Role Overview Help train and evaluate next-generation AI systems by creating rigorous, real-world assessments that test how advanced models learn, reason, and perform. This remote contract...  ...reviewer feedback while following project guidelines and quality standards.... 
    Training
    Hourly pay
    Contract work
    For contractors
    Remote work

    SaidGig

    United States
    5 days ago
  • $60 - $80 per hour

     ...Apply your marketing expertise to help train and evaluate AI systems through future remote contract projects aligned with your background and interests. Role Overview...  ...Key Responsibilities Train and evaluate AI models in marketing-related work. Create tasks and deliverables... 
    Training
    Hourly pay
    Contract work
    Remote work

    SaidGig

    United States
    more than 2 months ago
  • $10 - $20 per hour

     ...refine, and validate high-quality linguistic content that trains next-generation AI systems. This remote contract role focuses on translation,...  ...language, culture, and delivering high-quality work on every project. Preferred: degree in Linguistics, Translation, Malayalam... 
    Training
    Hourly pay
    Full time
    Contract work
    Part time
    For contractors
    Remote work

    SaidGig

    United States
    26 days ago
  • $15 per hour

     ...Evaluate AI-generated music and lyrics in Malayalam and English, helping assess outputs across a broad range of genres against detailed...  ...for critical listening. Preferred Qualifications Formal training in music performance, theory, or composition. Credited or... 
    Training
    Hourly pay
    Immediate start
    Remote work
    Flexible hours

    SaidGig

    United States
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Project Lead, AI Model Training. Be the first to apply!