Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Remote AI Research Scientist: LLM Evaluation & Experiments

$30 - $50 per hour

REX

New York, NY
  • Remote job

A tech company is seeking an AI Researcher to support end-to-end research for modern AI systems. This remote role involves designing experiments, defining evaluation protocols, and improving evaluation rigor for large language models. Key responsibilities include developing AI research experiments, performing error analysis, and supporting data quality practices. Ideal candidates will have a fundamental understanding of AI and machine learning, along with experience in NLP or computer vision. Competitive pay ranges from $30 to $50 per hour. #J-18808-Ljbffr REX

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Remote AI Research Scientist: LLM Evaluation & Experiments in New York, NY vacancy
  • $30 - $50 per hour

    A tech company is seeking an AI Researcher to support end-to-end research for modern AI systems. This remote role involves designing experiments, defining evaluation protocols, and improving evaluation rigor for large language models. Key responsibilities include developing... 
    Remote job
    Hourly pay

    Rex.zone

    New York, NY
    2 days ago
  • Rex.zone is seeking an AI Research Scientist to lead applied AI research projects for US-based...  ...-ended questions into measurable experiments in LLM evaluation and RLHF data design. You will evaluate...  ..., and usefulness. The role is remote in the United States, with compensation... 
    Remote job
    Hourly pay
    Flexible hours

    AIToolboard

    New York, NY
    3 days ago
  • $245k - $315k

     ...Applied Research Scientist, LLM Evaluation & Post-Training Innodata is expanding its GenAI research capability...  ...work across human-in-the-loop and AI-augmented workflows, partnering with...  ...outcomes, and you will design experiments that produce credible, actionable conclusions... 
    Remote work

    Innodata Inc.

    United States
    2 days ago
  • Research Scientist, LLM Evaluation & Post-Training page is loaded## Research Scientist,...  ...& Post-Traininglocations: Remote Work( USA)time type: Full...  ...Centific**Centific is a frontier AI data foundry that curates...  ...model improvement. Design experiments to study how evaluation... 
    Remote work
    Full time

    Centific Global Solutions, Inc.

    Seattle, WA
    2 days ago
  •  ...Accelerator Program seeks an early-career researcher to work on problems at the frontier of LLM reasoning and post-training methodology. You will run experiments, form independent hypotheses, and...  ...production constraints. This is a remote internship designed for current... 
    Remote job
    Internship

    binance

    New York, NY
    3 days ago
  •  ..., processes, and AI into a single, governed...  ...best company for remote...  ...an exceptional AI Research Scientist to join our growing...  ...optimised RAG, tool‑use evaluation, and multi‑agent...  ...RequirementsQualifications / Experience / Technical...  .../JAX and modern LLM frameworks.Strong... 
    Remote work
    Flexible hours

    Workato

    San Francisco, CA
    3 days ago
  • This AI Research Scientist will lead the design and build biological...  ...project work. Fully remote applicants will not...  ...AI systems (e.g., LLM-based tools, agentic...  ...and develop rigorous evaluation frameworks and benchmarks...  ...validation experiments. Lead by example in the... 
    Remote work
    3 days per week

    Q-state Biosciences

    Cambridge, MA
    1 day ago
  • AI Research Scientist, Learning & Evaluation Studyfetch Beverly Hills, California, United States About this position...  ...bar for the team. How we run experiments, what counts as a result, when a change...  ...methods, hypothesis testing AI/LLM: eval frameworks, LLM-as-judge and... 
    Work at office
    Worldwide

    Studyfetch

    Beverly Hills, CA
    1 day ago
  •  ...We are looking for an AI Evaluation Scientist to design and execute evaluation processes that...  ...tools, frameworks, metrics, and research related to LLM assessment and generative AI reliability...  ...or a related field and 5+ years of experience. Master's degree in Computer... 
    Remote work

    Convergenz

    United States
    3 days ago
  • $60 - $90 per hour

     ...technical talent with leading AI research labs. Headquartered in San...  ...Machine Learning Engineer — Model Evaluation & Experimentation...  ...90/hour Location: Remote Commitment: 35 hours...  ...multi-step tasks. Run experiments by implementing changes, executing... 
    Remote work
    Hourly pay
    Weekly pay
    Full time
    Contract work
    For contractors
    Summer work

    Mercor

    Remote
    14 hours ago
  •  ...accelerate next‑generation computing experiences—from AI and data centers to PCs, gaming...  ...and beyond. The Role Lead AI Research Scientist, Reinforcement Learning (LLM) and Post‑Training. You...  ...misspecification, variance reduction, and evaluation that reflects real constraints—... 

    AMD

    Santa Clara, CA
    2 days ago
  •  ...on behalf of UL Research Institutes. We have...  ...a Lead Research Scientist in AI at UL Research Institutes...  ...DSRI). This is a REMOTE opportunity....  ...in NLP and LLM safety including...  ...independent test and evaluation research programs...  .... What you’ll experience working at UL Research... 
    Remote work
    Worldwide
    Flexible hours

    Lucas James Talent Partners

    Evanston, IL
    a month ago
  •  ...first enterprise AI company. We...  ...Cohere is a team of researchers, engineers,...  ...!Why this role?Evaluation is critical to...  ...infrastructure to measure LLM progress.As a Senior Research Scientist, Model...  ...offices if you are remote, plus an annual...  ...with your experience, we still encourage... 
    Remote work
    Full time
    Work at office
    Local area
    Home office

    Cohere

    New York, NY
    2 days ago
  • Cohere is seeking a Senior Research Scientist, Model Evaluation, to create ambitious evaluation...  ...for enterprise AI. You will work with cross-functional...  ...to push the frontiers of LLM evaluation. The role emphasizes...  ...and engineering, with remote-friendly policies and opportunities... 
    Remote work

    cohere

    New York, NY
    1 day ago
  •  ...Working remotely in a full-time capacity, the AI Research Scientist will focus on developing agentic systems...  ...criteria for evaluating systems Required qualifications...  ...modern research literature Experience in training generative...  ...solid understanding of LLM training fundamentals... 
    Remote work
    Full time

    Virtual Vocations Inc

    United States
    5 days ago
  • $100 - $120 per hour

     ...machine learning research across computer vision...  ..., improving, evaluating, and deploying deep...  ...learning research experience. PhD research counts...  ...or comparable AI company, or an equivalent...  ...generators. LLM post-training and...  ...Work Terms Remote, hourly engagement... 
    Remote work
    Hourly pay
    Flexible hours

    SaidGig

    United States
    20 days ago
  •  ...Job Title AI Research Scientist Location Hybrid / Remote Employment Type Full-time Job Summary...  ...generative AI. Design, develop, and evaluate client AI models, algorithms,...  ...experimentation. Design and execute experiments to evaluate model performance,... 
    Remote work
    Full time

    Ova Technologies

    New York, NY
    2 days ago
  • $117.6k - $176.4k

     ...Data Scientist At Schneider Electric, we are committed...  .... Within our Global AI Hub we combine our...  ...fast-moving applied AI research scientist who loves working...  ...data preparation, experiment tracking, baselines, ablations...  ...baselines, ablations, evaluation protocols) and... 
    Remote work
    Full time
    Temporary work

    Schneider Electric

    United States
    5 days ago
  • $40 per hour

    A leading AI development company is seeking experienced quantitative...  ...-edge AI systems. This fully remote role allows for flexible hours and involves evaluating AI-generated work, designing quantitative...  ...should possess hands-on experience in fields like data science or statistics... 
    Remote work
    Hourly pay
    Flexible hours

    DataAnnotation

    United States
    15 hours ago
  • $167.8k - $209.7k

     ...AI Research Scientist II – Office of the CTO The mission of the Allen Institute...  ..., disseminating, and evaluating AI/ML/computational methods...  ...Required Education and Experience PhD in Computer Science...  ...currently able to work both remotely and onsite in a hybrid work... 
    Remote work
    Work experience placement
    Work at office
    Visa sponsorship
    Work visa
    Relocation package

    The Allen Institute

    Seattle, WA
    1 day ago
  • Codertal is hiring an AI Research Scientist for a remote opportunity in the European Union on a B2B contract...  ...Contract Type: B2B Contract Experience Level: Mid to Senior About the Role...  ...Scientist to explore, design, and evaluate next‑generation AI models. You’ll work... 
    Remote work
    Daily paid
    Contract work
    Flexible hours

    Codertal

    Union, NJ
    2 days ago
  • $146.6k - $183.25k

     ...AI Research Scientist - AI Biological Design The Allen Institute accelerates...  ...the AI Research Scientist evaluates model performance, limitations...  ...computational frameworks, experiments, and benchmarks Assess...  ...is 700 Dexter Ave N.; any remote work must be performed in Washington... 
    Remote work
    Work at office
    Local area
    Visa sponsorship
    Work visa
    Relocation package

    Allen Institute

    Washington DC
    a month ago
  • $120k - $170k

     ...WhatYou’llBeLaunching As an AI Research Scientist, you will join the AI and...  ...field + Minimum of 2 years' experience in similar role (or PhD with...  ..., simulation frameworks, evaluation harnesses) to refine simulation...  ...-ready, send it. Location: Remote, US Salary Range: $120,000... 
    Remote work
    Full time
    Work experience placement
    Currently hiring

    Slingshot Aerospace

    New York, NY
    3 days ago
  •  ...Global Solutions, Inc. is seeking a Research Scientist for LLM Evaluation & Post-Training in Seattle, WA. This...  ...evaluation frameworks and collaborating with AI industry stakeholders. Candidates...  ...relevant field, at least 5 years of experience in applied ML research, and strong... 
    Full time

    Centific Global Solutions, Inc.

    Seattle, WA
    2 days ago
  • $55 - $80 per hour

     ...technical talent with leading AI research labs. Headquartered in San...  ...55–$80/hour Location: Remote Role Responsibilities...  ...problems in your domain. Evaluate and improve AI model performance...  ...(or equivalent industry experience) in Medicine , Healthcare... 
    Remote work
    Contract work
    Summer work

    Mercor

    San Francisco, CA
    16 days ago
  • $175k - $275k

    Full-Time in Austin, TX Remote (any location) - Senior - Product...  ...175k - $275k Applied Data Scientist, LLM Evaluation Introduction At Driver, we’...  ...that provides a rich user experience. About Driver We’re an...  ...context layer for employees and AI agents alike to use in... 
    Remote job
    Full time
    Flexible hours

    Driverai

    Austin, TX
    2 days ago
  •  ...position of STEM Computational Scientific Software & Evaluation Design in Astrophysics & Cosmology. This position is remote with a commitment of 15-20 hours per week....  ...include designing computational problems, evaluating AI systems, and developing strategies for data... 
    Remote job
    Contract work

    Mercor

    New York, NY
    2 days ago
  • $159.75k - $255.6k

     ...skilled and innovative Senior AI Research Scientist to join a new team focusing...  ...to Protect Life but your experience doesn’t align perfectly...  ...with the flexibility to work remotely on Mondays, unless there is...  ...through model development, evaluation, deployment, and iteration... 
    Remote work
    Work experience placement
    Work at office

    Axon

    Seattle, WA
    1 day ago
  • $25 - $30 per hour

     ...Bilingual Traditional Chinese AI Evaluation Specialist is a remote Chinese specialist track for evaluating chinese...  ...Specialist work. Demonstrable experience editing, translating, or evaluating...  ...QA frameworks. Familiarity with LLM evaluation rubrics and inter-rater agreement... 
    Remote work
    For contractors
    10 hours per week

    AuraOne Human Data

    Remote
    29 days ago
  •  ...next-generation computing experiences—from AI and data centers, to PCs, gaming...  ...hiring Forward Deployed AI Research Scientist to help bring advanced AI...  ..., technical plans, evaluation criteria, and prototype paths...  ...human-in-the-loop review, and LLM-as-judge methods where appropriate... 

    AMD

    Santa Clara, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Remote AI Research Scientist: LLM Evaluation & Experiments. Be the first to apply!