Remote AI Research Scientist: LLM Evaluation & Experiments
$30 - $50 per hourREX
- Remote job
A tech company is seeking an AI Researcher to support end-to-end research for modern AI systems. This remote role involves designing experiments, defining evaluation protocols, and improving evaluation rigor for large language models. Key responsibilities include developing AI research experiments, performing error analysis, and supporting data quality practices. Ideal candidates will have a fundamental understanding of AI and machine learning, along with experience in NLP or computer vision. Competitive pay ranges from $30 to $50 per hour. #J-18808-Ljbffr REX
$30 - $50 per hour
A tech company is seeking an AI Researcher to support end-to-end research for modern AI systems. This remote role involves designing experiments, defining evaluation protocols, and improving evaluation rigor for large language models. Key responsibilities include developing...Remote jobHourly pay- Rex.zone is seeking an AI Research Scientist to lead applied AI research projects for US-based... ...-ended questions into measurable experiments in LLM evaluation and RLHF data design. You will evaluate... ..., and usefulness. The role is remote in the United States, with compensation...Remote jobHourly payFlexible hours
$245k - $315k
...Applied Research Scientist, LLM Evaluation & Post-Training Innodata is expanding its GenAI research capability... ...work across human-in-the-loop and AI-augmented workflows, partnering with... ...outcomes, and you will design experiments that produce credible, actionable conclusions...Remote work- Research Scientist, LLM Evaluation & Post-Training page is loaded## Research Scientist,... ...& Post-Traininglocations: Remote Work( USA)time type: Full... ...Centific**Centific is a frontier AI data foundry that curates... ...model improvement. Design experiments to study how evaluation...Remote workFull time
- ...Accelerator Program seeks an early-career researcher to work on problems at the frontier of LLM reasoning and post-training methodology. You will run experiments, form independent hypotheses, and... ...production constraints. This is a remote internship designed for current...Remote jobInternship
- ..., processes, and AI into a single, governed... ...best company for remote... ...an exceptional AI Research Scientist to join our growing... ...optimised RAG, tool‑use evaluation, and multi‑agent... ...RequirementsQualifications / Experience / Technical... .../JAX and modern LLM frameworks.Strong...Remote workFlexible hours
- This AI Research Scientist will lead the design and build biological... ...project work. Fully remote applicants will not... ...AI systems (e.g., LLM-based tools, agentic... ...and develop rigorous evaluation frameworks and benchmarks... ...validation experiments. Lead by example in the...Remote work3 days per week
- AI Research Scientist, Learning & Evaluation Studyfetch Beverly Hills, California, United States About this position... ...bar for the team. How we run experiments, what counts as a result, when a change... ...methods, hypothesis testing AI/LLM: eval frameworks, LLM-as-judge and...Work at officeWorldwide
- ...We are looking for an AI Evaluation Scientist to design and execute evaluation processes that... ...tools, frameworks, metrics, and research related to LLM assessment and generative AI reliability... ...or a related field and 5+ years of experience. Master's degree in Computer...Remote work
$60 - $90 per hour
...technical talent with leading AI research labs. Headquartered in San... ...Machine Learning Engineer — Model Evaluation & Experimentation... ...90/hour Location: Remote Commitment: 35 hours... ...multi-step tasks. Run experiments by implementing changes, executing...Remote workHourly payWeekly payFull timeContract workFor contractorsSummer work- ...accelerate next‑generation computing experiences—from AI and data centers to PCs, gaming... ...and beyond. The Role Lead AI Research Scientist, Reinforcement Learning (LLM) and Post‑Training. You... ...misspecification, variance reduction, and evaluation that reflects real constraints—...
- ...on behalf of UL Research Institutes. We have... ...a Lead Research Scientist in AI at UL Research Institutes... ...DSRI). This is a REMOTE opportunity.... ...in NLP and LLM safety including... ...independent test and evaluation research programs... .... What you’ll experience working at UL Research...Remote workWorldwideFlexible hours
- ...first enterprise AI company. We... ...Cohere is a team of researchers, engineers,... ...!Why this role?Evaluation is critical to... ...infrastructure to measure LLM progress.As a Senior Research Scientist, Model... ...offices if you are remote, plus an annual... ...with your experience, we still encourage...Remote workFull timeWork at officeLocal areaHome office
- Cohere is seeking a Senior Research Scientist, Model Evaluation, to create ambitious evaluation... ...for enterprise AI. You will work with cross-functional... ...to push the frontiers of LLM evaluation. The role emphasizes... ...and engineering, with remote-friendly policies and opportunities...Remote work
- ...Working remotely in a full-time capacity, the AI Research Scientist will focus on developing agentic systems... ...criteria for evaluating systems Required qualifications... ...modern research literature Experience in training generative... ...solid understanding of LLM training fundamentals...Remote workFull time
$100 - $120 per hour
...machine learning research across computer vision... ..., improving, evaluating, and deploying deep... ...learning research experience. PhD research counts... ...or comparable AI company, or an equivalent... ...generators. LLM post-training and... ...Work Terms Remote, hourly engagement...Remote workHourly payFlexible hours- ...Job Title AI Research Scientist Location Hybrid / Remote Employment Type Full-time Job Summary... ...generative AI. Design, develop, and evaluate client AI models, algorithms,... ...experimentation. Design and execute experiments to evaluate model performance,...Remote workFull time
$117.6k - $176.4k
...Data Scientist At Schneider Electric, we are committed... .... Within our Global AI Hub we combine our... ...fast-moving applied AI research scientist who loves working... ...data preparation, experiment tracking, baselines, ablations... ...baselines, ablations, evaluation protocols) and...Remote workFull timeTemporary work$40 per hour
A leading AI development company is seeking experienced quantitative... ...-edge AI systems. This fully remote role allows for flexible hours and involves evaluating AI-generated work, designing quantitative... ...should possess hands-on experience in fields like data science or statistics...Remote workHourly payFlexible hours$167.8k - $209.7k
...AI Research Scientist II – Office of the CTO The mission of the Allen Institute... ..., disseminating, and evaluating AI/ML/computational methods... ...Required Education and Experience PhD in Computer Science... ...currently able to work both remotely and onsite in a hybrid work...Remote workWork experience placementWork at officeVisa sponsorshipWork visaRelocation package- Codertal is hiring an AI Research Scientist for a remote opportunity in the European Union on a B2B contract... ...Contract Type: B2B Contract Experience Level: Mid to Senior About the Role... ...Scientist to explore, design, and evaluate next‑generation AI models. You’ll work...Remote workDaily paidContract workFlexible hours
$146.6k - $183.25k
...AI Research Scientist - AI Biological Design The Allen Institute accelerates... ...the AI Research Scientist evaluates model performance, limitations... ...computational frameworks, experiments, and benchmarks Assess... ...is 700 Dexter Ave N.; any remote work must be performed in Washington...Remote workWork at officeLocal areaVisa sponsorshipWork visaRelocation package$120k - $170k
...WhatYou’llBeLaunching As an AI Research Scientist, you will join the AI and... ...field + Minimum of 2 years' experience in similar role (or PhD with... ..., simulation frameworks, evaluation harnesses) to refine simulation... ...-ready, send it. Location: Remote, US Salary Range: $120,000...Remote workFull timeWork experience placementCurrently hiring- ...Global Solutions, Inc. is seeking a Research Scientist for LLM Evaluation & Post-Training in Seattle, WA. This... ...evaluation frameworks and collaborating with AI industry stakeholders. Candidates... ...relevant field, at least 5 years of experience in applied ML research, and strong...Full time
$55 - $80 per hour
...technical talent with leading AI research labs. Headquartered in San... ...55–$80/hour Location: Remote Role Responsibilities... ...problems in your domain. Evaluate and improve AI model performance... ...(or equivalent industry experience) in Medicine , Healthcare...Remote workContract workSummer work$175k - $275k
Full-Time in Austin, TX Remote (any location) - Senior - Product... ...175k - $275k Applied Data Scientist, LLM Evaluation Introduction At Driver, we’... ...that provides a rich user experience. About Driver We’re an... ...context layer for employees and AI agents alike to use in...Remote jobFull timeFlexible hours- ...position of STEM Computational Scientific Software & Evaluation Design in Astrophysics & Cosmology. This position is remote with a commitment of 15-20 hours per week.... ...include designing computational problems, evaluating AI systems, and developing strategies for data...Remote jobContract work
$159.75k - $255.6k
...skilled and innovative Senior AI Research Scientist to join a new team focusing... ...to Protect Life but your experience doesn’t align perfectly... ...with the flexibility to work remotely on Mondays, unless there is... ...through model development, evaluation, deployment, and iteration...Remote workWork experience placementWork at office$25 - $30 per hour
...Bilingual Traditional Chinese AI Evaluation Specialist is a remote Chinese specialist track for evaluating chinese... ...Specialist work. Demonstrable experience editing, translating, or evaluating... ...QA frameworks. Familiarity with LLM evaluation rubrics and inter-rater agreement...Remote workFor contractors10 hours per week- ...next-generation computing experiences—from AI and data centers, to PCs, gaming... ...hiring Forward Deployed AI Research Scientist to help bring advanced AI... ..., technical plans, evaluation criteria, and prototype paths... ...human-in-the-loop review, and LLM-as-judge methods where appropriate...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Remote AI Research Scientist: LLM Evaluation & Experiments. Be the first to apply!
- ai data scientist New York, NY
- ai scientist New York, NY
- molecular biology scientist New York, NY
- water quality scientist New York, NY
- cosmetic scientist New York, NY
- machine learning scientist New York, NY
- principal applied scientist New York, NY
- bioanalytical scientist New York, NY
- image scientist New York, NY
- machine learning research scientist New York, NY




