Get new jobs by email
$256k - $307k
.... You will work closely with a team of talented engineers and researchers in a fast-paced, collaborative environment.In this role, you will... ...simulation data, to extract actionable insights.Develop and evaluate novel algorithms and methodologies for data processing,...SuggestedFull timeTemporary workRelocation package$100k - $140k
...unmatched technical expertise. As the operator of a federally funded research and development center (FFRDC), we are broadly engaged across... ...for the role of Materials Physicist - Non-Destructive Evaluation (NDE). In this role, you will work alongside senior technical...SuggestedFull timeImmediate startRemote workRelocation packageFlexible hours$146k - $280k
Waabi is seeking a Senior Applied Data Scientist in San Francisco to shape evaluation methodologies for autonomous driving technology. Responsibilities include designing production frameworks, prototyping analyses, and developing analytical models to correlate simulation...SuggestedFlexible hours- Reflection Research Lab in San Francisco is seeking a candidate to conduct critical comparative analysis to advance our understanding of model capabilities. You will build and refine evaluation systems that create tight feedback loops between data, evals, and model behavior...Suggested
$180.6k - $225.75k
...to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers... ...expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation. This role is on the evaluation pod within the GenAI Research...SuggestedFull time$272k - $431.25k
...boundaries of multimodal AI, robotics, and world foundation models for Physical AI. We are looking for a Senior Research Manager to lead world-model evaluation and benchmarking across NVIDIA’s Physical AI model portfolio. This role will build the team and research agenda...SuggestedFull time$216k - $270k
Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities and safeguarding AI models and systems. Building on this expertise, Scale Labs...SuggestedFull time- AI Research Scientist, Learning & Evaluation Studyfetch Beverly Hills, California, United States About this position About Studyfetch StudyFetch is the #1 AI-native learning platform globally, transforming how millions of students learn through personalized AI-powered education...SuggestedWork at officeWorldwide
- ...expertise for next‑gen AI systems. Remote contractor role focusing on evaluating complex physics arguments and providing authoritative judgments. The ideal candidate has a strong independent research record and leadership in science. No prior AI experience is required;...SuggestedRemote jobFor contractors
- UCLA Center for Labor Research and Education seeks a Senior Research Analyst to manage components of the DEO High Road Training Partnership project. The role leads evaluation design, data collection, analyses, and reporting, while collaborating with external partners to...Suggested
$124.2k - $156k
A higher education institution in Berkeley, CA, is seeking an Assistant Researcher with a focus on nuclear data evaluation. The successful candidate will collaborate with Lawrence Berkeley National Laboratory and the International Atomic Energy Agency to analyze datasets...Suggested- ...company, is hiring a Member of Technical Staff in Data Analysis and Evaluation. You will design data-collection tasks, apply statistical methods, and assess model robustness. Collaborate with researchers and engineers to improve dataset quality and model performance,...Suggested
- Cohere is looking for a Member of Technical Staff in Data Analysis and Evaluation to ensure the quality and performance of large language models. The role involves designing data collection tasks, collaborating with teams, and applying statistical methods for data evaluation...Suggested
- Thinking Machines in San Francisco seeks a researcher to advance internal evaluations and signals for post-training models. You will collaborate with researchers and engineers across the research organization, shaping evaluation creation, usability, and auditing. Your...Suggested
- The Research Analyst will support the evaluation of the Economic Liberation Project for the POWER team. Key responsibilities include coordinating research and evaluation data-collection efforts; cleaning, coding, and analyzing data; writing reports and briefs; and disseminating...Suggested
$150k - $200k
...leading creative preference datasets that power benchmarking, evaluation, and post-training for the world's leading AI models and applications... .... Why this role exists Contra Labs is expanding its research work with frontier AI labs across evaluation, post-training data...- Scale AI, Inc. seeks Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation. The role focuses on building benchmarks and diagnosing model failure modes in text and multimodal modalities within the GenAI...
- The UCLA Center for Labor Research and Education is seeking a Senior Research Analyst to manage and lead evaluation activities for the DEO High Road Training Partnership project. You will oversee design, data collection, analysis, and reporting while collaborating with...
- What the role actually is Nuro is hiring an applied AI researcher focused on agent systems and evaluation. The work sits around the research and measurement loop for autonomous driving systems, with emphasis on how agent behavior is tested, compared, and improved. For...
- Anthropic in San Francisco seeks a Research Scientist to measure recursive-self-improvement in large models. You will design evaluations, build models of capability growth, and interpret results to guide research direction. Senior candidates will combine hands-on work with...
$184k - $287.5k
...advance the frontiers of accelerated computing.We’re looking for a research engineer to join our team to develop simulation tools built... .... Our goal is to build the industry's leading tool for evaluating robot foundation models in simulation. This mission builds on...Full time- Verita AI is seeking an Applied AI Researcher to work with clients on model evaluation and data strategy. You will assess model performance, identify failure modes, and design data-driven solutions, collaborating with operations and engineering to implement scalable data...
- Nuro is hiring an applied AI researcher focused on agent systems and evaluation. The work sits around the research and measurement loop for autonomous driving systems, with emphasis on how agent behavior is tested, compared, and improved. For an impact-focused role, the...
- Sanas is a leading force in real-time speech AI, advancing evaluation-driven research across accent translation, noise cancellation, and language translation. We seek a Research Scientist to define meaningful progress metrics and build rigorous evaluation infrastructure...
- California Department of Public Health is seeking a Home Visiting Performance Measurement & Evaluation Unit Supervisor (Research Scientist Supervisor I) to lead scientific and programmatic oversight of home visiting evaluations. You will guide a team in designing studies...Local area
- ...content, temporal coherence — to improve pretraining signal Build evaluation frameworks and benchmarks to measure causal video model... ...horizon rollout fidelity, and downstream robot task performance Research and implement data selection, mixing, and weighting strategies...
- Sanas in Palo Alto is seeking a Research Scientist focused on rigorous evaluation of speech AI models. You will define meaningful metrics, build scalable evaluation pipelines, and align research progress with product impact across Accent Translation, Noise Cancellation,...
- ...future of human communication. Founded by a team of Stanford researchers and entrepreneurs with deep industry experience, Sanas has developed... ...means across all of Sanas's model families, build the evaluation infrastructure to measure it rigorously, and close the loop between...
$196k - $230k
...deeply about giving our customers more time for their life’s work.About the Role:We’re seeking an experienced UX Researcher to define and scale how we evaluate Notion’s AI-powered experiences—focusing on what “good” looks like not only for model output quality, but for...Local areaShift work$70 - $110 per hour
...AI systems by bringing rigorous materials science and engineering judgment to the evaluation, design, and improvement of technical knowledge work. You will work closely with an AI research team to define what high quality materials reasoning looks like in practice and...Hourly payFull timeLive inRelocationRelocation package

