Get new jobs by email
  • $256k - $307k

     .... You will work closely with a team of talented engineers and researchers in a fast-paced, collaborative environment.In this role, you will...  ...simulation data, to extract actionable insights.Develop and evaluate novel algorithms and methodologies for data processing,... 
    Suggested
    Full time
    Temporary work
    Relocation package

    Zoox

    Foster, CA
    3 days ago
  • $100k - $140k

     ...unmatched technical expertise. As the operator of a federally funded research and development center (FFRDC), we are broadly engaged across...  ...for the role of Materials Physicist - Non-Destructive Evaluation (NDE). In this role, you will work alongside senior technical... 
    Suggested
    Full time
    Immediate start
    Remote work
    Relocation package
    Flexible hours

    The Aerospace Corporation

    El Segundo, CA
    1 day ago
  • $146k - $280k

    Waabi is seeking a Senior Applied Data Scientist in San Francisco to shape evaluation methodologies for autonomous driving technology. Responsibilities include designing production frameworks, prototyping analyses, and developing analytical models to correlate simulation... 
    Suggested
    Flexible hours

    Waabi

    San Francisco, CA
    4 days ago
  • Reflection Research Lab in San Francisco is seeking a candidate to conduct critical comparative analysis to advance our understanding of model capabilities. You will build and refine evaluation systems that create tight feedback loops between data, evals, and model behavior... 
    Suggested

    Visa Hunt

    San Francisco, CA
    1 day ago
  • $180.6k - $225.75k

     ...to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers...  ...expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation. This role is on the evaluation pod within the GenAI Research... 
    Suggested
    Full time

    Scale AI

    San Francisco, CA
    3 days ago
  • $272k - $431.25k

     ...boundaries of multimodal AI, robotics, and world foundation models for Physical AI. We are looking for a Senior Research Manager to lead world-model evaluation and benchmarking across NVIDIA’s Physical AI model portfolio. This role will build the team and research agenda... 
    Suggested
    Full time

    Nvidia

    Santa Clara, CA
    7 days ago
  • $216k - $270k

    Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities and safeguarding AI models and systems. Building on this expertise, Scale Labs... 
    Suggested
    Full time

    Scale AI

    San Francisco, CA
    4 days ago
  • AI Research Scientist, Learning & Evaluation Studyfetch Beverly Hills, California, United States About this position About Studyfetch StudyFetch is the #1 AI-native learning platform globally, transforming how millions of students learn through personalized AI-powered education... 
    Suggested
    Work at office
    Worldwide

    Studyfetch

    Beverly Hills, CA
    2 days ago
  •  ...expertise for next‑gen AI systems. Remote contractor role focusing on evaluating complex physics arguments and providing authoritative judgments. The ideal candidate has a strong independent research record and leadership in science. No prior AI experience is required;... 
    Suggested
    Remote job
    For contractors

    YO AI Labs

    Palo Alto, CA
    17 hours ago
  • UCLA Center for Labor Research and Education seeks a Senior Research Analyst to manage components of the DEO High Road Training Partnership project. The role leads evaluation design, data collection, analyses, and reporting, while collaborating with external partners to... 
    Suggested

    UCLA

    Los Angeles, CA
    17 hours ago
  • $124.2k - $156k

    A higher education institution in Berkeley, CA, is seeking an Assistant Researcher with a focus on nuclear data evaluation. The successful candidate will collaborate with Lawrence Berkeley National Laboratory and the International Atomic Energy Agency to analyze datasets... 
    Suggested

    Inside Higher Ed

    Berkeley, CA
    4 days ago
  •  ...company, is hiring a Member of Technical Staff in Data Analysis and Evaluation. You will design data-collection tasks, apply statistical methods, and assess model robustness. Collaborate with researchers and engineers to improve dataset quality and model performance,... 
    Suggested

    Cohere

    San Francisco, CA
    1 day ago
  • Cohere is looking for a Member of Technical Staff in Data Analysis and Evaluation to ensure the quality and performance of large language models. The role involves designing data collection tasks, collaborating with teams, and applying statistical methods for data evaluation... 
    Suggested

    Cohere

    San Francisco, CA
    3 days ago
  • Thinking Machines in San Francisco seeks a researcher to advance internal evaluations and signals for post-training models. You will collaborate with researchers and engineers across the research organization, shaping evaluation creation, usability, and auditing. Your... 
    Suggested

    Mosaic.tech

    San Francisco, CA
    2 days ago
  • The Research Analyst will support the evaluation of the Economic Liberation Project for the POWER team. Key responsibilities include coordinating research and evaluation data-collection efforts; cleaning, coding, and analyzing data; writing reports and briefs; and disseminating... 
    Suggested

    The Regents of the University of California on behalf of the...

    Los Angeles, CA
    17 hours ago
  • $150k - $200k

     ...leading creative preference datasets that power benchmarking, evaluation, and post-training for the world's leading AI models and applications...  .... Why this role exists Contra Labs is expanding its research work with frontier AI labs across evaluation, post-training data... 

    ConTra

    San Francisco, CA
    1 day ago
  • Scale AI, Inc. seeks Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation. The role focuses on building benchmarks and diagnosing model failure modes in text and multimodal modalities within the GenAI... 

    Scale AI, Inc.

    San Francisco, CA
    3 days ago
  • The UCLA Center for Labor Research and Education is seeking a Senior Research Analyst to manage and lead evaluation activities for the DEO High Road Training Partnership project. You will oversee design, data collection, analysis, and reporting while collaborating with... 

    University of California, Los Angeles

    Los Angeles, CA
    17 hours ago
  • What the role actually is Nuro is hiring an applied AI researcher focused on agent systems and evaluation. The work sits around the research and measurement loop for autonomous driving systems, with emphasis on how agent behavior is tested, compared, and improved. For... 

    The Robot Age

    Mountain View, CA
    7 hours agonew
  • Anthropic in San Francisco seeks a Research Scientist to measure recursive-self-improvement in large models. You will design evaluations, build models of capability growth, and interpret results to guide research direction. Senior candidates will combine hands-on work with... 

    Anthropic

    San Francisco, CA
    17 hours ago
  • $184k - $287.5k

     ...advance the frontiers of accelerated computing.We’re looking for a research engineer to join our team to develop simulation tools built...  .... Our goal is to build the industry's leading tool for evaluating robot foundation models in simulation. This mission builds on... 
    Full time

    Nvidia

    Santa Clara, CA
    6 days ago
  • Verita AI is seeking an Applied AI Researcher to work with clients on model evaluation and data strategy. You will assess model performance, identify failure modes, and design data-driven solutions, collaborating with operations and engineering to implement scalable data... 

    Verita AI

    San Francisco, CA
    4 days ago
  • Nuro is hiring an applied AI researcher focused on agent systems and evaluation. The work sits around the research and measurement loop for autonomous driving systems, with emphasis on how agent behavior is tested, compared, and improved. For an impact-focused role, the... 

    The Robot Age

    Mountain View, CA
    7 hours agonew
  • Sanas is a leading force in real-time speech AI, advancing evaluation-driven research across accent translation, noise cancellation, and language translation. We seek a Research Scientist to define meaningful progress metrics and build rigorous evaluation infrastructure... 

    Sanas.AI Inc.

    Palo Alto, CA
    2 days ago
  • California Department of Public Health is seeking a Home Visiting Performance Measurement & Evaluation Unit Supervisor (Research Scientist Supervisor I) to lead scientific and programmatic oversight of home visiting evaluations. You will guide a team in designing studies... 
    Local area

    California Department of Public Health

    Sacramento, CA
    2 days ago
  •  ...content, temporal coherence — to improve pretraining signal Build evaluation frameworks and benchmarks to measure causal video model...  ...horizon rollout fidelity, and downstream robot task performance Research and implement data selection, mixing, and weighting strategies... 

    Rhoda AI

    Mountain View, CA
    17 hours ago
  • Sanas in Palo Alto is seeking a Research Scientist focused on rigorous evaluation of speech AI models. You will define meaningful metrics, build scalable evaluation pipelines, and align research progress with product impact across Accent Translation, Noise Cancellation,... 

    Sanas

    Palo Alto, CA
    1 day ago
  •  ...future of human communication. Founded by a team of Stanford researchers and entrepreneurs with deep industry experience, Sanas has developed...  ...means across all of Sanas's model families, build the evaluation infrastructure to measure it rigorously, and close the loop between... 

    Sanas.AI Inc.

    Palo Alto, CA
    2 days ago
  • $196k - $230k

     ...deeply about giving our customers more time for their life’s work.About the Role:We’re seeking an experienced UX Researcher to define and scale how we evaluate Notion’s AI-powered experiences—focusing on what “good” looks like not only for model output quality, but for... 
    Local area
    Shift work

    Notion Labs

    San Francisco, CA
    3 days ago
  • $70 - $110 per hour

     ...AI systems by bringing rigorous materials science and engineering judgment to the evaluation, design, and improvement of technical knowledge work. You will work closely with an AI research team to define what high quality materials reasoning looks like in practice and... 
    Hourly pay
    Full time
    Live in
    Relocation
    Relocation package

    SaidGig

    California
    1 day ago