Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Scientist - Frontier Benchmarks

$200k - $325k

Snorkel AI

About Snorkel At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes between 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler! About the Role We're looking for a Research Scientist to collaborate with partners and lead the development of the next frontier benchmarks and datasets. This is a highly visible, customer-facing role at the intersection of research, company strategy, and go-to-market. You'll design datasets taking into account frontier model performance and work with our academic partners, and then partner with delivery, product and go-to-market to scale out production. You will also serve as a credible technical partner for our customers, prospects, and drive results that impact the broader research community. This role reports directly to the Head of Research and is ideal for someone who is energized by cross-functional work and wants to understand how startups operate across research, data operations, and commercial teams. Main Responsibilities Design state of the art datasets that drive frontier model training and evaluation based on current model performance and academic partnerships Translate benchmark insights into clear, compelling narratives that articulate the ROI of expert-curated data for customer-facing presentations, technical reports, and go-to-market materials. Work cross-functionally with data operations, product, engineering, and strategy to surface research findings that inform the company roadmap. Stay at the frontier of LLM evaluation research and bring best practices into Snorkel's workflows Represent Snorkel's research externally through publications, blog posts, conference talks, and customer engagements that advance the conversation around data-centric AI Preferred Qualifications Strong research background in AI/ML evaluation, NLP, or related fields, with a track record of rigorous experimental design — especially around measuring the impact of training and evaluation data on model behavior. Exceptional communication skills — able to present complex technical findings clearly to both technical and non-technical audiences Comfort operating in a fast-moving, cross-functional environment with ambiguous problem spaces Genuine interest in GTM strategy, startup dynamics, and the commercial side of AI data services. Ph.D. in machine learning, NLP, or a related field preferred; equivalent industry or research lab experience considered. Salary Range

$200,000 - $325,000 USD

Growth Opportunities Joining Snorkel AI means becoming part of a company that has market-proven solutions, robust funding, and is scaling rapidly—offering a unique combination of stability and the excitement of high growth. As a member of our team, you’ll have meaningful opportunities to shape priorities and initiatives, influence key strategic decisions, and directly impact our ongoing success. Whether you’re looking to deepen your technical expertise, explore leadership opportunities, or learn new skills across multiple functions, you’re fully supported in building your career in an environment designed for growth, learning, and shared success. Equal Employment Opportunity Statement Snorkel AI is proud to be an Equal Employment Opportunity employer and is committed to building a team that represents a variety of backgrounds, perspectives, and skills. Snorkel AI embraces diversity and provides equal employment opportunities to all employees and applicants for employment. Snorkel AI prohibits discrimination and harassment of any type on the basis of race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state, or local law. All employment is decided on the basis of qualifications, performance, merit, and business need. We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation. #J-18808-Ljbffr Snorkel AI

Vacancy posted 16 hours ago
Similar jobs that could be interesting for youBased on the Research Scientist - Frontier Benchmarks in San Francisco, CA vacancy
  • $216k - $270k

    Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner for frontier AI companies, Scale plays an integral...  ...team.Nice to have:Experience in crafting evaluations and benchmarks, or a background in data science roles related to LLM... 
    Suggested
    Full time

    Scale AI

    San Francisco, CA
    2 days ago
  • $250k

     ...data and evaluation infrastructure that frontier AI labs use to make their models better....  ...rigorous evaluations that go beyond static benchmarks. We are a small, early team (post Series...  ...and measured. Working directly with research teams at top AI labs, you’ll experiment... 
    Suggested

    AfterQuery

    San Francisco, CA
    2 days ago
  •  ...profound global impact. About the Role Frontier AI is moving toward scientific...  ...for this next era. We are looking for a Research Scientist who can help define Quantum AI: not just...  ...propose, test, and refine hypotheses. Benchmarking frameworks that reveal when a new computational... 
    Suggested
    Casual work
    Visa sponsorship

    Sygaldry

    San Francisco, CA
    4 days ago
  • $150k - $250k

     ...rearchitect critical operations for the frontier of AI. Our customers include the largest...  ...goods, and global social organizations.We research and deploy technologies that power AI-...  ...want to drive incremental improvements on benchmarks or optimize an existing process but... 
    Suggested
    Work at office
    3 days per week

    Distyl AI

    San Francisco, CA
    1 day ago
  •  ...Researcher Position at Hedra Hedra is building a world-class Physical AI research team...  ...researchers who are excited to go beyond benchmarks and build models that operate in the...  ...research into production Stay at the frontier of the field — synthesizing relevant literature... 
    Suggested
    Work at office

    HEDRA INC

    San Francisco, CA
    5 days ago
  • $160k - $220k

     ...year runway.About the RoleWe’re looking for an AI Research Scientist to advance the methodological frontier of AI in healthcare. This role is ideal for someone...  ...healthcare validation standards are higher than benchmark culture alone, and you are energized by the opportunity... 
    Temporary work
    Work at office
    Monday to Friday
    Monday to Thursday

    Sprinter Health

    San Francisco, CA
    3 days ago
  • Velvet is a data research company building the datasets that power the next...  ...quality audiovisual training data for frontier labs. We’re hiring a Research Scientist to develop and fine‑tune models...  ...Build evaluation frameworks and benchmarks to rigorously measure enhancement... 
    Immediate start
    Shift work

    Velvet

    San Francisco, CA
    1 day ago
  • $234.3k - $349k

     ...work with AI. About the roleAI research at WRITER isn't just about...  ...world. As an AI research scientist, you'll be at the center of...  ...workflowsBuild novel evaluation benchmarks and methodologies that push...  ...representing WRITER at the frontier of the field and contributing... 
    Full time
    Work at office
    Local area

    Writer

    San Francisco, CA
    3 days ago
  • The role As a research scientist, you will design, implement, and optimize the large-scale training infrastructure that powers our frontier reinforcement learning stack. This is systems work at...  ...partners like Kleiner Perkins, Benchmark, Sequoia, Lux, and Greenoaks. Who... 
    Work at office
    Visa sponsorship
    Relocation package

    Applied Compute

    San Francisco, CA
    1 day ago
  • $250k - $400k

     ...stealth AI start-up building frontier reasoning models for...  ...genuinely novel AI for Science research, combining frontier reasoning...  ...Building evaluation frameworks and benchmarks for complex, multi-step...  ...Researcher or experienced Research Scientist. What matters most is hands‑... 

    techire ai

    San Francisco, CA
    1 day ago
  • $125k - $225k

    FutureSearch is looking for exceptional Research Scientists to evaluate and improve state-of-the-...  ...of engineers and researchers at the frontier of AI epistemics. We have the best publicly...  ...ICLR Workshop paper in Jan 2026, and benchmarks like Deep Research Bench and Bench to... 
    Remote work
    Flexible hours

    Futuresearch

    San Francisco, CA
    2 days ago
  • $140k - $200k

     ...breakthrough AI models at leading research labs and enterprises. Since...  ...integrated solutions for frontier AI development: Enterprise...  ...not a traditional research scientist role. You will not spend months...  ...with evaluation and benchmarking of LLMs — designing metrics,... 
    Work at office
    Flexible hours
    2 days per week

    Labelbox

    San Francisco, CA
    16 hours ago
  • $300k - $320k

     ...growing group of committed researchers, engineers, policy experts,...  ...seeking an exceptional Research Scientist to join our Life Sciences...  ...model training objectives, benchmarks, and agentic workflows. You’...  ...accelerated biology while shaping how frontier models reason about and... 
    Visa sponsorship

    Anthropic

    San Francisco, CA
    3 days ago
  •  ...enterprises. We aim to push the frontier of AI that understands real,...  ...role is for an experienced scientist who thrives both in...  ...and deep content extraction. Research, evaluate, and integrate the...  ...product impact. Develop new benchmarks, datasets, and evaluation methodologies... 

    Tensorlake Inc.

    San Francisco, CA
    16 hours ago
  • $295k

     ...at OpenAI, and is guided by OpenAI's Preparedness Framework. Frontier AI models have the potential to benefit all of humanity, but also...  ...productivity can also accelerate exploitation. As a Researcher for cybersecurity risks, you will help design and implement an... 

    Slope

    San Francisco, CA
    4 days ago
  • DataAnnotation is seeking a Clinical Data Scientist for a remote contract role to evaluate AI-generated quantitative analyses and create benchmark problems for training AI systems. You will assess AI outputs, develop training problems across forecasting, experiment design... 
    Remote job
    Contract work

    DataAnnotation

    San Francisco, CA
    5 days ago
  • Wheel the World seeks a full-time ML Research Scientist in San Francisco to advance generative AI and quantum computing. This role involves developing theoretical and practical implementations for quantum acceleration in generative models. Ideal candidates should have... 
    Full time
    Visa sponsorship

    Wheel the World

    San Francisco, CA
    4 days ago
  •  ...and evaluation infrastructure that frontier AI labs use to improve their...  ...evaluations that go beyond static benchmarks. We're a small, early team (post-...  ...re building out our post-training research team and hiring 2-3 Research Scientists to work together on this mission.... 
    Internship
    Shift work

    Product Pulse

    San Francisco, CA
    1 day ago
  • $380k

    Research Scientist - Multimodal Agent, Consumer Devices | OpenAI Careers Research Scientist -...  ...the future of computing. We work at the frontier of multimodal AI, helping turn...  ...‑grounded: success is not just higher benchmark performance, but better model behaviour... 
    Work at office
    Immediate start
    Relocation package

    OpenAI

    San Francisco, CA
    5 days ago
  • $216k - $270k

    Scale Labs, Research Scientist - AI Controls and Monitoring As the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the...  ...researchers to establish standards and benchmarks for AI monitoring and escalation. Qualifications... 
    Full time

    Scale AI, Inc.

    San Francisco, CA
    16 hours ago
  • $218.4k - $273k

     ...round, we’re accelerating the abundance of frontier data to pave the road to Artificial...  ...Environments (ACE) team, part of Scale’s Research organization, brings together customer-...  ...environments and RL reward signals, benchmarking autonomous agent performance across real... 
    Full time

    Scale AI

    San Francisco, CA
    1 day ago
  •  ...the Role We’re looking for a Clinical Research Scientist to help lead and expand our work evaluating...  ...to build clinically grounded benchmarks, realistic multi-turn scenarios, scoring...  ...model labs are continuously pushing the frontier. The unicorn companies that will emerge... 
    Work experience placement
    Relocation package
    Shift work

    Vals AI

    San Francisco, CA
    5 days ago
  • Enam, Inc. is seeking a senior ML researcher to lead exploration of new neural network foundations and architectures, with a strong emphasis on LLM applications. You will design experiments, validate novel hypotheses, and translate research ideas into scalable systems... 

    Enam, Inc.

    San Francisco, CA
    2 days ago
  •  ...datasets. This is a rare intersection of frontier AI and real-world scientific impact....  ...mode. The Role We’re looking for research scientists who want to work at the intersection of...  ...Evaluation: Contributing to meaningful benchmarks and evaluation methods for domain-specific... 

    Xterraai

    San Francisco, CA
    1 day ago
  • $140k - $200k

    Research Engineer & Scientist The Center for AI Safety (CAIS) is a leading research and advocacy organization...  ...the first state-of-the-art benchmarks for measuring it. More recently, we...  ...regularly used by AI safety institutes and frontier AI labs, and they have shaped real... 
    Work at office
    Local area

    Center for Ai Safety

    San Francisco, CA
    3 days ago
  • $300k

    Research Scientist — Frontier World Models & RL A stealth, exceptionally well-backed applied AI lab is hiring Research Scientists to solve open problems at the frontier of world models and RL. Three weeks from public beta, 100M+ views before launch, backed by names you'... 
    Remote work
    Visa sponsorship
    Relocation package

    Harnham

    San Francisco, CA
    5 days ago
  •  ...workersResponsibilitiesWe are looking for an exceptional AI Research Scientist to join our growing team. In this role, you will be responsible...  ...use evaluation, and multi‑agent collaboration.Prototype and benchmark models; present findings internally and externally.... 
    Remote work
    Flexible hours

    Workato

    San Francisco, CA
    2 days ago
  •  ...Granica’s mission is to remove that inefficiency. We combine new research in information theory , probabilistic modeling , and...  ...focus on unstructured text or media, we are exploring the next frontier: systems that understand and reason over the information that runs... 
    Flexible hours

    Granica Computing, Inc.

    San Francisco, CA
    5 days ago
  • $225k - $300k

    Research Scientist About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access, and those...  ...small group of researchers and engineers focused on pushing the frontier while shipping real systems into production. We are a small... 
    Work at office
    Immediate start

    Latent

    San Francisco, CA
    5 days ago
  •  ...compounds. Accelerate change - Ship fast, adapt faster, and move frontier ideas into production. Create win-wins - Creatively turn trade...  ...fail. But succeed an unfair amount. Job: Our first dedicated research hire - you will answer the question: how to train and scale a... 

    Parallel Web Systems

    San Francisco, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Scientist - Frontier Benchmarks. Be the first to apply!