Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Scientist - AI Self-Improvement & Evaluation

United States Digital Space LLC

the company is seeking a Research Scientist to advance measurable recursive-self-improvement in large models. You will design evaluations, build models of capability growth, and interpret results to guide R&D decisions. Senior roles exist and involve hands-on work alongside strategy. Candidates should have hands-on LLM research experience, strong quantitative instincts, and a track record in evaluating AI systems. #J-18808-Ljbffr United States Digital Space LLC

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Research Scientist - AI Self-Improvement & Evaluation in San Francisco, CA vacancy
  • Anthropic is seeking a Research Scientist to measure and understand recursive-self-improvement in large models. You will design evaluations and models, run experiments, and interpret results to guide research direction. We hire at junior and senior levels; seniors lead... 
    Suggested
    Work at office

    Anthropic Limited

    San Francisco, CA
    21 hours ago
  • $150k - $250k

    About Distyl AI Distyl is an applied AI technology...  ...organizations.We research and deploy technologies...  ...work spans research into self-constructing systems, the...  ...want to drive incremental improvements on benchmarks or...  ...architectures that continuously evaluate and enhance their own... 
    Suggested
    Work at office
    3 days per week

    Distyl AI

    San Francisco, CA
    1 day ago
  • $117.2k - $313.7k

     ...SalesforceSalesforce is the #1 AI CRM, where humans...  ...AI Research is looking for outstanding...  ...AI Research Scientists and Research Engineers...  ...workflows, self-evolving agent systems...  ...model distillation, evaluation, causal inference,...  ...and simulation to improve model accuracy and... 
    Suggested
    Full time

    Salesforce

    San Francisco, CA
    1 day ago
  • Carnaby Fox is seeking a Member of Technical Staff (AI Research) in San Francisco to help shape the research direction for frontier...  ...with world-class researchers to design experiments, evaluate LLMs, and improve data quality for high-stakes AI benchmarks. The role emphasizes... 
    Suggested

    Carnaby Fox

    San Francisco, CA
    4 days ago
  • $160k - $250k

    Overview Research Scientist - Mountain View, CA at Granica. This...  ...does Granica is an AI research and systems company...  ...systems to design self-optimizing data infrastructure...  ...that continuously improve how information is...  ...model architectures, evaluate on live datasets, and... 
    Suggested
    Flexible hours

    Granica

    San Francisco, CA
    2 days ago
  •  ...learn, adapt, and improve over time; have a...  ...record of exceptional research or engineering...  ...systems… About P-1 AI At P-1 AI, we are...  ...AI Research Scientist to join our small...  ...data generation to evaluation to product integration...  ...architectures, meta-learning, self-improvement, and... 
    Relocation package

    Namely

    San Francisco, CA
    1 day ago
  • $196k - $230k

     ...We AreNotion is the collaborative AI workspace where teams and agents think...  ...:We’re seeking an experienced UX Researcher to define and scale how we evaluate Notion’s AI-powered experiences—focusing...  ...break down and how they can improve. You’ll help teams spot regressions... 
    Local area
    Shift work

    Notion Labs

    San Francisco, CA
    1 day ago
  •  ...trustworthy and reliable AI systems, changing...  ...aspect of the research life cycle, from...  ...functional team of data scientists, software...  ...through training, evaluation, validation, and implementation...  ...to identify and improve the status quo....  ...optimization, self‑supervised... 
    Flexible hours

    Capital One

    San Francisco, CA
    3 days ago
  • $262.5k - $299.6k

     ...Applied Researcher II (AI Foundations, LLM Core and Agentic...  ...functional team of data scientists, software engineers,...  ...through training, evaluation, validation, and implementation...  ...to identify and improve the status quo. You’...  ...optimization, self‑supervised learning,... 
    Full time
    Part time
    Local area
    Flexible hours

    Capital One

    San Francisco, CA
    5 days ago
  • $180k - $260k

     ...looking for an Applied Scientist, AI to turn messy, high-...  ...models and AI systems that improve access to care and...  ...at the intersection of research, product, engineering,...  ..., design honest evaluations, run careful error analysis...  ...experiments end to end and self-serve deployments or... 
    Temporary work
    Work at office
    Monday to Friday
    Monday to Thursday

    Sprinter Health

    San Francisco, CA
    7 hours ago
  •  ...applications, processes, and AI into a single, governed platform...  ...balancing productivity with self-care. That’s why we offer all...  ...for an exceptional AI Research Scientist to join our growing team. In...  ...cost‑optimised RAG, tool‑use evaluation, and multi‑agent collaboration... 
    Remote work
    Flexible hours

    Workato

    San Francisco, CA
    2 days ago
  • $234.3k - $349k

     ...enterprises orchestrate AI-powered work. Our...  ...AI. About the roleAI research at WRITER isn't just about...  ...world. As an AI research scientist, you'll be at the...  ...through model training, evaluation, and production deploymentDesign...  ...— with a focus on improving multi-step reasoning,... 
    Full time
    Work at office
    Local area

    Writer

    San Francisco, CA
    3 days ago
  • Distyl in San Francisco is seeking an Applied AI Researcher for the System Self-Construction team. You will design architectures enabling autonomous generation and refinement of sub-systems, pushing the frontier of self-constructing AI. Hybrid in-office collaboration is... 
    Work at office
    3 days per week

    SupportFinity™

    San Francisco, CA
    5 days ago
  • CLERA in San Francisco, CA is seeking a Medical AI Researcher to bridge benchmark results with real-world reliability. You will own customer engagements, define evaluation questions, and deliver evidence to support FDA submissions. The role blends ML rigor with clinical... 

    CLERA

    San Francisco, CA
    5 days ago
  • $188k - $215k

     ...Collate   Collate is an AI document generation...  ...of Lever. Our AI researchers, engineers, and designers...  ...for an AI Research Scientist to push the boundaries...  ...production. Develop evaluation frameworks to measure...  ...simplicity of maintaining and improving real world AI... 

    Collate

    San Francisco, CA
    11 days ago
  •  ...site in the specified location(s).As an AI Researcher within Schwab’s AI Strategy &...  ...problem formulation through modeling, evaluation, deployment, and iteration, contributing...  ...evaluation and monitoring practices, and improving performance under real‑world constraints... 
    Full time
    Work at office

    The Charles Schwab Corporation

    San Francisco, CA
    7 hours ago
  • $216.3k - $280.8k

     ...received.Meet the TeamAt Foundation AI, we are leading frontier AI research across Cisco. Our mission is to...  ...systems, scalable training algorithms, evaluation science, inference optimization,...  ...evaluation benchmarks and frameworks that improve the reliability, transparency and... 
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Francisco, CA
    4 days ago
  • Anthropic is seeking an exceptional Research Scientist to join our Life Sciences team in San Francisco. This role focuses on improving AI models capabilities on scientific tasks. You will build bioinformatics tools and evaluation benchmarks while collaborating with product... 

    Anthropic

    San Francisco, CA
    21 hours ago
  • About the Role We’re looking for a Clinical Research Scientist to help lead and expand our work evaluating AI systems in mental health and other clinically sensitive...  ...Research in adolescent mental health, suicide or self-harm, psychosis, eating disorders, trauma, or... 
    Work experience placement
    Relocation package
    Shift work

    Vals AI

    San Francisco, CA
    5 days ago
  • $192.6k - $344.85k

     ...Research Lead / Principal Scientist & ManagerPost-Training...  ...LearningAutodesk AI Lab: London · San Francisco...  ...is still an open research problem.Autodesk touches...  ...tasks, and real-world evaluation grounded in...  ...novel algorithms that improve model reliability, controllability... 
    Full time
    For contractors
    Remote work

    Autodesk

    San Francisco, CA
    1 day ago
  • $216k - $270k

    Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities and safeguarding AI models and systems. Building on this expertise, Scale Labs... 
    Full time

    Scale AI

    San Francisco, CA
    2 days ago
  •  ...the Team The Proactivity Research team, within OpenAI’s...  ...technical foundations for AI that can anticipate what...  ...As a Research Engineer / Scientist, you will research and develop improvements to our models’...  ...learning, dataset creation, evaluations, and other post-training... 
    Work at office
    Relocation package
    Shift work

    Neura Market

    San Francisco, CA
    3 days ago
  • $192.6k - $344.85k

    ## AI Research Manager/Scientist, Reinforcement LearningApplylocations: San Francisco, CA, USA: AMER -...  ...scalable post-training workflows### ### Evaluation, Alignment & *Model* Quality* Design...  ...models demonstrate measurable improvements in reliability, alignment, and usefulness... 
    Remote work

    Autodesk, Inc.

    San Francisco, CA
    21 hours ago
  • $160k - $220k

     ...enjoy multi-year runway.About the RoleWe’re looking for an AI Research Scientist to advance the methodological frontier of AI in healthcare....  ...Your work may include novel architectures, new training or evaluation techniques, long-horizon research bets, peer-reviewed validation... 
    Temporary work
    Work at office
    Monday to Friday
    Monday to Thursday

    Sprinter Health

    San Francisco, CA
    3 days ago
  • $84.13 - $91.34 per hour

    AI Researcher - Efficient AI (Contractor) Step into the innovative world of LG Electronics....  ...into working prototypes and measurable improvements. You will have the opportunity to collaborate...  ...deployment workflows. • Propose and evaluate novel compression methods (PTQ, QAT,... 
    Full time
    Contract work
    Temporary work
    For contractors
    Local area
    Immediate start

    LG Electronics

    San Francisco, CA
    21 hours ago
  • $204k - $259k

     ...start as the Google Self‑Driving Car...  ...Experienced Driver—to improve access to mobility...  ...mission of the Waymo AI Foundations team...  ...collaborations with other research teams in Alphabet....  ..., and robust evaluation. Role Summary In...  ...to a Principal Scientist. Responsibilities... 
    Temporary work
    Remote work

    Waymo

    San Francisco, CA
    3 days ago
  • Ivo Inc. in San Francisco seeks an exceptional AI Researcher to push the state of the art in LLMs and deep...  ...text. You will develop novel techniques, build evaluative benchmarks, and ship production-ready solutions that improve accuracy, explainability, and robustness for... 

    Ivo Inc.

    San Francisco, CA
    21 hours ago
  •  ...that operationalizes responsible AI governance at scale. We're a 4...  ...Principal AI Security & Risk Researcher to join our founding research...  ...agentic systems Develop risk evaluation methodologies that adapt as threats...  ...fields Critical Attributes: Self-directed: You identify threats... 
    Part time
    Remote work
    Flexible hours

    Ciph Lab

    San Francisco, CA
    21 hours ago
  •  ...to help fine-tune large language models for medical reasoning. You will evaluate AI responses, design clinical scenarios across endocrine subspecialties, and provide structured feedback to improve AI systems for clinical and educational use. Requirements include MD or... 
    Remote job
    Contract work
    For contractors

    Turing

    San Francisco, CA
    2 days ago
  • $300k - $405k

     ...interpretable, and steerable AI systems. We want AI to be...  ...growing group of committed researchers, engineers, policy experts,...  ...designing and running capability evaluations against frontier models,...  ...identify gaps, and prioritize improvements Design and run red-teaming... 
    Visa sponsorship
    Shift work

    United States Digital Space LLC

    San Francisco, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Scientist - AI Self-Improvement & Evaluation. Be the first to apply!