Research Scientist - AI Self-Improvement & Evaluation
United States Digital Space LLC
the company is seeking a Research Scientist to advance measurable recursive-self-improvement in large models. You will design evaluations, build models of capability growth, and interpret results to guide R&D decisions. Senior roles exist and involve hands-on work alongside strategy. Candidates should have hands-on LLM research experience, strong quantitative instincts, and a track record in evaluating AI systems. #J-18808-Ljbffr United States Digital Space LLC
- Anthropic is seeking a Research Scientist to measure and understand recursive-self-improvement in large models. You will design evaluations and models, run experiments, and interpret results to guide research direction. We hire at junior and senior levels; seniors lead...SuggestedWork at office
$150k - $250k
About Distyl AI Distyl is an applied AI technology... ...organizations.We research and deploy technologies... ...work spans research into self-constructing systems, the... ...want to drive incremental improvements on benchmarks or... ...architectures that continuously evaluate and enhance their own...SuggestedWork at office3 days per week$117.2k - $313.7k
...SalesforceSalesforce is the #1 AI CRM, where humans... ...AI Research is looking for outstanding... ...AI Research Scientists and Research Engineers... ...workflows, self-evolving agent systems... ...model distillation, evaluation, causal inference,... ...and simulation to improve model accuracy and...SuggestedFull time- Carnaby Fox is seeking a Member of Technical Staff (AI Research) in San Francisco to help shape the research direction for frontier... ...with world-class researchers to design experiments, evaluate LLMs, and improve data quality for high-stakes AI benchmarks. The role emphasizes...Suggested
$160k - $250k
Overview Research Scientist - Mountain View, CA at Granica. This... ...does Granica is an AI research and systems company... ...systems to design self-optimizing data infrastructure... ...that continuously improve how information is... ...model architectures, evaluate on live datasets, and...SuggestedFlexible hours- ...learn, adapt, and improve over time; have a... ...record of exceptional research or engineering... ...systems… About P-1 AI At P-1 AI, we are... ...AI Research Scientist to join our small... ...data generation to evaluation to product integration... ...architectures, meta-learning, self-improvement, and...Relocation package
$196k - $230k
...We AreNotion is the collaborative AI workspace where teams and agents think... ...:We’re seeking an experienced UX Researcher to define and scale how we evaluate Notion’s AI-powered experiences—focusing... ...break down and how they can improve. You’ll help teams spot regressions...Local areaShift work- ...trustworthy and reliable AI systems, changing... ...aspect of the research life cycle, from... ...functional team of data scientists, software... ...through training, evaluation, validation, and implementation... ...to identify and improve the status quo.... ...optimization, self‑supervised...Flexible hours
$262.5k - $299.6k
...Applied Researcher II (AI Foundations, LLM Core and Agentic... ...functional team of data scientists, software engineers,... ...through training, evaluation, validation, and implementation... ...to identify and improve the status quo. You’... ...optimization, self‑supervised learning,...Full timePart timeLocal areaFlexible hours$180k - $260k
...looking for an Applied Scientist, AI to turn messy, high-... ...models and AI systems that improve access to care and... ...at the intersection of research, product, engineering,... ..., design honest evaluations, run careful error analysis... ...experiments end to end and self-serve deployments or...Temporary workWork at officeMonday to FridayMonday to Thursday- ...applications, processes, and AI into a single, governed platform... ...balancing productivity with self-care. That’s why we offer all... ...for an exceptional AI Research Scientist to join our growing team. In... ...cost‑optimised RAG, tool‑use evaluation, and multi‑agent collaboration...Remote workFlexible hours
$234.3k - $349k
...enterprises orchestrate AI-powered work. Our... ...AI. About the roleAI research at WRITER isn't just about... ...world. As an AI research scientist, you'll be at the... ...through model training, evaluation, and production deploymentDesign... ...— with a focus on improving multi-step reasoning,...Full timeWork at officeLocal area- Distyl in San Francisco is seeking an Applied AI Researcher for the System Self-Construction team. You will design architectures enabling autonomous generation and refinement of sub-systems, pushing the frontier of self-constructing AI. Hybrid in-office collaboration is...Work at office3 days per week
- CLERA in San Francisco, CA is seeking a Medical AI Researcher to bridge benchmark results with real-world reliability. You will own customer engagements, define evaluation questions, and deliver evidence to support FDA submissions. The role blends ML rigor with clinical...
$188k - $215k
...Collate Collate is an AI document generation... ...of Lever. Our AI researchers, engineers, and designers... ...for an AI Research Scientist to push the boundaries... ...production. Develop evaluation frameworks to measure... ...simplicity of maintaining and improving real world AI...- ...site in the specified location(s).As an AI Researcher within Schwab’s AI Strategy &... ...problem formulation through modeling, evaluation, deployment, and iteration, contributing... ...evaluation and monitoring practices, and improving performance under real‑world constraints...Full timeWork at office
$216.3k - $280.8k
...received.Meet the TeamAt Foundation AI, we are leading frontier AI research across Cisco. Our mission is to... ...systems, scalable training algorithms, evaluation science, inference optimization,... ...evaluation benchmarks and frameworks that improve the reliability, transparency and...Full timeTemporary workLocal areaFlexible hours- Anthropic is seeking an exceptional Research Scientist to join our Life Sciences team in San Francisco. This role focuses on improving AI models capabilities on scientific tasks. You will build bioinformatics tools and evaluation benchmarks while collaborating with product...
- About the Role We’re looking for a Clinical Research Scientist to help lead and expand our work evaluating AI systems in mental health and other clinically sensitive... ...Research in adolescent mental health, suicide or self-harm, psychosis, eating disorders, trauma, or...Work experience placementRelocation packageShift work
$192.6k - $344.85k
...Research Lead / Principal Scientist & ManagerPost-Training... ...LearningAutodesk AI Lab: London · San Francisco... ...is still an open research problem.Autodesk touches... ...tasks, and real-world evaluation grounded in... ...novel algorithms that improve model reliability, controllability...Full timeFor contractorsRemote work$216k - $270k
Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities and safeguarding AI models and systems. Building on this expertise, Scale Labs...Full time- ...the Team The Proactivity Research team, within OpenAI’s... ...technical foundations for AI that can anticipate what... ...As a Research Engineer / Scientist, you will research and develop improvements to our models’... ...learning, dataset creation, evaluations, and other post-training...Work at officeRelocation packageShift work
$192.6k - $344.85k
## AI Research Manager/Scientist, Reinforcement LearningApplylocations: San Francisco, CA, USA: AMER -... ...scalable post-training workflows### ### Evaluation, Alignment & *Model* Quality* Design... ...models demonstrate measurable improvements in reliability, alignment, and usefulness...Remote work$160k - $220k
...enjoy multi-year runway.About the RoleWe’re looking for an AI Research Scientist to advance the methodological frontier of AI in healthcare.... ...Your work may include novel architectures, new training or evaluation techniques, long-horizon research bets, peer-reviewed validation...Temporary workWork at officeMonday to FridayMonday to Thursday$84.13 - $91.34 per hour
AI Researcher - Efficient AI (Contractor) Step into the innovative world of LG Electronics.... ...into working prototypes and measurable improvements. You will have the opportunity to collaborate... ...deployment workflows. • Propose and evaluate novel compression methods (PTQ, QAT,...Full timeContract workTemporary workFor contractorsLocal areaImmediate start$204k - $259k
...start as the Google Self‑Driving Car... ...Experienced Driver—to improve access to mobility... ...mission of the Waymo AI Foundations team... ...collaborations with other research teams in Alphabet.... ..., and robust evaluation. Role Summary In... ...to a Principal Scientist. Responsibilities...Temporary workRemote work- Ivo Inc. in San Francisco seeks an exceptional AI Researcher to push the state of the art in LLMs and deep... ...text. You will develop novel techniques, build evaluative benchmarks, and ship production-ready solutions that improve accuracy, explainability, and robustness for...
- ...that operationalizes responsible AI governance at scale. We're a 4... ...Principal AI Security & Risk Researcher to join our founding research... ...agentic systems Develop risk evaluation methodologies that adapt as threats... ...fields Critical Attributes: Self-directed: You identify threats...Part timeRemote workFlexible hours
- ...to help fine-tune large language models for medical reasoning. You will evaluate AI responses, design clinical scenarios across endocrine subspecialties, and provide structured feedback to improve AI systems for clinical and educational use. Requirements include MD or...Remote jobContract workFor contractors
$300k - $405k
...interpretable, and steerable AI systems. We want AI to be... ...growing group of committed researchers, engineers, policy experts,... ...designing and running capability evaluations against frontier models,... ...identify gaps, and prioritize improvements Design and run red-teaming...Visa sponsorshipShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist - AI Self-Improvement & Evaluation. Be the first to apply!
- scientist ii San Francisco, CA
- scientist 1 San Francisco, CA
- image scientist San Francisco, CA
- downstream processing scientist San Francisco, CA
- qc scientist San Francisco, CA
- research scientist San Francisco, CA
- analytical scientist San Francisco, CA
- research scientist - biology San Francisco, CA
- genomics scientist San Francisco, CA
- research associate scientist San Francisco, CA


