Research Scientist - AI Capabilities & Safety
$250kMETR
A nonprofit research organization in Berkeley is seeking a Research Engineer to help develop scientific methods for assessing AI capabilities and risks. The ideal candidate will have a strong research background, excellent execution skills, and experience in software engineering. The role offers a salary range of $250,000 to $450,000 annually and includes benefits like unlimited PTO, relocation support, and a collaborative work environment focused on high-quality research. #J-18808-Ljbffr METR
- Anthropic in San Francisco is seeking a Research Scientist to join the Alignment team, focusing on keeping frontier AI systems safe, helpful, and honest as capabilities scale. You will design empirical experiments, develop novel alignment methods like RLHF and Constitutional...Suggested
- A non-profit AI research organization based in Berkeley, CA, is looking for experienced research scientists to lead and accelerate AI alignment research. The role allows for publishing... ...for remote work. Ideal for those keen to innovate in AI safety. #J-18808-Ljbffr AisafetySuggestedRemote jobFull time
- Astera Institute is seeking a Senior Research Scientist to develop fundamental theories of intelligence. The role demands rigorous mathematical skills and the ability to engage with complex models and data across multiple fields. The ideal candidate will contribute to...Suggested
- Scale Labs in San Francisco is seeking a Research Scientist focused on Agent Robustness to advance safe, aligned AI. You will tackle fundamental challenges in evaluating agent capabilities, safety, and risk, and help design benchmarks and protocols. You will design harnesses...Suggested
$140k - $200k
Research Engineer & Scientist The Center for AI Safety (CAIS) is a leading research and advocacy organization focused on mitigating societal-scale risks from... ...24 , we turned to malicious use and weaponization capabilities, introducing the first state-of-the-art benchmarks...SuggestedWork at officeLocal area$300k - $405k
...interpretable, and steerable AI systems. We want AI to be safe... ...growing group of committed researchers, engineers, policy experts, and... ...engineers to build the safety and oversight mechanisms that... ...time: designing and running capability evaluations against frontier...Work at officeVisa sponsorshipFlexible hoursShift work- Anthropic is seeking a Capabilities Researcher to join the team building Claude Security. You will identify security capabilities in frontier models, measure performance, and ensure usability for non-experts. The role blends research with product impact, offering latitude...
- Anthropic is seeking a Capabilities Researcher to join the Claude Security team in San Francisco. You will identify which security capabilities... ...research role on a product team offers broad latitude to explore AI security capabilities, design rigorous evaluations, and help...
$293k - $405k
About the team Preparedness is a critical Safety Research team at OpenAI, which is focused on mitigating AI threats to global security that could scale to an extreme... .... Monitoring and predicting the evolving capabilities of frontier AI systems. Mitigation. Keeping misuse...- United States Digital Space LLC is seeking a Capabilities Researcher to join the team building Claude Security. You will identify security capabilities in frontier models, measure performance, and determine how to make them usable for non-experts. In this research role...
- the company is an AI research and deployment company dedicated to ensuring that general-purpose... .... We push the boundaries of the capabilities of AI systems and seek to safely deploy... ...powerful tool that must be created with safety and human needs at its core, and to achieve...
- United States Digital Space LLC is seeking a Capabilities Researcher to join the team building Claude Security. You will identify security capabilities in frontier models, measure performance, and determine how to make them useful to non-experts. This research role sits...
$207k - $300k
...practical experience. 3 years of research experience in frontier AI safety or governance (e.g., in threat modeling... ..., safety frameworks, or dangerous capability/alignment evaluations). Preferred... ...The Job We are hiring a research scientist to lead governance research and...Shift work- Anthropic is seeking a Research Scientist to measure and understand recursive-self-improvement in large models. You will design evaluations and models, run experiments, and interpret results to guide research direction. We hire at junior and senior levels; seniors lead...Work at office
$150k
...Master's vs PhD in AI (2026): Which Is Right for You? A comprehensive... ...looks like SOC 15-1221-style research, SOC 15-1252-style software... ...you\'re building. Research scientists who publish papers, advance... ...work at AI labs creating new capabilities need a PhD. ML engineers,...Full timeWork at office- OpenAI is seeking an experienced security researcher to help mitigate AI threats and safeguard systems as AI agents become more capable. The role focuses on designing robust defenses and coordinating with teams to maintain safeguards across our platform. You will identify...
- Scale Labs in San Francisco seeks a Research Scientist focused on Safety Post-Training to advance post-training methods and interpretability for frontier AI systems. You will design pipelines, evaluate safety properties, and help translate findings into practical guidelines...
- Scale Labs is seeking a Research Scientist focused on Agent Robustness to advance safe and aligned AI agents. You will contribute to evaluating risks, building testing harnesses, and prototyping mitigation strategies across agent interactions and environments. The role...
- ...Francisco is building a team to monitor and control AI systems, with emphasis on real‑time observability and safety. You will prototype monitoring, containment... ...standards for responsible AI oversight. The role seeks researchers with a track record in ML, especially generative...
- A leading AI research organization in San Francisco is seeking a Research Scientist to work on agent robustness. The role involves conducting research on AI safety, designing tests for AI agents, and developing mitigations for potential risks. Candidates should have a Ph...Full time
$216k - $270k
Scale Labs, Research Scientist — Safety Post TrainingAs the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities and safeguarding AI models and systems. Building on this expertise, Scale Labs has launched...Full time- ...About Us FAR.AI is a non-profit AI research institute dedicated to ensuring advanced AI is safe and beneficial... ...is to facilitate breakthrough AI safety research, advance global... ...Staff, with significant overlap between scientist and engineer roles. As a scientist, you...Full timeRemote workVisa sponsorship
$234.3k - $349k
...leading enterprises orchestrate AI-powered work. Our vision is... ...with AI. About the roleAI research at WRITER isn't just about publishing... ...the world. As an AI research scientist, you'll be at the center of... ..., and the system-level capabilities that make AI genuinely useful...Full timeWork at officeLocal area$150k - $275k
...Research Scientist, AI Substrate is addressing one of the most important technological problems facing the United States. At the intersection... ...and modeling while simultaneously building internal AI capabilities across the organization. This role sits at the...Local area- OpenAI is at the center of high-impact multimodal AI. The Chat and Multimodal Safety team builds safe, scalable models and evaluations for text, vision, and audio tasks. As a Researcher on the Chat and Multimodal Safety team in San Francisco, you will shape model perception...Work at officeRelocation package
- About the Team Preparedness is a critical Safety Research team at the company, which is focused on mitigating AI threats to global security that could scale to an... ...Measurement. Monitoring and predicting the evolving capabilities of frontier AI systems. Mitigation. Keeping...Permanent employmentTemporary work
- OpenAI in San Francisco seeks exceptional researchers to push the frontier of safety mitigations, helping derisk frontier models and advance techniques from... ...role requires deep technical experience, 2+ years in AI safety and a PhD or equivalent, plus 4+ years of research...
- About the team Preparedness is a critical Safety Research team at the company, which is focused on mitigating AI threats to global security that could scale to an... ...Measurement. Monitoring and predicting the evolving capabilities of frontier AI systems. Mitigation. Keeping...
- United States Digital Space LLC is seeking a Researcher for cybersecurity risks to design and implement an end-to-end mitigation stack across products, advancing safeguards as AI models evolve. You will work with cross-functional teams to span prevention, monitoring, detection...
- OpenAI is looking for a Senior Researcher focused on AI safety, located in San Francisco. This role involves setting research directions for ensuring safe AGI, developing AI monitor models, and collaborating with cross-functional teams to uphold safety standards. The ideal...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist - AI Capabilities & Safety. Be the first to apply!
- molecular biology scientist Berkeley, CA
- water quality scientist Berkeley, CA
- machine learning scientist Berkeley, CA
- scientist antibody discovery Berkeley, CA
- machine learning research scientist Berkeley, CA
- materials scientist Berkeley, CA
- health scientist Berkeley, CA
- scientist Berkeley, CA
- quality control scientist Berkeley, CA
- scientist biology Berkeley, CA

