Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Agent Safety & Robustness Research Scientist

Scale AI

A leading AI research organization in San Francisco is seeking a Research Scientist to work on agent robustness. The role involves conducting research on AI safety, designing tests for AI agents, and developing mitigations for potential risks. Candidates should have a Ph.D. and a strong background in machine learning, with at least three years of experience in sophisticated ML problems. This full-time position offers a competitive salary and comprehensive benefits including health coverage, retirement plans, and generous PTO. #J-18808-Ljbffr Scale AI

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the AI Agent Safety & Robustness Research Scientist in San Francisco, CA vacancy
  • Scale Labs in San Francisco is seeking a Research Scientist focused on Agent Robustness to advance safe, aligned AI. You will tackle fundamental challenges in evaluating agent capabilities, safety, and risk, and help design benchmarks and protocols. You will design harnesses... 
    Suggested

    Scale AI, Inc.

    San Francisco, CA
    2 days ago
  • $216k - $270k

    Scale Labs, Research Scientist - Agent Robustness As the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities...  ...with a focus on how they relate to safety, risk factors, and methodologies for benchmarking... 
    Suggested
    Full time

    Scale AI, Inc.

    San Francisco, CA
    2 days ago
  • $200.8k - $251k

     ...Scale AI in San Francisco is seeking an AI Researcher focused on advancing intelligent agents through data strategy and research publications. This role requires expertise in...  ...salary range of $200,800 – $251,000, along with equity and robust benefits.#J-18808-Ljbffr... 
    Suggested

    Scale AI

    San Francisco, CA
    15 hours ago
  •  ...Scale AI is seeking a Senior/Staff Machine Learning Research Scientist on the Agents team in Seattle. You will advance practical ML systems by prototyping ideas from literature and turning them into deployable prototypes, with a strong emphasis on LLMs and agent frameworks... 
    Suggested

    Scale

    San Francisco, CA
    15 hours ago
  • $218.4k - $273k

    About ScaleAt Scale AI, our mission is to accelerate the development...  ....About the ACE team The Agent Capabilities & Environments (ACE) team, part of Scale’s Research organization, brings together customer...  ...and environments, creating robust data programs to improve Large... 
    Suggested
    Full time

    Scale AI

    San Francisco, CA
    2 days ago
  • Edison Scientific is seeking a Senior/Principal AI Research Scientist in San Francisco. This role focuses on developing models and algorithms to enhance scientific research. You will tackle significant research challenges and lead projects that drive innovation in the field... 

    Edison Scientific

    San Francisco, CA
    2 days ago
  • $140k - $200k

    Research Engineer & Scientist The Center for AI Safety (CAIS) is a leading research and advocacy organization focused on mitigating societal-scale risks from...  ...them first. In 2022-2023 , we focused on AI honesty, robustness, transparency, and trojan/backdoor behaviors. In 2... 
    Work at office
    Local area

    Center for Ai Safety

    San Francisco, CA
    4 days ago
  • $259.2k - $324k

     ...About Scale At Scale AI, our mission is to accelerate the development...  .... About the ACE team The Agent Capabilities & Environments (ACE) team, part of Scale’s Research organization, brings together...  ...and environments, creating robust data programs to improve Large... 
    Full time

    Scale AI, Inc.

    San Francisco, CA
    23 hours ago
  • Atom Matrix Innovation LLC seeks aTechnical PM to own the AI agent platform roadmap, shaping instructions, tools, knowledge, guardrails...  ...You will partner with AI and platform engineers to define scope, safety, and sequencing, turning high-stakes problems into tangible... 

    Atom Matrix Innovation LLC

    San Francisco, CA
    4 days ago
  • OpenAI is seeking a Researcher for cybersecurity risks to design and implement an end-to-end mitigation stack mitigating severe cyber misuse...  ...with cross-functional teams to ensure safeguards are robust, scalable, and effective as models and attacker tactics evolve.... 

    Neura Market

    San Francisco, CA
    2 days ago
  •  ...in San Francisco seeks exceptional researchers to push the frontier of safety mitigations, helping derisk frontier...  ...techniques from interpretability, robustness and alignment to ensure safe deployments...  ...technical experience, 2+ years in AI safety and a PhD or equivalent,... 

    Triwill Group

    San Francisco, CA
    3 days ago
  • $300k - $405k

     ...interpretable, and steerable AI systems. We want AI to...  ...group of committed researchers, engineers, policy...  ...engineers to build the safety and oversight mechanisms...  ...safeguards need to be robust against sophisticated actors...  ..., such as select agent regulations, the Biological... 
    Work at office
    Visa sponsorship
    Flexible hours
    Shift work

    Anthropic

    San Francisco, CA
    2 days ago
  • OpenAI is seeking an experienced security researcher to help mitigate AI threats and safeguard systems as AI agents become more capable. The role focuses on designing robust defenses and coordinating with teams to maintain safeguards across our platform. You will identify... 

    Neura Market

    San Francisco, CA
    4 days ago
  • Anthropic in San Francisco is seeking a Research Scientist to join the Alignment team, focusing on keeping frontier AI systems safe, helpful, and honest as capabilities scale. You will design empirical experiments, develop novel alignment methods like RLHF and Constitutional... 

    AI Breaking Wire

    San Francisco, CA
    2 days ago
  •  ...Francisco is building a team to monitor and control AI systems, with emphasis on real‑time observability and safety. You will prototype monitoring, containment...  ...standards for responsible AI oversight. The role seeks researchers with a track record in ML, especially generative... 

    Scale AI, Inc.

    San Francisco, CA
    2 days ago
  • Scale Labs in San Francisco seeks a Research Scientist focused on Safety Post-Training to advance post-training methods and interpretability for frontier AI systems. You will design pipelines, evaluate safety properties, and help translate findings into practical guidelines... 

    Scale AI, Inc.

    San Francisco, CA
    22 hours ago
  • $216k - $270k

    Scale Labs, Research Scientist — Safety Post TrainingAs the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities...  ...tackles the hardest problems in agent robustness, AI control protocols, and AI risk... 
    Full time

    Scale AI

    San Francisco, CA
    3 days ago
  • $320k

     ...interpretable, and steerable AI systems. We want AI to...  ...group of committed researchers, engineers, policy...  ...researching and ensuring safety with self-improving, highly...  ...signs that LLMs and agents are increasingly capable...  ...surface. As a Research Scientist on FRT focusing on cyber... 
    Work at office
    Relocation
    Visa sponsorship
    Flexible hours

    Neura Market

    San Francisco, CA
    2 days ago
  •  ...Team The Computer-Using Agent team is responsible for developing...  ...long-horizon reliability, safety and more. Combining rigorous research with high-quality...  ...designing and maintaining robust and secure systems that facilitate...  ...deep experience building AI infrastructure and who are... 
    Full time
    Work at office
    Relocation package

    OpenAI

    San Francisco, CA
    10 hours ago
  •  ...About the Job We’re seeking an Agent Engineer to design and build agentic features in our...  ...passionate about building agent systems and making AI easy for developers to adopt. The ideal...  ...and monitoring systems Ensure APIs are robust, developer-friendly, and enterprise-ready... 
    Full time
    Worldwide
    Flexible hours

    FriendliAI

    San Francisco, CA
    10 hours ago
  •  ...emerging trillion-dollar Voice AI economy, providing real-time...  ...building production-grade voice agents at scale. More than 200,000...  ...working alongside Deepgram’s core research teams. We’ll be working...  ...focus areas include: Ultra-robust ASR in noisy, multi-mic environments... 
    Full time
    Home office
    Flexible hours

    Deepgram

    San Francisco, CA
    10 hours ago
  • $230k - $315k

     ...democratizing access to cutting-edge AI innovation to enable any...  ....SuperApp is our flagship AI agent workspace. It unifies communication...  ...Engineering: Engineer robust system-level guardrails to detect...  ...status, or disability status. Research shows that in order to apply for... 
    Work at office
    Flexible hours

    Instabase

    San Francisco, CA
    1 day ago
  • $124k - $218.3k

     ...leading enterprises orchestrate AI-powered work. Our vision is to...  ...are building and deploying AI agents that are grounded in their...  ...frameworks.Develop and enhance robust infrastructure and high-throughput...  ...with product managers, AI researchers, data engineers, and UX teams... 
    Full time
    Work at office
    Local area

    Writer

    San Francisco, CA
    10 hours ago
  •  ...About the Team OpenAI's research training infrastructure...  ...and evaluated. The Agent Harness Bridge team sits...  ...research; it is building robust infrastructure that...  ...OpenAI OpenAI is an AI research and deployment...  ...that must be created with safety and human needs at its... 
    Full time

    OpenAI

    San Francisco, CA
    10 hours ago
  • OpenAI is at the center of high-impact multimodal AI. The Chat and Multimodal Safety team builds safe, scalable models and evaluations for text, vision, and audio tasks. As a Researcher on the Chat and Multimodal Safety team in San Francisco, you will shape model perception... 
    Work at office
    Relocation package

    United States Digital Space LLC

    San Francisco, CA
    4 days ago
  • About the Team Preparedness is a critical Safety Research team at the company, which is focused on mitigating AI threats to global security that could scale to an extreme...  ...the ability to think outside the box and have a robust “red-teaming mindset” Have experience in ML... 
    Permanent employment
    Temporary work

    United States Digital Space LLC

    San Francisco, CA
    4 days ago
  • $293k - $405k

    About the team Preparedness is a critical Safety Research team at OpenAI, which is focused on mitigating AI threats to global security that could scale to an extreme...  ...company and for society. About the role As AI agents become more capable at software engineering, and... 

    Slope

    San Francisco, CA
    2 days ago
  •  ...team Preparedness is a critical Safety Research team at the company, which is focused on mitigating AI threats to global security...  ...from tools that assist humans to agents that can plan, execute, and...  ...building protections that remain robust as products, model... 

    United States Digital Space LLC

    San Francisco, CA
    4 days ago
  • the company is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of...  ...products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we... 

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  • United States Digital Space LLC is seeking a Researcher for cybersecurity risks to design and implement an end-to-end mitigation stack across products, advancing safeguards as AI models evolve. You will work with cross-functional teams to span prevention, monitoring, detection... 

    United States Digital Space LLC

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Agent Safety & Robustness Research Scientist. Be the first to apply!