Research Scientist, Agent Robustness
$216k - $270kScale
Scale Labs, Research Scientist — Agent Robustness As the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities and safeguarding AI models and systems. Building on this expertise, Scale Labs has launched a new team focused on policy research, to bridge the gap between AI research and global policymakers to make informed, scientific decisions about AI risks and capabilities. Our research tackles the hardest problems in agent robustness, AI control protocols, and AI risk evaluations to help governments, industry, and the public understand and mitigate AI risk while maximizing AI adoption. This team collaborates broadly across industry, the public sector, and academia and regularly publishes our findings. We are actively seeking talented researchers to join us in shaping this vision. As a Research Scientist working on Agent Robustness you will work on the fundamental challenges of building AI agents that are safe and aligned with humans. For example, you might: Research the science of AI agent capabilities with a focus on how they relate to safety, risk factors, and methodologies for benchmarking them; Design and build harnesses to test AI agents’ tendency to take harmful actions when pressured to do so by users or tricked into doing so by elements of their environment; Design and build exploits and mitigations for new and unique failure modes that arise as AI agents gain affordances like coding, web browsing, and computer use; Characterize and design mitigations for potential failure modes or broader risks of systems involving multiple interacting AI agents. Ideally you’d have: Commitment to our mission of promoting safe, secure, and trustworthy AI deployments in the industry as frontier AI capabilities continue to advance. Practical experience conducting technical research collaboratively. You should be comfortable building and leveraging agent scaffolding, designing evaluation harnesses, and quickly turning new ideas from the research literature into working prototypes. Experience with post‑training and RL techniques such as RLHF, DPO, GRPO, and similar approaches. A track record of published research in machine learning, particularly in generative AI. At least three years of experience addressing sophisticated ML problems, whether in a research setting or in product development. Strong written and verbal communication skills to operate in a cross‑functional team. Nice to have: Hands‑on experience with agent evaluation frameworks such as SWE‑bench, WebArena, OSWorld, Inspect, or similar tools. Experience with red‑teaming, prompt injection, or adversarial testing of AI systems. Our research interviews are crafted to assess candidates' skills in practical ML prototyping and debugging, their grasp of research concepts, and their alignment with our organizational culture. We will not ask any LeetCode‑style questions. If you’re excited about advancing AI safety and contributing to our mission, we encourage you to apply, even if your experience doesn’t perfectly align with every requirement. Compensation packages at Scale for eligible roles include base salary, equity, and benefits. The range displayed on each job posting reflects the minimum and maximum target for new hire salaries for the position and may be inclusive of several career levels at Scale; it will be determined during the interview process based on work location and additional factors, including job‑related skills, experience, qualifications, interview performance, and relevant education or training. Scale employees in eligible roles are also granted equity based compensation, subject to Board of Director approval. Your recruiter can share more about the specific salary range for your preferred location during the hiring process, and confirm whether the hired role will be eligible for equity grant. You'll also receive benefits including, but not limited to: comprehensive health, dental and vision coverage, retirement benefits, a learning and development stipend, and generous PTO. Additionally, this role may be eligible for additional benefits such as a commuter stipend. The base salary range for this full‑time position in the locations of San Francisco, New York, Seattle is:
$216,000 - $270,000 USD
PLEASE NOTE: Our policy requires a 90‑day waiting period before reconsidering candidates for the same role. This allows us to ensure a fair and thorough evaluation of all applicants. About Us: At Scale, our mission is to develop reliable AI systems for the world’s most important decisions. Our products provide the high‑quality data and full‑stack technologies that power the world’s leading models, and help enterprises and governments build, deploy, and oversee AI applications that deliver real impact. We work closely with industry leaders like Meta, Ernst & Young, Mayo Clinic, Time Inc., the Government of Qatar, and U.S. government agencies including the Army and Air Force. We are expanding our team to accelerate the development of AI applications. We believe that everyone should be able to bring their whole selves to work, which is why we are proud to be an inclusive and equal opportunity workplace. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability status, gender identity or Veteran status. We are committed to working with and providing reasonable accommodations to applicants with physical and mental disabilities. If you need assistance and/or a reasonable accommodation in the application or recruiting process due to a disability, please contact us at View email address on click.appcast.io. Please see the United States Department of Labor’s Know Your Rights poster for additional information. We comply with the United States Department of Labor’s Pay Transparency provision. PLEASE NOTE: We collect, retain and use personal data for our professional business purposes, including notifying you of job opportunities that may be of interest and sharing with our affiliates. We limit the personal data we collect to that which we believe is appropriate and necessary to manage applicants’ needs, provide our services, and comply with applicable laws. Any information we collect in connection with your application will be treated in accordance with our internal policies and programs designed to protect personal data. Please see our privacy policy for additional information. #J-18808-Ljbffr Scale$216k - $270k
Scale Labs, Research Scientist - Agent Robustness As the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities and safeguarding AI models and systems. Building on this expertise, Scale Labs has launched...SuggestedFull time- Scale Labs is seeking a Research Scientist focused on Agent Robustness to advance safe and aligned AI agents. You will contribute to evaluating risks, building testing harnesses, and prototyping mitigation strategies across agent interactions and environments. The role...Suggested
$218.4k - $273k
...and private evaluations.About the ACE team The Agent Capabilities & Environments (ACE) team, part of Scale’s Research organization, brings together customer-facing Researchers... ...real-world scenarios and environments, creating robust data programs to improve Large Language Models (...SuggestedFull time- Scale Labs seeks a Research Scientist to advance agent robustness, safety, and alignment in frontier AI systems. You will develop evaluation harnesses, prototype inspired by research, and collaborate across teams to push practical ML solutions. We prioritize candidates...Suggested
- Scale Labs seeks a Research Scientist focused on Agent Robustness in a multi-location setting. You will tackle fundamental challenges in building safe AI agents and aligning them with humans, including benchmarking methods and mitigation strategies. Ideal candidates have...Suggested
- Scale AI is seeking a candidate at the intersection of cutting-edge AI research and practical application to study data types essential for building state-of-the-art agents. You will explore data landscapes for intelligent, adaptable AI agents and guide Scale's data strategy...
- ...The organization prioritizes research in areas poised for impact including... ...partner with human scientists on the next generation of discovery... ...labor market. About the AI Agents pilot program Research into... ...this program will enable more robust and trustworthy agent design,...Local area
- Autonomous Technologies Group in New York, NY, seeks a senior researcher to drive original AI research at the boundaries of reasoning systems. You will explore new models, algorithms, and architectures for complex environments, with autonomy to pursue fundamental research...
- cursor is hiring a Research Scientist to drive research in reinforcement learning at their New York office. Candidates should have a deep background... ...for model training, and executing realtime RL for coding agents. This position offers significant scope and autonomy within a...Work at office
$302.4k - $378k
..., Jax, or Tensorflow. You should also be adept at interpreting research literature and quickly turning new ideas into prototypes. A track... ...-tuning projects using Pytorch/Jax. Hands-on experience with agent frameworks such as OpenHands, Swarm, LangGraph, etc....Full time- ...are human-aligned, hiring for potential and rewarding impactful research. The role sits in San Francisco or New York, full-time, with... ...experiments, and build simulations for behavioral intelligence agents. Expect to publish work, contribute to open-source projects, and...Full time
- Schmidt Sciences invites applications for a Program Scientist to lead the AI Agents pilot program, design new AI agents, and support cross‑cutting AI Center initiatives. You will work with grantmaking, research teams, and external advisers to advance responsible AI in...
$180.6k - $225.75k
...and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise... ...modes in frontier LLMs and Agents. You’ll identify everything from... ...capability gaps and reasoning errors to robustness and alignment issues, all...Full time$180.6k - $225.75k
...high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training... ...and propose solutions for bias mitigation and model robustness.Publish research findings in top-tier AI conferences...Full time$97.3k - $125.54k
...exempted from the federal civilian hiring freeze. As an FBI special agent, you'll directly impact national security. By harnessing your... ...to professional growth, a supportive work environment, and a robust benefits package that prioritizes you. Set yourself apart. Apply...Work at officeLocal area- ...INDEPENDENT FREIGHT AGENT R.B. Humphreys, Inc. — Rome, NY (Remote) ROLE OVERVIEW As an Independent Freight Agent with R.B. Humphreys, you will build and manage your own book of business while operating under our established authority. Backed by more than 60 years...Remote workRelocation package
- ...for our customers. Cohere is a team of researchers, engineers, designers, and more, who are... ...platform designed to securely deploy AI agents and automations within organizations' infrastructure... ...and ensure the execution environment is robust enough for the most demanding enterprise...Full timeWork at officeLocal areaRemote workHome officeFlexible hours
$58.5k - $73.36k
...has trained thousands of physicians and scientists who have helped to shape the course of medical... ...through medical education, scientific research, and direct patient care. At NYU Langone... ...package. Our offerings provide a robust support system for any stage of life, whether...$204k - $259k
...initiate and foster collaborations with other research teams in Alphabet. AI Foundations areas... ...inference, hierarchical learning, and robust evaluation. Role Summary In this hybrid role, you will report to a Principal Scientist. Responsibilities Participate in Waymo’s...Temporary workRemote work$65k - $120k
...has trained thousands of physicians and scientists who have helped to shape the course of medical... ...through medical education, scientific research, and direct patient care. At NYU Langone... ...package. Our offerings provide a robust support system for any stage of life, whether...- ...collaborative organization that puts human values first. About the Role Research scientists lead Basis’ efforts to develop a deeper understanding of the... ...scientists/engineers aspire to do rigorous, high-quality, robust science, but are not afraid to tinker, make mistakes, and...Full time
$200k - $320k
...should be. By integrating AI agents deeply into existing... ...to strengthen our expansive research team. We are looking for candidates... ...mathematical and quantitative skills, robust programming expertise, and... ...physicists, and computer scientists—who are passionate about...Work at officeLocal areaRelocation$65k - $85k
...has trained thousands of physicians and scientists who have helped to shape the course of medical... ...through medical education, scientific research, and direct patient care. At NYU Langone... ...package. Our offerings provide a robust support system for any stage of life, whether...- ...About the Role You'll drive original research at the boundaries of AI, working on new... ...deep learning, reinforcement learning, and agent-based AI. Develop novel architectures,... ..., JAX). Ability to turn theory into robust, practical code. Up-to-date on the latest...
- ...Basis is a nonprofit applied AI research organization with two... ...first. About the Role Research scientists lead Basis’ efforts to develop... ...to do rigorous, high-quality, robust science, but are not afraid to... ...might reliably detect if an agent possesses one, and crucially,...Full timeWork at officeImmediate start
$252k - $315k
...warehouses, internal APIs)Build robust data connectors and ETL... ...and configure AI models and agents within customer security and... ...collaborating with customer data scientists, ML engineers, and software developers... ...current on the latest AI/ML research and tools, bringing...Full time$93.75k - $133.2k
...hiring freeze. Use your STEM background to become an FBI special agent! The transition from physical sciences to special agent is more... ...commitment to professional growth, a supportive work environment, and a robust benefits package that prioritizes you.Set yourself apart. Apply...Work at officeLocal area- CB Talents SP is seeking a Customer Success Agent who can operate in German and Spanish (B2) to join our team in Madrid or Valencia. The... ...environment. The position emphasizes lively city settings, robust team collaboration, and a customer-first approach, with shifts aligning...Local areaShift work
$26 - $32 per hour
...As an Agent Experience Manager, you are the first person our customers meet when they join Compass and will be their account manager... ...opportunity employer, we offer competitive compensation packages, robust benefits and professional growth opportunities aimed at helping...Hourly payMinimum wageWork at officeFlexible hours- ...temporal intelligence. We are hiring a Research Scientist to architect and train proprietary foundation... ...architectures, and developing highly robust optical flow models tailored for the... ...how the next generation of embodied AI agents perceive and move through the physical...Shift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist, Agent Robustness. Be the first to apply!
- materials scientist New York, NY
- cosmetic scientist New York, NY
- scientist assay development New York, NY
- entry level research scientist New York, NY
- health scientist New York, NY
- quality control scientist New York, NY
- deep learning scientist New York, NY
- research associate scientist New York, NY
- application scientist New York, NY
- scientist antibody discovery New York, NY


