Research Scientist, Agent Robustness
Gravity Engineering Services Pvt Ltd.
About the Role As a Research Scientist working on Agent Robustness, you will work on the fundamental challenges of building AI agents that are safe and aligned with humans. Scale Labs has launched a new team focused on policy research, bridging the gap between AI research and global policymakers to make informed, scientific decisions about AI risks and capabilities. This team collaborates broadly across industry, the public sector, and academia and regularly publishes its findings. Responsibilities Research the science of AI agent capabilities with a focus on how they relate to safety, risk factors, and methodologies for benchmarking them. Design and build harnesses to test AI agents’ tendency to take harmful actions when pressured to do so by users or tricked into doing so by elements of their environment. Design and build exploits and mitigations for new and unique failure modes that arise as AI agents gain affordances like coding, web browsing, and computer use. Characterize and design mitigations for potential failure modes or broader risks of systems involving multiple interacting AI agents. Requirements Commitment to our mission: Promoting safe, secure, and trustworthy AI deployments as frontier AI capabilities continue to advance. Practical research experience: Comfortable building and leveraging agent scaffolding, designing evaluation harnesses, and quickly turning new ideas from the research literature into working prototypes. Experience with post-training and RL techniques: Such as RLHF, DPO, GRPO, and similar approaches. Track record of published research: In machine learning, particularly in generative AI. Experience: At least three years addressing sophisticated ML problems, whether in a research setting or in product development. Communication skills: Strong written and verbal communication skills to operate in a cross-functional team. Nice to Have Hands-on experience with agent evaluation frameworks such as SWE-bench, WebArena, OSWorld, Inspect, or similar tools. Experience with red-teaming, prompt injection, or adversarial testing of AI systems. #J-18808-Ljbffr Gravity Engineering Services Pvt Ltd.
- Scale Labs in San Francisco is seeking a Research Scientist focused on Agent Robustness to advance safe, aligned AI. You will tackle fundamental challenges in evaluating agent capabilities, safety, and risk, and help design benchmarks and protocols. You will design harnesses...Suggested
- Gravity Engineering Services Pvt Ltd. is seeking a Research Scientist focused on AI agent robustness in San Francisco. You will tackle the fundamental challenges of building safe AI agents by researching their capabilities and potential risks. Ideal candidates will have...Suggested
- A leading AI research organization in San Francisco is seeking a Research Scientist to work on agent robustness. The role involves conducting research on AI safety, designing tests for AI agents, and developing mitigations for potential risks. Candidates should have a...SuggestedFull time
$200.8k - $251k
Scale AI in San Francisco is seeking an AI Researcher focused on advancing intelligent agents through data strategy and research publications. This role requires... ...competitive salary range of $200,800 - $251,000, along with equity and robust benefits. #J-18808-Ljbffr Scale AISuggested$250k
Research Scientist / Engineer - Multimodal Agent SF Bay Area, CA • Remote, International • London, UK | Research Remote • Hybrid Full-time About Luma AI Luma... ...new tasks through data. Design, implement, and run robust data pipelines for constructing, enriching, and...SuggestedFull timeRemote workWorldwide$259.2k - $324k
...and private evaluations. About the ACE team The Agent Capabilities & Environments (ACE) team, part of Scale’s Research organization, brings together customer-facing... ...real-world scenarios and environments, creating robust data programs to improve Large Language Models (...Full time$170k - $220k
A leading AI safety nonprofit in San Francisco is seeking a Research Scientist to lead diverse projects aimed at enhancing AI safety. You will... ...directions in critical areas such as AI honesty and robustness while collaborating with leading academics. The ideal candidate...- MakerMaker.AI is seeking a researcher for autonomous systems focused on multi-agent machine learning. You will design methods and frameworks to improve agent capabilities and lead rigorous experiments in an open-ended research environment. The ideal candidate has over...
- Rip up the playbook and step into uncharted territory. If you've been building long-horizon multi-agent systems and pushing the boundaries of AI research, this is the kind of role where curiosity and ambition meet real execution, exploring truly novel problems at the frontier...
$400k
A dynamic technology company in San Francisco is seeking innovative individuals to push the boundaries of AI research. Candidates should have a PhD and experience in long-horizon reasoning and reinforcement learning. The role involves building systems to outperform existing...- cursor is hiring a Research Scientist to drive research in reinforcement learning at their New York office. Candidates should have a deep background... ...for model training, and executing realtime RL for coding agents. This position offers significant scope and autonomy within a...Work at office
$180k - $250k
Optimized, Inc. is seeking a research scientist in San Francisco to advance the intelligence of supply chain agents. Responsibilities include developing planning approaches, designing experiments, publishing findings, and collaborating with ML engineers. Candidates should...$180k - $250k
We're looking for a research scientist to advance the core intelligence of our supply chain agents. You'll work on the fundamental problems of multi-agent coordination, planning under uncertainty, and grounding LLM reasoning in real-world supply chain constraints, then...$250k - $300k
Role Overview As an Applied Research Engineer at Labelbox, you’ll sit at the junction of advanced AI research and real product impact, with a focus on the data that makes modern agents work—browser interactions, SWE/code traces, GUI sessions, and multi-turn workflows....Work at office3 days per week- A leading AI technology company is hiring AI Research Scientists and Engineers to enhance their AI products, focusing on state-of-the-art models, user experience, and deep research. Candidates should have strong backgrounds in large-scale LLMs, programming expertise in...
- Gravity Engineering Services Pvt Ltd. is looking for top-tier AI Research Scientists and Engineers to advance innovative AI products. Candidates will work on improving large-scale models that power AI experiences through various teams focused on research, implementation...
- Edison Scientific is seeking a Senior/Principal AI Research Scientist in San Francisco. This role focuses on developing models and algorithms to enhance scientific research. You will tackle significant research challenges and lead projects that drive innovation in the...
- ...TikTok AI, Google DeepMind, xAI, Microsoft Research, etc.), where we built large-scale... ...blog . Role overview As an AI Researcher, Agent, you will own the full research lifecycle... ...agentic RL—that produce models capable of robust multi‑step reasoning, tool use, and long‑...Full timeWork at office
$147.6k - $274k
...drug discovery and development. Roche’s Research and Early Development organisations at Genentech... ...Intelligence (AI) to assist our scientists in both pRED and gRED to deliver more innovative... ...Machine Learning Scientist building agents for applied small‑molecule drug design....Local areaWorldwideRelocation package- Gravity Engineering Services Pvt Ltd. is seeking a candidate to work at the forefront of AI research, focusing on reasoning within large language models (LLMs). You will play a crucial role in shaping data strategy and collaborate closely with engineering teams. The ideal...
$158k - $269k
...learn more visit: The Behaviors team at Waabi develops cutting-edge simulation agents and scenario generation algorithms for Waabi World, our simulation platform. As a Research Scientist on the Behaviors team, you will work closely with our multidisciplinary team of...Full timeWork at officeWork from homeFlexible hours$150k - $250k
...grasping and more dexterous behaviors in unstructured environments Research and implement state-of-the-art robot learning policies,... ...stack optimized for inference performance Design and maintain robust data collection and curation pipelines for production robot fleets...- ...the role We’re looking for a top‑tier Research Scientist to join our tech team. Your core responsibilities... ...retrieval and ranking systems for AI agents Prototype, train, and evaluate new... ...and multimodal understanding Build robust research pipelines, develop, deploy, and...Remote workFlexible hours
$200k - $320k
...should be. By integrating AI agents deeply into existing workflows... ...to strengthen our expansive research team. We are looking for candidates... ...and quantitative skills, robust programming expertise, and... ...theoretical physicists, and computer scientists—who are passionate about...Work at officeLocal areaRelocation- ...well known in the AI community for seminal research accomplishments at top AI labs, have run... ...a highly experienced AI Research Scientist to play a crucial role in the development... ...models for biomolecular design. Develop robust and scalable AI pipelines to process, analyze...
$97.3k - $125.54k
...Join to apply for the Special Agent role at Federal Bureau of Investigation (FBI) . 2 days ago Be among the first 25 applicants Federal... ...to professional growth, a supportive work environment, and a robust benefits package that prioritizes you. Set yourself apart. Apply...Work experience placementWork at officeLocal area$25 - $28 per hour
...Front Office Agent (Part-Time) Position Overview: The Front Desk Agent is responsible for processing all guest check-ins, check-outs,... ...Well and Do Good. The Barnes Hotel offers competitive salaries and robust benefit plans. Part Time Hourly Benefits: Part-time Medical...Hourly payPart timeLocal areaFlexible hoursWeekend work- Research Scientist, Real-Time Interactivity / Inference San Francisco · Research · Full Time Real... ...ll invent the methods that let people, agents, and robots drive video and world models... ...running on our stack and formulating a robust, generalizable solution Proposing a new...Full timeVisa sponsorshipRelocation package
- ...mission to make robots commonplace. The team is looking for a Research Scientist to help architect and deploy the machine learning models that... ...using imitation learning, RL, and VLA approaches Developing robust teleoperation and data collection frameworks Extending state...
- ...high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post‑training... ...and propose solutions for bias mitigation and model robustness. Publish research findings in top‑tier AI...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist, Agent Robustness. Be the first to apply!
- research associate scientist San Francisco, CA
- scientist antibody discovery San Francisco, CA
- protein scientist San Francisco, CA
- lead scientist San Francisco, CA
- applied sports scientist San Francisco, CA
- deep learning scientist San Francisco, CA
- lab scientist San Francisco, CA
- principal scientist San Francisco, CA
- molecular biology scientist San Francisco, CA
- senior analytical scientist San Francisco, CA

