Researcher, Agent Safety, Oversight and System Mitigations
AI Chopping Block
About the Team The Agent Safety team works to ensure that increasingly capable AI agents act safely, exercise sound judgment, and remain aligned with user intent. Our mission is to reduce the probability of severe unintended outcomes from increasingly capable AI agents while preserving their ability to act effectively and autonomously. Our work spans three areas: Training: Create training methods, environments and data that teach agents to make better decisions in consequential situations. We turn real-world failures into training signals that prevent similar incidents, and identify precursor behaviors and mitigations to address emerging risks. Measurements: Build evaluations and production metrics that identify emerging risks and measure whether our interventions work. Oversight : Develop oversight and system mitigation mechanisms that reduce harmful actions while preserving useful agent autonomy (for example future versions of auto-review). About the Role This role focuses on oversight and system-level mitigations that enable increasingly capable agents to operate safely and autonomously in real environments. We prioritize building oversight systems that are used in practice today, both internally and externally (see our recent work on action monitoring for codex and former code review). We also study longer-term questions about how increasingly capable agentis systems can be supervised, constrained, and corrected. We’re looking for a safety&security minded researcher or engineer who can reason rigorously about security boundaries and agent behavior, then build and test practical mitigations. A background in AI control or security is welcome but not required. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, build, and evaluate system-level controls for agent actions like agent-based review. Plan how they fit in a broader system including sandboxing with process isolation and permission boundaries. Work closely with a Codex harness engineering team to productionize the AI controls. Red-team end-to-end agentic systems to measure whether controls prevent data exfiltration, unsafe tool use, and other harmful outcomes. Improve the safety-productivity tradeoff by measuring and reducing missed harmful actions, unnecessary blocks, approval burden, and latency. You might thrive in this role if you: Have strong systems or security instincts and can reason concretely about isolation boundaries, permissions, attack surfaces, and failure modes in complex systems. Enjoy turning ambiguous safety questions into concrete threat models, reproducible experiments, and practical mitigations, and revising your approach based on evidence from deployment. Can build robust experimental infrastructure and design evaluations that distinguish promising mitigations from brittle ones. Are deeply interested in frontier AI alignment, safety and control. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s AffIrmative Action and Equal Employment Opportunity Policy Statement. Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non‑public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non‑compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link. OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology. #J-18808-Ljbffr AI Chopping Block
$52 per hour
...leader in security and threat mitigation. We specialize in risk... ...passionate about security and safety. Join our innovative team committed... ...such as executive protection agents, intelligence analysts, armed... ...Operating surveillance systems, alarms, and security technology...SuggestedHourly payWork at officeLocal areaFlexible hoursShift workNight shiftWeekend work- About the Company We're building autonomous research agents for recursive self-improvement (multi-agent systems that propose, run, and analyze machine learning experiments). We're a small team based in San Francisco, on-site. About the Role You'll be researching the agents...SuggestedShift work
- Senior AI Architect - Multi-Agent Systems & Platform Infrastructure Senior AI Architect - Multi-Agent Systems & Platform Infrastructure... ...matter • Hands-on coding ability (Python preferred), not just oversight • Strong architectural thinking paired with iterative, collaborative...SuggestedFull timeWork at officeRemote work
$150k - $250k
Applied AI Researcher, System Self-Construction Distyl | Distyl | Posted Mar 2 Apply Full-time Unknown About Distyl AI Distyl is an applied... ...components interact dynamically (e.g. graph-based planners, agent orchestration frameworks, workflow engines, or automated...SuggestedFull timeWork at office3 days per week- Distyl in San Francisco is seeking an Applied AI Researcher for the System Self-Construction team. You will design architectures enabling autonomous generation and refinement of sub-systems, pushing the frontier of self-constructing AI. Hybrid in-office collaboration is...SuggestedWork at office3 days per week
$380k
...tools that assist humans to agents that can plan, execute,... ...in the real world. Mitigating the frontier risks resulting... ...of frontier AI systems.Mitigation. Keeping misuse... ...are seeking exceptional researchers who can push the frontier of safety mitigations. You will help...Work at officeLocal areaFlexible hours$380k
About the TeamThe Safety Systems team is responsible for various safety work to ensure our best... ...and transparency.The Model Safety Research team aims to fundamentally advance our... ...identifying areas of risk and proposing mitigation strategies.You might thrive in this role...Work at officeLocal areaFlexible hours- ...that affect reliability and safety. Gridware’s advanced Active Grid... ...maintenance and fault mitigation. This comprehensive approach... ...Gridware's first dedicated design research role, and it is a foundational... ...users, while also mapping the systemic journeys that cut across...Local areaShift work
$62 per hour
...value that allows companies to prioritize safety. Our success is built on forging... ...Position Summary: The Residential Patrol Agent is responsible for providing high-level,... ...Monitor and report on CCTV and access control systems Assist with emergency evacuations or incidents...Hourly payFull timePart timeFlexible hoursShift workNight shift- Innovaccer in San Francisco, CA is seeking a Senior AI Researcher to design, develop, and deploy AI-powered... ...engineering, and data teams to build production-grade systems including LLM-based solutions, AI agents, and RAG workflows. The role is hands-on across the...
- Distyl AI is seeking researchers to pioneer AI-native operations, building across diverse AI system paradigms and deploying prototypes that redefine how software is used. You will work in a fast-moving, research-driven environment that values practical impact and real-...
- ...looking for a Reinforcement Learning Researcher with a strong product mindset to... ...research team. You'll develop RL systems that enable autonomous programming agents to learn and improve from real-... ...to measure agent performance and safety Collaborate with product teams to...
$23 - $23.5 per hour
...Office Agent Office Agents are responsible for helping customers during the shipping process... ...enter data into company's information system as required including international cargo... ...nights, weekends, and holidays. Safety, Security and Compliance: Take reasonable...Hourly payFull timePart timeWork experience placementWork at officeImmediate startFlexible hoursNight shift$22 - $28 per hour
...regulations set forth by Expeditors, with safety being the priority. Ensure that shipments... ...and advise supervisor of abuses Agents will closely work with the shift lead to... ...habits. Must be familiar with warehouse systems. Fork-lift certification may be required...Hourly payTemporary workPart timeLocal areaFlexible hoursShift work- OpenAI is seeking a highly capable professional for the Agent Post-Training team to advance frontier agent capabilities. You will design... ...translating findings into product improvements. The role spans research, engineering, data, and product work, with opportunities to own...
$23 - $23.5 per hour
...Description:Do you enjoy working in a fast-paced, safety-obsessed aviation environment?As a Warehouse Agent, you will be essential to increase operational efficiency... ...and enter data into company's information system as required.Follow company procedures and protocols...Full timePart timeWork experience placementWork at officeImmediate startFlexible hoursNight shift- ...Legacy carriers rely on static applications and disconnected systems Brokers chase carriers through calls, emails, and resubmissions... ..., no human intervention until the last mile. With Shepherd, safety, speed, and quality no longer trade off against one another —...Full timeWork experience placementWork at office
- ...Proximal is building the research systems needed to identify what models can’t yet do, build the... .... Our early team has built coding agents and RL infrastructure at companies like... ...problems end-to-end without heavy process or oversight Execution with excellence, even at high...
$150k - $220k
...About the Role You'll build and own the core agent infrastructure powering Doctours' AI systems—the orchestration layer, tool integrations, retrieval, and... ...fast self-directed learning. Personal projects, research, or open-source work building agent or LLM systems—orchestration...Full time$60 per hour
...a global leader in security and threat mitigation. We specialize in risk consulting, executive... ...passionate about security and safety. Join our innovative team committed to excellence... ...key roles, such as executive protection agents, intelligence analysts, armed security...Hourly payFlexible hoursShift work$248.4k - $310.5k
...ScaleScale's mission is to develop reliable AI systems for the world's most important... ...in production, paired with applied ML research, design, and evaluation to ensure these... ...About the RoleAs Engineering Manager for Agent Oversight, you'll lead the team building our platform...Full time$155k - $190k
...Healthcare That Remembers Every time you call a doctor's office, the system forgets you the moment you hang up. You wait on hold, explain... ...building one continuous conversation for every patient. An AI agent that knows who you are, remembers how you like to be reached,...Full timeWork at officeFlexible hours$250k - $325k
...first-to-market with: An AI agent that lives in MS Word and edits... ...: Why, What, and Who Why AI Researchers are the engine of innovation... ...Evaluate emerging work in agentic systems, long-context modeling, and... ...decoding), hallucination mitigation, novel architectures in deep...Contract workWork at officeImmediate startRemote work$380k
About the TeamThe Safety Training research team aims to fundamentally advance our capabilities for precisely implementing safe behavior in AI models... ...humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our...Work at officeLocal areaRemote workFlexible hours$75k - $155k
...organizations, we use our knowledge to advance safety and performance, set industry benchmarks,... .... Group also includes our strategic Research and Development unit, which provides... ...the safety and reliability of embodied AI systems - intelligent systems that interact with...Temporary workWork at officeFlexible hours- Lead Security Researcher Location: San Francisco, CA or Israel Type:... ...led Microsoft's vulnerability mitigation efforts and Tesla's offensive... ...platforms). You understand how these systems actually break. Builder... ...ways of putting LLMs or agents to work inside real engineering...Permanent employmentFull time
$150k - $220k
...Job Description Job Description About the Role You will own the core agent infrastructure that powers AI systems, including the orchestration layer, tool integrations, retrieval, and evaluation that support patient-facing and internal workflows. This is a systems...Full time- ...outreach clients. Assure optimal safety and efficiency of all... ..., and laboratory information system (LIS) functions. Involves... ...STANDARDS Personnel Oversight/Management # Direct supervision... ...through advanced biomedical research, graduate-level education in...TraineeshipWork experience placementLocal areaImmediate startWorldwide
- ...Corporate Security Agent/Emergency Response AgentCrisis24 is a global... ...this role:Ensure the overall safety and security of protectees/... ...employees.Monitoring security systems and technology tools for various... ...to proactively identify and mitigate threats.Detect and report suspicious...Local areaShift workNight shift
- ...Primarily On-site We are seeking a highly technical Enterprise Agent Systems Engineer to build and deploy advanced agent systems for... ...role combines strong software engineering with applied AI research judgment. You will work directly with technically sophisticated...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Researcher, Agent Safety, Oversight and System Mitigations. Be the first to apply!
- design researcher San Francisco, CA
- online researcher San Francisco, CA
- independent researcher San Francisco, CA
- product researcher San Francisco, CA
- survey researcher San Francisco, CA
- security researcher San Francisco, CA
- senior design researcher San Francisco, CA
- field researcher San Francisco, CA
- court researcher San Francisco, CA
- qualitative researcher San Francisco, CA





