Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Researcher, Agent Safety, Oversight and System Mitigations

AI Chopping Block

About the Team The Agent Safety team works to ensure that increasingly capable AI agents act safely, exercise sound judgment, and remain aligned with user intent. Our mission is to reduce the probability of severe unintended outcomes from increasingly capable AI agents while preserving their ability to act effectively and autonomously. Our work spans three areas: Training: Create training methods, environments and data that teach agents to make better decisions in consequential situations. We turn real-world failures into training signals that prevent similar incidents, and identify precursor behaviors and mitigations to address emerging risks. Measurements: Build evaluations and production metrics that identify emerging risks and measure whether our interventions work. Oversight : Develop oversight and system mitigation mechanisms that reduce harmful actions while preserving useful agent autonomy (for example future versions of auto-review). About the Role This role focuses on oversight and system-level mitigations that enable increasingly capable agents to operate safely and autonomously in real environments. We prioritize building oversight systems that are used in practice today, both internally and externally (see our recent work on action monitoring for codex and former code review). We also study longer-term questions about how increasingly capable agentis systems can be supervised, constrained, and corrected. We’re looking for a safety&security minded researcher or engineer who can reason rigorously about security boundaries and agent behavior, then build and test practical mitigations. A background in AI control or security is welcome but not required. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, build, and evaluate system-level controls for agent actions like agent-based review. Plan how they fit in a broader system including sandboxing with process isolation and permission boundaries. Work closely with a Codex harness engineering team to productionize the AI controls. Red-team end-to-end agentic systems to measure whether controls prevent data exfiltration, unsafe tool use, and other harmful outcomes. Improve the safety-productivity tradeoff by measuring and reducing missed harmful actions, unnecessary blocks, approval burden, and latency. You might thrive in this role if you: Have strong systems or security instincts and can reason concretely about isolation boundaries, permissions, attack surfaces, and failure modes in complex systems. Enjoy turning ambiguous safety questions into concrete threat models, reproducible experiments, and practical mitigations, and revising your approach based on evidence from deployment. Can build robust experimental infrastructure and design evaluations that distinguish promising mitigations from brittle ones. Are deeply interested in frontier AI alignment, safety and control. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s AffIrmative Action and Equal Employment Opportunity Policy Statement. Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non‑public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non‑compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link. OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology. #J-18808-Ljbffr AI Chopping Block

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Researcher, Agent Safety, Oversight and System Mitigations in San Francisco, CA vacancy
  • $52 per hour

     ...leader in security and threat mitigation. We specialize in risk...  ...passionate about security and safety. Join our innovative team committed...  ...such as executive protection agents, intelligence analysts, armed...  ...Operating surveillance systems, alarms, and security technology... 
    Suggested
    Hourly pay
    Work at office
    Local area
    Flexible hours
    Shift work
    Night shift
    Weekend work

    Enhanced Protection Services

    San Francisco, CA
    2 days ago
  • About the Company We're building autonomous research agents for recursive self-improvement (multi-agent systems that propose, run, and analyze machine learning experiments). We're a small team based in San Francisco, on-site. About the Role You'll be researching the agents... 
    Suggested
    Shift work

    MakerMaker.AI

    San Francisco, CA
    6 hours ago
  • Senior AI Architect - Multi-Agent Systems & Platform Infrastructure Senior AI Architect - Multi-Agent Systems & Platform Infrastructure...  ...matter • Hands-on coding ability (Python preferred), not just oversight • Strong architectural thinking paired with iterative, collaborative... 
    Suggested
    Full time
    Work at office
    Remote work

    Nivalto

    San Francisco, CA
    2 days ago
  • $150k - $250k

    Applied AI Researcher, System Self-Construction Distyl | Distyl | Posted Mar 2 Apply Full-time Unknown About Distyl AI Distyl is an applied...  ...components interact dynamically (e.g. graph-based planners, agent orchestration frameworks, workflow engines, or automated... 
    Suggested
    Full time
    Work at office
    3 days per week

    Distyl

    San Francisco, CA
    3 days ago
  • Distyl in San Francisco is seeking an Applied AI Researcher for the System Self-Construction team. You will design architectures enabling autonomous generation and refinement of sub-systems, pushing the frontier of self-constructing AI. Hybrid in-office collaboration is... 
    Suggested
    Work at office
    3 days per week

    SupportFinity™

    San Francisco, CA
    3 days ago
  • $380k

     ...tools that assist humans to agents that can plan, execute,...  ...in the real world. Mitigating the frontier risks resulting...  ...of frontier AI systems.Mitigation. Keeping misuse...  ...are seeking exceptional researchers who can push the frontier of safety mitigations. You will help... 
    Work at office
    Local area
    Flexible hours

    OpenAI

    San Francisco, CA
    4 days ago
  • $380k

    About the TeamThe Safety Systems team is responsible for various safety work to ensure our best...  ...and transparency.The Model Safety Research team aims to fundamentally advance our...  ...identifying areas of risk and proposing mitigation strategies.You might thrive in this role... 
    Work at office
    Local area
    Flexible hours

    OpenAI

    San Francisco, CA
    4 days ago
  •  ...that affect reliability and safety. Gridware’s advanced Active Grid...  ...maintenance and fault mitigation. This comprehensive approach...  ...Gridware's first dedicated design research role, and it is a foundational...  ...users, while also mapping the systemic journeys that cut across... 
    Local area
    Shift work

    Gridware

    San Francisco, CA
    12 days ago
  • $62 per hour

     ...value that allows companies to prioritize safety. Our success is built on forging...  ...Position Summary: The Residential Patrol Agent is responsible for providing high-level,...  ...Monitor and report on CCTV and access control systems Assist with emergency evacuations or incidents... 
    Hourly pay
    Full time
    Part time
    Flexible hours
    Shift work
    Night shift

    Accomplished Security Inc.

    San Francisco, CA
    1 day ago
  • Innovaccer in San Francisco, CA is seeking a Senior AI Researcher to design, develop, and deploy AI-powered...  ...engineering, and data teams to build production-grade systems including LLM-based solutions, AI agents, and RAG workflows. The role is hands-on across the... 

    Innovaccer

    San Francisco, CA
    3 days ago
  • Distyl AI is seeking researchers to pioneer AI-native operations, building across diverse AI system paradigms and deploying prototypes that redefine how software is used. You will work in a fast-moving, research-driven environment that values practical impact and real-... 

    Distyl

    San Francisco, CA
    1 day ago
  •  ...looking for a Reinforcement Learning Researcher with a strong product mindset to...  ...research team. You'll develop RL systems that enable autonomous programming agents to learn and improve from real-...  ...to measure agent performance and safety Collaborate with product teams to... 

    Autohand AI Ltd.

    San Francisco, CA
    3 days ago
  • $23 - $23.5 per hour

     ...Office Agent Office Agents are responsible for helping customers during the shipping process...  ...enter data into company's information system as required including international cargo...  ...nights, weekends, and holidays. Safety, Security and Compliance: Take reasonable... 
    Hourly pay
    Full time
    Part time
    Work experience placement
    Work at office
    Immediate start
    Flexible hours
    Night shift

    AGI

    San Francisco, CA
    1 day ago
  • $22 - $28 per hour

     ...regulations set forth by Expeditors, with safety being the priority. Ensure that shipments...  ...and advise supervisor of abuses Agents will closely work with the shift lead to...  ...habits. Must be familiar with warehouse systems. Fork-lift certification may be required... 
    Hourly pay
    Temporary work
    Part time
    Local area
    Flexible hours
    Shift work

    Expeditors

    Brisbane, CA
    3 days ago
  • OpenAI is seeking a highly capable professional for the Agent Post-Training team to advance frontier agent capabilities. You will design...  ...translating findings into product improvements. The role spans research, engineering, data, and product work, with opportunities to own... 

    OpenAI

    San Francisco, CA
    3 days ago
  • $23 - $23.5 per hour

     ...Description:Do you enjoy working in a fast-paced, safety-obsessed aviation environment?As a Warehouse Agent, you will be essential to increase operational efficiency...  ...and enter data into company's information system as required.Follow company procedures and protocols... 
    Full time
    Part time
    Work experience placement
    Work at office
    Immediate start
    Flexible hours
    Night shift

    lliance Ground International

    San Francisco, CA
    2 days ago
  •  ...Legacy carriers rely on static applications and disconnected systems Brokers chase carriers through calls, emails, and resubmissions...  ..., no human intervention until the last mile. With Shepherd, safety, speed, and quality no longer trade off against one another —... 
    Full time
    Work experience placement
    Work at office

    Shepherd

    San Francisco, CA
    more than 2 months ago
  •  ...Proximal is building the research systems needed to identify what models can’t yet do, build the...  .... Our early team has built coding agents and RL infrastructure at companies like...  ...problems end-to-end without heavy process or oversight Execution with excellence, even at high... 

    Proximal LLC

    San Francisco, CA
    2 days ago
  • $150k - $220k

     ...About the Role You'll build and own the core agent infrastructure powering Doctours' AI systems—the orchestration layer, tool integrations, retrieval, and...  ...fast self-directed learning. Personal projects, research, or open-source work building agent or LLM systems—orchestration... 
    Full time

    Doctours

    San Francisco, CA
    9 days ago
  • $60 per hour

     ...a global leader in security and threat mitigation. We specialize in risk consulting, executive...  ...passionate about security and safety. Join our innovative team committed to excellence...  ...key roles, such as executive protection agents, intelligence analysts, armed security... 
    Hourly pay
    Flexible hours
    Shift work

    Allied Universal

    San Francisco, CA
    3 days ago
  • $248.4k - $310.5k

     ...ScaleScale's mission is to develop reliable AI systems for the world's most important...  ...in production, paired with applied ML research, design, and evaluation to ensure these...  ...About the RoleAs Engineering Manager for Agent Oversight, you'll lead the team building our platform... 
    Full time

    Scale AI

    San Francisco, CA
    4 days ago
  • $155k - $190k

     ...Healthcare That Remembers Every time you call a doctor's office, the system forgets you the moment you hang up. You wait on hold, explain...  ...building one continuous conversation for every patient. An AI agent that knows who you are, remembers how you like to be reached,... 
    Full time
    Work at office
    Flexible hours

    Assort Health

    San Francisco, CA
    2 days ago
  • $250k - $325k

     ...first-to-market with: An AI agent that lives in MS Word and edits...  ...: Why, What, and Who Why AI Researchers are the engine of innovation...  ...Evaluate emerging work in agentic systems, long-context modeling, and...  ...decoding), hallucination mitigation, novel architectures in deep... 
    Contract work
    Work at office
    Immediate start
    Remote work

    Ivo Inc.

    San Francisco, CA
    3 days ago
  • $380k

    About the TeamThe Safety Training research team aims to fundamentally advance our capabilities for precisely implementing safe behavior in AI models...  ...humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our... 
    Work at office
    Local area
    Remote work
    Flexible hours

    OpenAI

    San Francisco, CA
    3 days ago
  • $75k - $155k

     ...organizations, we use our knowledge to advance safety and performance, set industry benchmarks,...  .... Group also includes our strategic Research and Development unit, which provides...  ...the safety and reliability of embodied AI systems - intelligent systems that interact with... 
    Temporary work
    Work at office
    Flexible hours

    DNV

    Oakland, CA
    3 days ago
  • Lead Security Researcher Location: San Francisco, CA or Israel Type:...  ...led Microsoft's vulnerability mitigation efforts and Tesla's offensive...  ...platforms). You understand how these systems actually break. Builder...  ...ways of putting LLMs or agents to work inside real engineering... 
    Permanent employment
    Full time

    PI Security LLC

    San Francisco, CA
    3 days ago
  • $150k - $220k

     ...Job Description Job Description About the Role You will own the core agent infrastructure that powers AI systems, including the orchestration layer, tool integrations, retrieval, and evaluation that support patient-facing and internal workflows. This is a systems... 
    Full time

    Clera

    San Francisco, CA
    7 days ago
  •  ...outreach clients. Assure optimal safety and efficiency of all...  ..., and laboratory information system (LIS) functions. Involves...  ...STANDARDS Personnel Oversight/Management # Direct supervision...  ...through advanced biomedical research, graduate-level education in... 
    Traineeship
    Work experience placement
    Local area
    Immediate start
    Worldwide

    University of California , San Francisco

    San Francisco, CA
    2 days ago
  •  ...Corporate Security Agent/Emergency Response AgentCrisis24 is a global...  ...this role:Ensure the overall safety and security of protectees/...  ...employees.Monitoring security systems and technology tools for various...  ...to proactively identify and mitigate threats.Detect and report suspicious... 
    Local area
    Shift work
    Night shift

    Crisis24

    San Francisco, CA
    3 days ago
  •  ...Primarily On-site We are seeking a highly technical Enterprise Agent Systems Engineer to build and deploy advanced agent systems for...  ...role combines strong software engineering with applied AI research judgment. You will work directly with technically sophisticated... 

    MaxIT Consulting - Max Corporate Group

    San Francisco, CA
    25 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Researcher, Agent Safety, Oversight and System Mitigations. Be the first to apply!