Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Researcher, Recursive Self-Improvement Safety

$295k

OpenAI

About the teamModels are becoming increasingly capable—moving from tools that assist humans to agents that can plan, execute, and adapt in the real world. Mitigating the frontier risks resulting from these capabilities is paramount to OpenAI’s ability to continue deploying models safely.The Preparedness team is dedicated to addressing these critical risks. Our work includes:Measurement. Monitoring and predicting the evolving capabilities of frontier AI systems.Mitigation. Keeping misalignment safeguards, alignment tools, and on track to adequately address extreme threats that might arise in the future.Coordination. Setting mitigation targets by maintaining OpenAI’s preparedness framework, and partnering with other staff to achieve these targets.This is urgent, fast-paced work that has far-reaching implications for the company and for society.About the rolePreparedness is hiring strong technical executors to support preparations for accelerated AI development, which may culminate in recursive self-improvement. This work relies on anticipating misalignment risks that might exist in the future, but might not exist now; so it’s especially important that people in this role are tasteful and strategic.The role is wide-ranging, covering any mitigation for loss of control risk, spanning the design and implementation of better pre-deployment risk-assessment, control measures, RSI-relevant training interventions, and turning one’s technical work into established institutional practices and external-facing communications.Below is a subset of our focus areas:Scalable oversight: Establishing practices for model misbehavior monitoring and oversight which remain effective in superhuman model capability regimes, with a focus on bridging from today’s monitoring approaches to future-proof ones.Automated auditing: As model capabilities increase, we’ll increasingly rely on automated approaches for finding the most severe forms of model misalignments. We’ll both need to sift through large swaths of production traffic to find the most egregious misalignments, and reliably elicit tail risks before deployment.Rigorous monitorability: Rigorous testing and red-teaming of our measurements of model misbehavior related to loss-of-control (e.g. reward hacking, sandbagging, scheming). This includes better understanding monitorability, and e.g. preparing for potential losses of Chain-of-Thought monitorability.Model behavior science: Design experiments and evaluations to understand the extent to which models are problematically misaligned, or their safety-relevant capabilities lag behind dangerous capabilities. This may include training model organisms of misbehavior for behaviors not currently present in production, or training interventions to increase safety-relevant capabilities.Coordination and verification: Prototype technical mechanisms for verifying compliance with potential future AI safety agreements.AI R&D risk measurement: Track progress toward automation of technical staff to inform OpenAI’s near-term investments in alignment and security.Maintaining and strengthening RSI safety cases: We’re especially interested in identifying and addressing blindspots of mitigation areas which we may have missed.Generally, our team alternates between performing rigorous hypothesis-driven research and turning our insights into interventions or control systems which impact production models, with occasional support of engineering teams.In this role, you will:Carefully consider the problems OpenAI might face in the future and how to prepare for them.Turn an open-ended objective like “prepare for future misalignment threats” into a much more concrete direction (e.g. “stress-test monitors for scheming”) – prioritizing the work that is most useful to start right now.Execute quickly, building scrappy prototypes, and then improving them iteratively until they become established components of our safety pipelines.Secure buy-in from other staff at OpenAI when necessary, and communicate your work clearly.Collaborate with or manage other staff as needed, since we might need to rapidly scale to tackle these problems quickly.You might thrive in this role if you:Are an exceptional technical executor.Have strong strategic and research taste: you can prioritize effectively in domains with weak feedback loops.Are passionate about mitigating the risks associated with recursive self-improvement.Are driven by a desire to do whatever work most positively impacts the future of AI development.Bonus: you have already done work in one of the domains listed above (ML research, AI alignment, AI verification etc).About OpenAIOpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Researcher, Recursive Self-Improvement Safety in San Francisco, CA vacancy
  • $150k - $250k

     ...goods, and global social organizations.We research and deploy technologies that power AI-...  ...Distyl itself. Our work spans research into self-constructing systems, the development of...  ...don’t just want to drive incremental improvements on benchmarks or optimize an existing... 
    Suggested
    Work at office
    3 days per week

    Distyl AI

    San Francisco, CA
    4 days ago
  • $295k

    About the team The Safety Systems org ensures that OpenAI's most capable models can be...  ...Teaming (ART) effort: building scalable, research-driven systems that continuously discover...  ...into actionable, production-facing improvements. The goal is to maximize counterfactual... 
    Suggested

    OpenAI

    San Francisco, CA
    1 day ago
  •  ...growing team comprised of quantitative researchers, software engineers, product managers, designers...  ...costs, continuously researching improvements to our optimization formulations and ensuring...  ...in an object oriented language. A self-starter who embraces ownership and... 
    Suggested
    Work at office
    Visa sponsorship
    Flexible hours

    Frec

    San Francisco, CA
    3 days ago
  • $218.7k - $249.6k

     ...life. Our work touches every aspect of the research life cycle, from partnering with Academia...  ...work with stakeholders to identify and improve the status quo. You’re passionate about talent...  ...of the following: training optimization, self‑supervised learning, robustness,... 
    Suggested
    Full time
    Part time
    Local area
    Flexible hours

    Capital One

    San Francisco, CA
    21 hours ago
  •  ...life. Our work touches every aspect of the research life cycle, from partnering with academia...  ...work with stakeholders to identify and improve the status quo. You’re passionate about talent...  ...of the following: training optimization, self‑supervised learning, robustness,... 
    Suggested
    Flexible hours

    Capital One

    San Francisco, CA
    1 day ago
  • $262.5k - $299.6k

     ...Applied Researcher II (AI Foundations, LLM Core and Agentic AI) Overview At Capital One, we...  ...and work with stakeholders to identify and improve the status quo. You’re passionate about talent...  ...of the following: training optimization, self‑supervised learning, robustness,... 
    Full time
    Part time
    Local area
    Flexible hours

    Capital One

    San Francisco, CA
    3 days ago
  •  ...advance. We believe this can be dramatically improved. At Axiom, we generate and curate...  .... We are looking for a machine learning researcher to help define and build the core AI systems...  ...systems. Work on contrastive learning, self-supervised learning, semi-supervised... 

    axiombio

    San Francisco, CA
    3 days ago
  •  ...of the grid that affect reliability and safety. Gridware’s advanced Active Grid Response...  ...mitigation. This comprehensive approach helps improve safety, reduce outages, and ensure the...  ...is Gridware's first dedicated design research role, and it is a foundational hire for our... 
    Local area
    Shift work

    Gridware

    San Francisco, CA
    22 days ago
  • $162.7k - $263.18k

     ...across our enterprise customers’ networks.As a Principal Security Researcher, you will play a key technical leadership role in shaping how...  ..., exploit, and attack technique areas where new or improved protections are needed.Drive innovative detection ideas from concept... 
    Full time
    Work at office

    Palo Alto Networks

    San Francisco, CA
    4 days ago
  •  ...this role to work on site in the specified location(s).As an AI Researcher within Schwab’s AI Strategy & Transformation (AI.x)...  ...establishing robust evaluation and monitoring practices, and improving performance under real‑world constraints. You’ll collaborate closely... 
    Full time
    Work at office

    The Charles Schwab Corporation

    San Francisco, CA
    3 days ago
  • $142.7k - $270.95k

     ...novices or experienced engineers, on a user-friendly platform improved by groundbreaking AI capabilities. The AI Foundations team sits...  ...optimization, and production deployment.Collaborate closely with Adobe Research, engineering, and product teams to bring AI-powered features to... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    2 days ago
  •  ...connected app. We've helped millions of people understand and improve their health by providing daily insights and practical steps to...  ...out of the office. We're looking for a Senior Human Factors Researcher to join our Design Research team. This role is ideal for a researcher... 
    Work at office
    Local area
    Remote work
    Flexible hours
    2 days per week
    3 days per week

    Oura

    San Francisco, CA
    4 days ago
  • $139.6k - $225.78k

     ...relationships, and the kind of precision that drives great outcomes.Job SummaryJob SummaryWe are seeking a passionate and self-driven Sr. Staff Researcher to join our Cloud-Delivered Security Services team. In this role, you will be pivotal in developing and refining the... 
    Full time
    Work experience placement
    Work at office
    Visa sponsorship
    Work visa

    Palo Alto Networks

    San Francisco, CA
    3 days ago
  • $96.7k - $126.4k

    Conduct independent and collaborative research using human data, human biospecimens, and induced pluripotent stem cell (iPSC)-based models...  ...aimed at understanding cardiovascular health disparities and improving health equity in understudied populations.Demonstrated success... 

    University of California, San Francisco

    San Francisco, CA
    1 day ago
  •  ...Security Researcher We believe that software is the foundation of modern civilization - yet vulnerabilities threaten its integrity,...  ...engineers to understand limitations and design new methodologies to improve our system Publish internal technical reports and... 
    Full time
    Work at office

    DepthFirst

    San Francisco, CA
    4 days ago
  • $150k

     ...apps out of POCs and into production. We eliminate the risk and improve the reliability of LLM apps by haizing them – i.e. rigorously,...  ...proactively, and continuously fuzz-testing them. We are looking for Research Engineers to help develop our reliability platform, with a... 
    Visa sponsorship

    Enboarder

    San Francisco, CA
    21 hours ago
  •  ...hardware). Train and tune policies, curate data, and iterate to improve real‑world performance. Write production‑quality code that...  ...understand use cases and observe robots in deployment contexts. Bridge research and operations: translate research advances into deployable... 

    Physical Intelligence

    San Francisco, CA
    1 day ago
  • $96.7k - $126.4k

    Assistant Professional Researcher position available in the Dept. of Laboratory Medicine, University of California, San FranciscoAvailability...  ...goals. Candidates should have a strong interest in directly improving patient care through their research and gaining experience in... 
    Traineeship

    University of California, San Francisco

    San Francisco, CA
    3 days ago
  • Tilde Research is a moonshot AI lab advancing mechanistic interpretability, new architectures, and pretraining science. We build foundational...  ...and engineers to advance interpretability as a tool for improving performance and control. What you might work on: Designing,... 

    Tilde Research

    San Francisco, CA
    3 days ago
  • $285k - $380k

     ...quickly growing group of committed researchers, engineers, policy experts,...  ...science at a leading AI safety company. Responsibilities...  ...psychometric techniques to analyze and improve interviewer calibration and...  ...~ Experience building self-service analytics tools or dashboards... 
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    14 days ago
  •  .... Our work is at the cutting edge of AI research, focusing on developing methodologies that...  ...of our approach are: (1) harnessing improved capabilities into alignment, making sure...  ...powerful tool that must be created with safety and human needs at its core, and to achieve... 
    Work at office
    Relocation package

    OpenAI

    San Francisco, CA
    4 days ago
  •  ...Foundation AI, we are leading frontier AI research across Cisco. Our mission is to advance...  ...AI infrastructure. You will drive improvements to training algorithms, curate and optimize...  ...systems for security-focused use cases, safety, and robust machine learning #J-18808-Ljbffr... 

    Cisco Systems, Inc.

    San Francisco, CA
    1 day ago
  • $204k - $259k

     ...trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has...  ...The World's Most Experienced Driver™—to improve access to mobility while saving thousands...  ...and foster collaborations with other research teams in Alphabet. AI Foundations areas that... 
    Full time
    Temporary work
    Remote work

    Waymo

    San Francisco, CA
    4 days ago
  • $306.3k - $349.5k

     ...Overview Distinguished Applied Researcher At Capital One, we are creating trustworthy...  ...and work with stakeholders to identify and improve the status quo. You’re passionate about...  ...of the following: training optimization, self-supervised learning, robustness, explainability... 
    Full time
    Part time
    Local area
    Flexible hours

    Capital One

    San Francisco, CA
    29 days ago
  • $180k - $235k

     ...error. Nauto technology is designed to predict risk and alerts the driver with advance warning to help prevent collisions, improve driver safety, and save lives. Our customers and prospects include many of the largest commercial fleets in the world along with vehicle... 
    Full time
    Remote work
    Flexible hours

    NAUTO

    San Francisco, CA
    3 days ago
  • $82.7k - $129.8k

     ...their communities thrive, while proactively reducing harm and ensuring support is accessible when needed. Our organization encompasses safety policy & enforcement, fraud prevention, and customer experience. Within Customer Trust, the Strategic Intelligence & Governance (... 
    Flexible hours

    Twitch

    San Francisco, CA
    1 day ago
  •  ...Executive Recruiting Researcher The Executive Recruiting Team (ERT) helps find, engage and hire the next generation of leadership...  ...with passive candidates and talent. Contribute to continuous improvement and innovation in the efficiency and effectiveness of our... 

    Google

    San Francisco, CA
    3 days ago
  • $192.6k - $344.85k

     ...domain-capable systems is still an open research problem.Autodesk touches more of the physical...  ...reasoningDevelop novel algorithms that improve model reliability, controllability, and...  ...horizon reasoning, tool use, agentic behavior, safety, and real-world workflow completionLead... 
    Full time
    For contractors
    Remote work

    Autodesk

    San Francisco, CA
    4 days ago
  • $159.35k - $254.97k

     ...great work and bringing your best, authentic self to everything you do. We value your...  ...new ones “I can succeed as a Quantitative Research Associate at Capital Group”As a member of...  ...and demonstrate commitment to continuously improve skills and self. You are curious about... 
    Full time
    Temporary work
    Local area
    Flexible hours

    Capital Group

    San Francisco, CA
    1 day ago
  • $65.08 - $73.21 per hour

     ...to the appropriate individuals.Safety: Maintains a clean, neat, and...  ...and helps enact processes that improve safety for all laboratory...  ...telephone courtesy, identifying self and department when answering...  ...management by participation in research, development, and implementation... 
    Immediate start
    Shift work

    Ahmc Healthcare

    Daly City, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Researcher, Recursive Self-Improvement Safety. Be the first to apply!