Researcher, Recursive Self-Improvement Safety
$295kOpenAI
About the teamModels are becoming increasingly capable—moving from tools that assist humans to agents that can plan, execute, and adapt in the real world. Mitigating the frontier risks resulting from these capabilities is paramount to OpenAI’s ability to continue deploying models safely.The Preparedness team is dedicated to addressing these critical risks. Our work includes:Measurement. Monitoring and predicting the evolving capabilities of frontier AI systems.Mitigation. Keeping misalignment safeguards, alignment tools, and on track to adequately address extreme threats that might arise in the future.Coordination. Setting mitigation targets by maintaining OpenAI’s preparedness framework, and partnering with other staff to achieve these targets.This is urgent, fast-paced work that has far-reaching implications for the company and for society.About the rolePreparedness is hiring strong technical executors to support preparations for accelerated AI development, which may culminate in recursive self-improvement. This work relies on anticipating misalignment risks that might exist in the future, but might not exist now; so it’s especially important that people in this role are tasteful and strategic.The role is wide-ranging, covering any mitigation for loss of control risk, spanning the design and implementation of better pre-deployment risk-assessment, control measures, RSI-relevant training interventions, and turning one’s technical work into established institutional practices and external-facing communications.Below is a subset of our focus areas:Scalable oversight: Establishing practices for model misbehavior monitoring and oversight which remain effective in superhuman model capability regimes, with a focus on bridging from today’s monitoring approaches to future-proof ones.Automated auditing: As model capabilities increase, we’ll increasingly rely on automated approaches for finding the most severe forms of model misalignments. We’ll both need to sift through large swaths of production traffic to find the most egregious misalignments, and reliably elicit tail risks before deployment.Rigorous monitorability: Rigorous testing and red-teaming of our measurements of model misbehavior related to loss-of-control (e.g. reward hacking, sandbagging, scheming). This includes better understanding monitorability, and e.g. preparing for potential losses of Chain-of-Thought monitorability.Model behavior science: Design experiments and evaluations to understand the extent to which models are problematically misaligned, or their safety-relevant capabilities lag behind dangerous capabilities. This may include training model organisms of misbehavior for behaviors not currently present in production, or training interventions to increase safety-relevant capabilities.Coordination and verification: Prototype technical mechanisms for verifying compliance with potential future AI safety agreements.AI R&D risk measurement: Track progress toward automation of technical staff to inform OpenAI’s near-term investments in alignment and security.Maintaining and strengthening RSI safety cases: We’re especially interested in identifying and addressing blindspots of mitigation areas which we may have missed.Generally, our team alternates between performing rigorous hypothesis-driven research and turning our insights into interventions or control systems which impact production models, with occasional support of engineering teams.In this role, you will:Carefully consider the problems OpenAI might face in the future and how to prepare for them.Turn an open-ended objective like “prepare for future misalignment threats” into a much more concrete direction (e.g. “stress-test monitors for scheming”) – prioritizing the work that is most useful to start right now.Execute quickly, building scrappy prototypes, and then improving them iteratively until they become established components of our safety pipelines.Secure buy-in from other staff at OpenAI when necessary, and communicate your work clearly.Collaborate with or manage other staff as needed, since we might need to rapidly scale to tackle these problems quickly.You might thrive in this role if you:Are an exceptional technical executor.Have strong strategic and research taste: you can prioritize effectively in domains with weak feedback loops.Are passionate about mitigating the risks associated with recursive self-improvement.Are driven by a desire to do whatever work most positively impacts the future of AI development.Bonus: you have already done work in one of the domains listed above (ML research, AI alignment, AI verification etc).About OpenAIOpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an
$150k - $250k
...goods, and global social organizations.We research and deploy technologies that power AI-... ...Distyl itself. Our work spans research into self-constructing systems, the development of... ...don’t just want to drive incremental improvements on benchmarks or optimize an existing...SuggestedWork at office3 days per week$295k
About the team The Safety Systems org ensures that OpenAI's most capable models can be... ...Teaming (ART) effort: building scalable, research-driven systems that continuously discover... ...into actionable, production-facing improvements. The goal is to maximize counterfactual...Suggested- ...growing team comprised of quantitative researchers, software engineers, product managers, designers... ...costs, continuously researching improvements to our optimization formulations and ensuring... ...in an object oriented language. A self-starter who embraces ownership and...SuggestedWork at officeVisa sponsorshipFlexible hours
$218.7k - $249.6k
...life. Our work touches every aspect of the research life cycle, from partnering with Academia... ...work with stakeholders to identify and improve the status quo. You’re passionate about talent... ...of the following: training optimization, self‑supervised learning, robustness,...SuggestedFull timePart timeLocal areaFlexible hours- ...life. Our work touches every aspect of the research life cycle, from partnering with academia... ...work with stakeholders to identify and improve the status quo. You’re passionate about talent... ...of the following: training optimization, self‑supervised learning, robustness,...SuggestedFlexible hours
$262.5k - $299.6k
...Applied Researcher II (AI Foundations, LLM Core and Agentic AI) Overview At Capital One, we... ...and work with stakeholders to identify and improve the status quo. You’re passionate about talent... ...of the following: training optimization, self‑supervised learning, robustness,...Full timePart timeLocal areaFlexible hours- ...advance. We believe this can be dramatically improved. At Axiom, we generate and curate... .... We are looking for a machine learning researcher to help define and build the core AI systems... ...systems. Work on contrastive learning, self-supervised learning, semi-supervised...
- ...of the grid that affect reliability and safety. Gridware’s advanced Active Grid Response... ...mitigation. This comprehensive approach helps improve safety, reduce outages, and ensure the... ...is Gridware's first dedicated design research role, and it is a foundational hire for our...Local areaShift work
$162.7k - $263.18k
...across our enterprise customers’ networks.As a Principal Security Researcher, you will play a key technical leadership role in shaping how... ..., exploit, and attack technique areas where new or improved protections are needed.Drive innovative detection ideas from concept...Full timeWork at office- ...this role to work on site in the specified location(s).As an AI Researcher within Schwab’s AI Strategy & Transformation (AI.x)... ...establishing robust evaluation and monitoring practices, and improving performance under real‑world constraints. You’ll collaborate closely...Full timeWork at office
$142.7k - $270.95k
...novices or experienced engineers, on a user-friendly platform improved by groundbreaking AI capabilities. The AI Foundations team sits... ...optimization, and production deployment.Collaborate closely with Adobe Research, engineering, and product teams to bring AI-powered features to...Full timeTemporary workLocal areaWorldwide- ...connected app. We've helped millions of people understand and improve their health by providing daily insights and practical steps to... ...out of the office. We're looking for a Senior Human Factors Researcher to join our Design Research team. This role is ideal for a researcher...Work at officeLocal areaRemote workFlexible hours2 days per week3 days per week
$139.6k - $225.78k
...relationships, and the kind of precision that drives great outcomes.Job SummaryJob SummaryWe are seeking a passionate and self-driven Sr. Staff Researcher to join our Cloud-Delivered Security Services team. In this role, you will be pivotal in developing and refining the...Full timeWork experience placementWork at officeVisa sponsorshipWork visa$96.7k - $126.4k
Conduct independent and collaborative research using human data, human biospecimens, and induced pluripotent stem cell (iPSC)-based models... ...aimed at understanding cardiovascular health disparities and improving health equity in understudied populations.Demonstrated success...- ...Security Researcher We believe that software is the foundation of modern civilization - yet vulnerabilities threaten its integrity,... ...engineers to understand limitations and design new methodologies to improve our system Publish internal technical reports and...Full timeWork at office
$150k
...apps out of POCs and into production. We eliminate the risk and improve the reliability of LLM apps by haizing them – i.e. rigorously,... ...proactively, and continuously fuzz-testing them. We are looking for Research Engineers to help develop our reliability platform, with a...Visa sponsorship- ...hardware). Train and tune policies, curate data, and iterate to improve real‑world performance. Write production‑quality code that... ...understand use cases and observe robots in deployment contexts. Bridge research and operations: translate research advances into deployable...
$96.7k - $126.4k
Assistant Professional Researcher position available in the Dept. of Laboratory Medicine, University of California, San FranciscoAvailability... ...goals. Candidates should have a strong interest in directly improving patient care through their research and gaining experience in...Traineeship- Tilde Research is a moonshot AI lab advancing mechanistic interpretability, new architectures, and pretraining science. We build foundational... ...and engineers to advance interpretability as a tool for improving performance and control. What you might work on: Designing,...
$285k - $380k
...quickly growing group of committed researchers, engineers, policy experts,... ...science at a leading AI safety company. Responsibilities... ...psychometric techniques to analyze and improve interviewer calibration and... ...~ Experience building self-service analytics tools or dashboards...Full timeWork at officeVisa sponsorshipFlexible hours- .... Our work is at the cutting edge of AI research, focusing on developing methodologies that... ...of our approach are: (1) harnessing improved capabilities into alignment, making sure... ...powerful tool that must be created with safety and human needs at its core, and to achieve...Work at officeRelocation package
- ...Foundation AI, we are leading frontier AI research across Cisco. Our mission is to advance... ...AI infrastructure. You will drive improvements to training algorithms, curate and optimize... ...systems for security-focused use cases, safety, and robust machine learning #J-18808-Ljbffr...
$204k - $259k
...trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has... ...The World's Most Experienced Driver™—to improve access to mobility while saving thousands... ...and foster collaborations with other research teams in Alphabet. AI Foundations areas that...Full timeTemporary workRemote work$306.3k - $349.5k
...Overview Distinguished Applied Researcher At Capital One, we are creating trustworthy... ...and work with stakeholders to identify and improve the status quo. You’re passionate about... ...of the following: training optimization, self-supervised learning, robustness, explainability...Full timePart timeLocal areaFlexible hours$180k - $235k
...error. Nauto technology is designed to predict risk and alerts the driver with advance warning to help prevent collisions, improve driver safety, and save lives. Our customers and prospects include many of the largest commercial fleets in the world along with vehicle...Full timeRemote workFlexible hours$82.7k - $129.8k
...their communities thrive, while proactively reducing harm and ensuring support is accessible when needed. Our organization encompasses safety policy & enforcement, fraud prevention, and customer experience. Within Customer Trust, the Strategic Intelligence & Governance (...Flexible hours- ...Executive Recruiting Researcher The Executive Recruiting Team (ERT) helps find, engage and hire the next generation of leadership... ...with passive candidates and talent. Contribute to continuous improvement and innovation in the efficiency and effectiveness of our...
$192.6k - $344.85k
...domain-capable systems is still an open research problem.Autodesk touches more of the physical... ...reasoningDevelop novel algorithms that improve model reliability, controllability, and... ...horizon reasoning, tool use, agentic behavior, safety, and real-world workflow completionLead...Full timeFor contractorsRemote work$159.35k - $254.97k
...great work and bringing your best, authentic self to everything you do. We value your... ...new ones “I can succeed as a Quantitative Research Associate at Capital Group”As a member of... ...and demonstrate commitment to continuously improve skills and self. You are curious about...Full timeTemporary workLocal areaFlexible hours$65.08 - $73.21 per hour
...to the appropriate individuals.Safety: Maintains a clean, neat, and... ...and helps enact processes that improve safety for all laboratory... ...telephone courtesy, identifying self and department when answering... ...management by participation in research, development, and implementation...Immediate startShift work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Researcher, Recursive Self-Improvement Safety. Be the first to apply!
- qualitative researcher San Francisco, CA
- vulnerability researcher San Francisco, CA
- independent researcher San Francisco, CA
- human factors researcher San Francisco, CA
- product researcher San Francisco, CA
- machine learning researcher San Francisco, CA
- music researcher San Francisco, CA
- court researcher San Francisco, CA
- remote researcher San Francisco, CA
- data collection researcher San Francisco, CA

