Researcher, Robustness & Safety Training
$380kOpenAI
About the TeamThe Safety Systems team is responsible for various safety work to ensure our best models can be safely deployed to the real world to benefit the society and is at the forefront of OpenAI's mission to build and deploy safe AGI, driving our commitment to AI safety and fostering a culture of trust and transparency.The Model Safety Research team aims to fundamentally advance our capabilities for precisely implementing robust, safe behavior in AI models, and to leverage these advances to make OpenAI’s deployed models safe and beneficial. This requires a breadth of new ML research to address the growing set of safety challenges as AI becomes more powerful and used in more settings. Key focus areas include how to enforce nuanced safety policies without trading off helpfulness and capabilities, how to make the model robust to adversaries, how to address privacy and security risks, and how to make the model trustworthy in safety-critical domains. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. About the RoleOpenAI is seeking a senior researcher with passion for AI safety and experience in safety research. Your role will set directions for research to enable and empower safe AGI and work on research projects to make our AI systems safer, more aligned and more robust to adversarial or malicious use cases. You will play a critical role in shaping how a safe AI system should look like in the future at OpenAI, making a significant impact on our mission to build and deploy safe AGI.In this role, you will:Conduct state-of-the-art research on AI safety topics such as RLHF, adversarial training, robustness, and more.Implement new methods in OpenAI’s core model training and launch safety improvements in OpenAI’s products.Set the research directions and strategies to make our AI systems safer, more aligned and more robust.Coordinate and collaborate with cross-functional teams, including T&S, legal, policy and other research teams, to ensure that our products meet the highest safety standards.Actively evaluate and understand the safety of our models and systems, identifying areas of risk and proposing mitigation strategies.You might thrive in this role if you:Are excited about OpenAI’s mission of building safe, universally beneficial AGI and are aligned with OpenAI’s charterDemonstrate a passion for AI safety and making cutting-edge AI models safer for real-world use.Bring 4+ years of experience in the field of AI safety, especially in areas like RLHF, adversarial training, robustness, fairness & biases.Hold a Ph.D. or other degree in computer science, machine learning, or a related field.Possess experience in safety work for AI model deploymentHave an in-depth understanding of deep learning research and/or strong engineering skills.Are a team player who enjoys collaborative work environments.About OpenAIOpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement.Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations.To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form. No response will be provided to inquiries unrelated to job posting compliance.We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link.OpenAI Global Applicant Privacy PolicyAt OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.Compensation Range: $380K - $500KLocationSan Francisco; London, UKEmployment TypeFull timeDepartmentSafety SystemsCompensation$380K – $500K • Offers EquityThe base pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience. If the role is non-exempt, overtime pay will be provided consistent with applicable laws. In addition to the salary range listed above, total compensation also includes generous equity, performance-related bonus(es) for eligible employees, and the following benefits.Medical, dental, and vision insurance for you and your family, with employer contributions to Health Savings AccountsPre-tax accounts for Health FSA, Dependent Care FSA, and commuter expenses (parking and transit)401(k) retirement plan with employer matchPaid parental leave (up to 24 weeks for birth parents and 20 weeks for non-birthing parents), plus paid medical and caregiver leave (up to 8 weeks)Paid time off: flexible PTO for exempt employees and up to 15 days annually for non-exempt employees13+ paid company holidays, and multiple paid coordinated company office closures throughout the year for focus and recharge, plus paid sick or safe time (1 hour per 30 hours worked, or more, as required by applicable state or local law)Mental health and wellness supportEmployer-paid basic life and disability coverageAnnual learning and development stipend to fuel your professional growthDaily meals in our offices, and meal delivery credits as eligibleRelocation support for eligible employeesAdditional taxable fringe benefits, such as charitable donation matching and wellness stipends, may also be provided.More details about our benefits are available to candidates during the hiring process.This role is at-will and OpenAI reserves the right to modify base pay and other compensation components at any time based on individual performance, team or company results, or market conditions.
$380k
About the TeamThe Safety Training research team aims to fundamentally advance our capabilities for precisely implementing safe behavior in AI models... ...to train nuanced safety behaviors, how to make the model robust to bad actors, how to address privacy and security risks,...TrainingWork at officeLocal areaRemote workFlexible hours$150k - $250k
...goods, and global social organizations.We research and deploy technologies that power AI-... ...evaluating intelligent systems: adversarial robustness testing, longitudinal performance... ...intelligent systems using models rather than training or fine-tuning them. Ideal candidates...TrainingWork at office3 days per week$216.3k - $280.8k
...Foundation AI, we are leading frontier AI research across Cisco. Our mission is to... ...learning, reasoning systems, scalable training algorithms, evaluation science, inference... ...systems for security-focused use cases, safety, and robust machine learning.Demonstrated...TrainingFull timeTemporary workLocal areaFlexible hours$139.6k - $225.78k
...a passionate and self-driven Sr. Staff Researcher to join our Cloud-Delivered Security Services... ...adversarial activities and implement robust, proactive protections.Building high-... ...TensorFlow, PyTorch, Scikit-Learn) - including training and testing workflows, and technologies....TrainingFull timeWork experience placementWork at officeVisa sponsorshipWork visa$218.7k - $249.6k
Applied Researcher I Overview: At Capital One, we are creating trustworthy and reliable... ...phases of development, from design through training, evaluation, validation, and... ...optimization, self-supervised learning, robustness, explainability, RLHF. An engineering...TrainingFull timePart timeLocal areaFlexible hours$250k - $350k
...boundaries of what's possible in LLM post-training. If you love training models, exploring... ..., running experiments, and turning research insights into products that ship, we'd love... ...findings in production settings Create robust benchmarks and evaluation frameworks...TrainingWork at office$262.5k - $299.6k
Applied Researcher II (AI Foundations, LLM Core and Agentic AI) Overview: At Capital One... ...phases of development, from design through training, evaluation, validation, and... ...optimization, self-supervised learning, robustness, explainability, RLHF. An engineering...TrainingFull timePart timeLocal areaFlexible hours$216k - $270k
Scale Labs, Research Scientist — Safety Post TrainingAs the leading data and evaluation partner for frontier... ...the hardest problems in agent robustness, AI control protocols, and AI risk evaluations... ...Scientist working on Safety Post-Training you will develop and apply post-...TrainingFull time$141.2k - $257.1k
Full Professional Researcher - Krummel labThe Krummel lab in the Department of Pathology seeks a computational scientist to help lead the... ...computational predictions.Work with project leads to co-mentor and train junior bioinformaticians, and establish best practices for...Training- DescriptionAbout the RoleWe are looking for a Senior AI Researcher to join our growing AI team and help build intelligent, production-grade... ..., PyTorch fluently, and enough systems sense to know why your training run is slow.You design experiments. You state the hypothesis,...TrainingFull timeTemporary workInternshipImmediate start
$96.7k - $257.1k
Professional Researcher in Molecular Imaging and Neurobiology of Neurodegenerative ProteinopathiesThe... ...and statistics, including personnel training, experimental design, and project... ...will be required to verify all department safety training is complete and perform...TrainingTraineeship$96.7k - $257.1k
...Sciences (BTS) seeks individuals with strong understanding of research development who can help provide the leadership required to implement... ...with this faculty level and at least 2 years of experience in training and supervising scientists in the laboratory setting are...TrainingImmediate startWorldwide$245k - $285k
...growing group of committed researchers, engineers, policy experts,... ...biological scientists to help build safety and oversight mechanisms for... ...modeling experts to develop training data for our safety systems,... ..., optimizing for both robustness against adversarial attacks...TrainingFull timeWork at officeVisa sponsorshipFlexible hoursShift work$116.2k - $146.7k
Associate Professional Researcher - Andino LabThe Andino Laboratory at UCSF is seeking an Associate Professional Researcher to join the lab... ...antiviral immunity.The Associate Professional Researcher will train and mentor new members of the laboratory, with particular...Training$96.7k - $126.4k
Assistant Professional Researcher - Okada labThe Okada lab in the Department of Neurological... ...scientific guidance, technical training, and oversight of laboratory and computational... ...mentor new laboratory members in laboratory safety, cell culture techniques, specimen...TrainingTraineeshipFlexible hoursAfternoon shift$204k - $300k
...you do your best work.The Advanced Technology Group (ATG) is the research division of the company. ATG’s mission is to look ahead,... ...experience, market demands, internal parity, and relevant education or training. Your recruiter can share more about the specific salary range...TrainingFull timeLocal areaWorldwideFlexible hours$96.7k - $126.4k
Assistant Professional Researcher position available in the Dept. of Laboratory Medicine, University of California, San FranciscoAvailability... ...available for participating in abstracts and publications. Training goals and specific projects will be tailored to the candidate’...TrainingTraineeship$118k - $309.8k
Educational Research Scientist Department of Surgery and the Center for Faculty Educators: University of California, San FranciscoThe... ...and facilitate collaboration with international surgical skills training programs. This is a 1.0 FTE position. - 70% Department of Surgery...TrainingFull timeWork at office- ...Proximal is building the research systems needed to identify what models can’t yet do, build the tasks required to teach them, measure... ...help them build evals, identify model weaknesses and curate post-training data to fix these weaknesses. This role is perfect for...Training
- ...Tilde Research is a moonshot AI lab advancing mechanistic interpretability, new architectures, and pretraining science. We build foundational... ...—and use those insights to make them better. You'll work on training, analyzing, and evaluating cutting-edge models, collaborating...Training
$150k - $300k
...Join to apply for the Applied Researcher role at Variant Overview Variant is hiring an applied researcher to join our team in San Francisco... ...with creativity and taste. This role involves designing, training, and evaluating deep neural networks and large language models...TrainingFull time$247k - $396k
...Computer Science, or Robotics (Top-tier research lab background) Deep expertise in Vision... ..., qualifications, relevant education or training, and market conditions. These ranges may... ...everything we do is our commitment to safety. Building best-in-class self-driving...TrainingWork at officeLocal area3 days per week$350k
...Research Engineer / Scientist, AlignmentSan Francisco, CAAbout AnthropicAnthropic... ...experimental research on AI safety, with a focus on risks from... ...Research: Developing robust defenses against adversarial... ...of our safety techniques by training language models to subvert...TrainingWork at officeVisa sponsorshipFlexible hours$117.8k - $176.8k
Grounded in safety, quality, and ethics, our experts lead their fields with dedication, a... ...working from home), flexible work hours and a robust compensation and benefits package.Your... ...of high performers through coaching and training staff to expand understanding of business...TrainingFull timeTemporary workPart timeCasual workWork at officeLocal areaRemote workWork from homeWorldwideFlexible hoursNight shift$190k - $205k
...of the grid that affect reliability and safety. Gridware’s advanced Active Grid Response... ...characteristics. Design models that are robust, interpretable, and deployable in... ...including data processing, feature extraction, training, evaluation, and deployment. Optimize...TrainingFull timeLive in- ...of the grid that affect reliability and safety. Gridware’s advanced Active Grid Response... ...intelligence. This role blends applied research, model optimization, and low-level implementation... ...analysis, feature engineering, model training, evaluation, and optimization. Design...TrainingLocal area
- ...this role to work on site in the specified location(s).As an AI Researcher within Schwab’s AI Strategy & Transformation (AI.x)... ...large‑language‑model‑based and agent‑driven systems, establishing robust evaluation and monitoring practices, and improving performance...Full timeWork at office
$380k
...for society.About the RoleWe are seeking exceptional researchers who can push the frontier of safety mitigations. You will help derisk frontier models by... ...red-teaming pipelines to examine the end-to-end robustness of our safety systems, and identify areas for future...Work at officeLocal areaFlexible hours- ...staffingfuture.comSummary of Position: The US Patient Safety Sr. PV Clinical Scientist -... ...potential compliance impact Responsible for training internal USPS and/or vendors staff or... ...reconciliation of outgoing communications Maintains robust knowledge of Technical Product Complaints...TrainingLocal areaFlexible hours
$285k - $380k
...People Research Scientist, RecruitingSan Francisco, CA | New York City, NY | Seattle, WAAbout... ...field of people science at a leading AI safety company.ResponsibilitiesResearch Design... ...an equivalent combination of education, training, and/or experienceRequired field of...TrainingWork at officeVisa sponsorshipFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Researcher, Robustness & Safety Training. Be the first to apply!
- survey researcher San Francisco, CA
- remote researcher San Francisco, CA
- independent researcher San Francisco, CA
- music researcher San Francisco, CA
- vulnerability researcher San Francisco, CA
- machine learning researcher San Francisco, CA
- security researcher San Francisco, CA
- design researcher San Francisco, CA
- product researcher San Francisco, CA
- qualitative researcher San Francisco, CA



