Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Head of AI Safety

Full-time

Moonshot

Moonshot believes that marginalised people in society — including minority ethnic people, people from working class backgrounds, women, Disabled and LGBTQIA+ people — must be centred in the work we do. We strongly encourage applications from people with these identities or who are members of other communities who are currently underrepresented in our workforce. We know a diverse workforce will enable us to understand drivers behind violent extremism and online harms in an in-depth way and do better work to counter them.

About the role

Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and growth of our AI Safety portfolio. The role combines Moonshot's expertise in violence prevention, behavioural risk, and online harms with the emerging practice of evaluating and improving the safety of AI systems. The portfolio addresses harm categories including pathways to violence, extremism, child sexual exploitation, abuse and grooming (CSEA), mental health and crisis, and risks affecting children and teenagers.

The Head of AI Safety will serve as Moonshot's primary applied AI safety counterpart for frontier AI companies, governments, and regulators. The role will work closely with model, policy, trust and safety, product, research, and engineering teams. This is not an engineering or data-science role, but it is a hands-on position requiring the successful candidate to lead and participate directly in red teaming and adversarial evaluation, working in detail with evaluation methodologies, test scenarios, model responses, safety policies, and intervention frameworks. The role holds responsibility for client and partner relationships, project and staff management, methodological quality, and business development. The Head of AI Safety will build and maintain relationships across the wider AI safety ecosystem, including with governments, foundations, regulators, academics, researchers, and civil society organisations.

Your responsibilities will include:

Applied AI Safety, Evaluation, and Advisory

  • Lead and quality-assure Moonshot's applied AI safety work across harm categories including pathways to violence, extremism, CSEA, abuse and grooming, mental health and crisis, and risks affecting children and teens, using methods such as red teaming and adversarial evaluation of AI systems.
  • Advise frontier AI companies on how to improve the safety of their models, products, policies, and intervention systems.
  • Translate insights from psychologists, child-safety specialists, violence-prevention practitioners, safeguarding experts, and other subject-matter experts into clear, actionable guidance for model safety, policy, product, research, and engineering teams.
  • Set the methodological approach for the portfolio, translating violence-prevention, safeguarding, and behavioural-risk expertise into structured and testable evaluation frameworks.
  • Lead and participate directly in red teaming and adversarial evaluation , working in detail with test scenarios, model responses, scoring criteria, safety policies, and evaluation results.
  • Identify patterns, edge cases, and potential safety failures, and develop practical recommendations for improving model behaviour and user protections.
  • Maintain rigour and clear documentation across the team's technical deliverables, suitable for technical, government, and foundation audiences.
  • Ensure work is delivered within a clear ethical framework and in compliance with contractual, legal, data protection, and ethics obligations.
  • Identify, manage, and escalate operational, reputational, delivery, and partnership risks.

Client & Partner Management

  • Serve as Moonshot's primary applied AI safety counterpart for frontier AI company partners, governments, regulators, and the wider ecosystem invested in AI safety.
  • Build trusted relationships with model, policy, trust and safety, product, research, and engineering teams.
  • Build and sustain relationships across the wider AI safety ecosystem, including governments, foundations, regulators, academics, researchers, civil society organisations, and specialist practitioners.
  • Represent Moonshot externally in meetings, briefings, workshops, and sector engagement, including with regulators and policymaker audiences.

Team Leadership & Management

  • Provide direct leadership, coaching, and management to Moonshot's AI safety team.
  • Foster a collaborative, accountable, and mission-driven team culture, with particular attention to wellbeing given the sensitive nature of the work.
  • Support workforce planning, performance management, and professional development across the team.
  • Ensure effective coordination with internal teams supporting the portfolio, including operations, finance, research, and technical teams.

Portfolio Development & Growth

  • Develop Moonshot's AI safety portfolio, identifying strategic opportunities, partnerships, and funding.
  • Lead proposal development, scoping, and renewals with technical credibility, using precise, defensible language suited to technical and government audiences.
  • Develop repeatable methodologies, service offerings, and partnerships that allow the portfolio to grow while maintaining methodological rigour and delivery quality.
  • Support external communications, publications, briefings, and thought leadership that establish Moonshot as a credible voice in applied AI safety.
  • Oversee project planning, staffing, budgeting, forecasting, and delivery timelines across the portfolio.

Requirements

Essential:

  • Experience in trust & safety , online harms, or a closely related field such as violence prevention, safeguarding, or public health, and the ability to adapt that knowledge to AI systems.
  • Curiosity about AI and the ability to build technical fluency quickly, enough to engage credibly with technical counterparts at AI companies. Much of this work is new, so comfort learning as you go matters more than existing AI safety expertise.
  • Experience designing research, evaluation frameworks, or interventions for harm categories such as violent extremism, CSEA, self-harm and crisis, or targeted violence.
  • Demonstrated experience managing projects, teams, budgets, partners, and clients, with strong people management skills.
  • Excellent written communication, with experience producing credible (not promotional) material for government, foundation, or enterprise audiences.
  • Comfort and demonstrated resilience working with highly sensitive or graphic content (CSEA, extremist material, crisis content), with awareness of wellbeing practices for this kind of work.
  • Strong judgment and the ability to navigate ambiguity, competing priorities, and sensitive stakeholder environments, including representing organisations externally.
  • Willingness to travel and work outside regular hours where needed to accommodate clients or respond to incidents.
  • Highly trustworthy, with discretion and diplomacy, and willing to undertake relevant security clearance procedures.
  • Experience supporting business development, grant funding, or procurement.
  • Commitment to Moonshot's mission.
  • We require and will check candidates' eligibility to work in the UK and pass any relevant security clearance procedures per the needs of clients.

Desirable:

  • Direct experience in model safety , red teaming , or adversarial evaluation of LLMs or other AI systems.
  • Understanding of LLM architecture , safety tooling, or trust & safety policy .
  • Prior experience in child safety evaluation, teen-safety product work, or grooming and CSEA detection.
  • Familiarity with government or regulatory engagement, such as briefing officials or supporting policy submissions.
  • Experience with intervention or diversion programme design that can transfer to AI-mediated interventions.
  • Academic or applied background in radicalisation studies, forensic psychology, or violence risk assessment.
  • Familiarity with taxonomy or classifier development, including how testing data feeds a classifier.

Benefits

  • 30 days' paid annual leave, excluding public holidays.
  • Flexible public holiday policy with the option to work public holidays in exchange for a day off at another time.
  • Private healthcare package with access to specialist mental health cover, including coverage for partners and children.
  • Dental and Vision Insurance.
  • Life Insurance & Income Protection.
  • Employee Assistance Programme providing access to mental health support.
  • Generous maternity and paternity leave: 26 weeks paid maternity leave, 8 weeks paid paternity leave.
  • All permanent employees are granted share options upon employment.
Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the Head of AI Safety in Remote vacancy
  • $147.1k - $213.3k

    Job Title AVP, Head of AI Solutions Division: Corp IT Location: Hudson Yards, NYC Reports To: VP, IT Platform Lead Americas Who We Are: For...  ...the world the best of beauty in terms of quality, efficacy, safety, sincerity and responsibility to satisfy all beauty needs and desires... 
    Suggested
    Contract work
    Temporary work
    Work experience placement
    Summer work
    Work at office
    Local area
    Work from home
    Flexible hours
    Shift work

    Loreal

    New York, NY
    4 days ago
  • $180k - $230k

    Head of AI - Digital Content PlatformSalary Range: $180,000 to $230,000Location: New York/Remote (EST Working Timezone Considered) Are you driven by the challenge of applying AI and data science to revolutionize how digital content is created and generates revenue? We... 
    Suggested
    Remote work

    Consortia

    New York, NY
    4 days ago
  •  ...Head Of Ai Enablement Remote Full-Time Research & Development Make Coinbax an AI-native organization. Architect and implement LLM, chat, and agentic integrations across engineering, go-to-market, and executive functions. Democratize software development so teams... 
    Suggested
    Full time
    Remote work

    Coinbax

    United States
    2 days ago
  •  ...Job Title Design and implement the AI integration layer across Coinbax's software stack and operational tech stack. Establish Claude Code as the primary development paradigm, creating standards, templates, and guardrails for AI-assisted engineering Build custom... 
    Suggested
    Remote work

    Banktech Ventures

    United States
    5 days ago
  • $244k - $282k

     ...learning, and data science improve customer outcomes, internal productivity, product differentiation, and operational leverage. The Head of AI manages AI workstreams across the company, turns scattered AI experiments into governed and measurable operating capability, and... 
    Suggested
    Full time
    Remote work
    Flexible hours

    AssetWatch

    Remote
    4 days ago
  •  ...control , and you are interested in applying this knowledge to AI-driven solutions, this is a unique opportunity to operate at the...  ...cost, labour, and asset performance The Role We are seeking a Head of Asset Management & AI Transformation to lead the definition... 
    Remote work

    Max Accelerate Technology Group

    Netherlands
    2 days ago
  •  ...committed to a secure and decarbonized future and provides tailored AI-powered, end-to-end solutions for all industries. Atos Group is...  ..., in a safe and secure information space. Position :  Head of Artificial Intelligence Role Summary The Head of Artificial... 
    For subcontractor
    Casual work
    Remote work

    Atos

    United States
    5 days ago
  •  ...good health starts with what you eat. We provide tools, resources and support to enable users to reach their health goals. As Head of Cal AI , you will be responsible for an important growth area. You will operate as the owner of the business – defining strategy,... 
    Full time
    Temporary work

    MyFitnessPal

    Remote
    22 days ago
  •  ...Head Of AI Enable Impact at the Heart of Technology Development at BCD! Head of AI (Remote) Full time, United Kingdom, Netherlands, Spain This is a rare opportunity to build and lead the AI capability of one of the world's largest travel management companies. Reporting... 
    Full time
    Remote work
    Flexible hours

    BCD TripTech

    United States
    3 days ago
  •  ...Mongo is seeking a Head of AI Platform, GM to lead the development and scaling of a new AI Applications Platform. With millions of developers trusting MongoDB for their workloads, MongoDB is building the future to enable them to create, deploy and manage AI Applications... 
    Local area
    Remote work
    Worldwide
    Day shift

    MongoDB

    New York, NY
    3 days ago
  •  ...good health starts with what you eat. We provide tools, resources and support to enable users to reach their health goals. As Head of Cal AI , you will be responsible for an important growth area. You will operate as the owner of the business - defining strategy, roadmap... 
    Full time

    MyFitnessPal

    Remote
    13 days ago
  • $180k - $220k

     ...platform that combines intelligent cameras, sensors, and AI analytics to help organizations improve safety and operations at scale. We have a solid product-...  ...accessible to any organization. \n Who You Are As the Head of Channel-East , you will own the strategy and... 
    Full time
    Remote work
    Work from home
    Flexible hours
    Night shift

    Rhombus

    Remote
    3 days ago
  •  ...The Sei Foundation is seeking an innovative and strategic Head of Artificial Intelligence to spearhead the growth and integration of AI-focused applications within the Sei blockchain ecosystem. This role will be pivotal in advancing use cases such as machine learning,... 
    Remote work

    Blockchain Works

    San Francisco, CA
    2 days ago
  •  ...Head of AI City: London Division: Corporate - CFO Why work for us? A career at Janus Henderson is more than a job, it's about investing in a brighter future together. Our Mission at Janus Henderson is to help clients define and achieve superior financial outcomes... 
    Temporary work
    Remote work
    Flexible hours

    Janus Henderson Investors

    United States
    1 day ago
  •  ...Head Of AI Maze is building an AI-native vulnerability management platform. Our autonomous agents investigate, triage, and remediate security findings the way a senior analyst would, only faster and at scale. As Head of AI, you'll own the intelligence that makes those... 
    Temporary work
    Part time
    Immediate start
    Remote work

    Maze

    United States
    3 days ago
  • Role Description As Head of Paid Search & AI Advertising (f/m/d), you will take ownership of the strategic direction of our entire Paid Search practice while building our capabilities across emerging AI advertising channels. You will develop future-proof strategies for... 
    Full time
    Remote work
    Flexible hours

    hurra.com

    Remote
    a month ago
  •  ...firm managing ~$500M of its own capital across public markets. We are hiring a Head of Algo to lead our dedicated algorithmic effort to turn our unique market insights into systematic, AI-driven alpha. This is a hands-on leadership role at the intersection of research... 
    Full time

    Deeter Analytics

    Remote
    1 day ago
  •  ...into high-value training data for the next generation of physical AI. This is a rare opportunity to build a new AI business from...  ...infrastructure, operational footprint, and resources to give it a meaningful head start. ~A structural data advantage: Direct access to one of... 
    Full time

    Stord

    Remote
    6 days ago
  • Role Description 3Cloud is seeking a Head of Advisory & AI to lead the strategy, performance, growth, and continued evolution of our Advisory organization while also providing enterprise leadership for 3Cloud AI. This executive will be accountable for building Advisory... 
    Full time
    Visa sponsorship
    Work visa

    3Cloud

    Remote
    2 days ago
  •  ...Upwork is seeking a senior leader to head the Search & Recommendations engineering organization, shaping technical direction, scaling...  ...executive leadership to drive architecture, reliability, and responsible AI across a remote-first, globally distributed team. #J-18808-... 
    Freelance
    Remote work

    Jobleads-US

    Chicago, IL
    5 days ago
  •  ...About Us At Fireworks, we’re building the future of generative AI infrastructure. Our platform delivers the highest-quality models...  ...of Meta PyTorch and Google Vertex AI. About This Role The Head of Developer Relations owns how developers discover, learn, adopt... 
    Full time
    Remote work
    Shift work

    Fireworks AI

    Remote
    13 days ago
  •  ...Lenovo is seeking a Senior Leader to head Core Intelligence, owning AI brain systems across devices and cloud. You will steer agentic orchestration, memory, RAG retrieval, and model optimization, delivering proactive AI experiences to millions of devices worldwide.... 
    Remote work
    Worldwide

    Lenovo

    North Carolina
    2 days ago
  •  ...to Boston and New York City yet provides the natural beauty and safety of a rural campus, with access to outdoor activities, the arts,...  ...implementing systems that ensure long-term sustainability. The Head Coach is responsible for supporting student-athlete academic success... 
    Temporary work
    Relocation
    Flexible hours

    Franklin Pierce University

    Rindge, NH
    1 day ago
  • $400k

    1 day ago Be among the first 25 applicants Get AI-powered advice on this job and more exclusive features. AI Strategy Formulation and...  ...chances of interviewing at DerbySoft by 2x Get notified about new Head of Artificial Intelligence jobs in Dallas, TX . Director, Data... 
    Full time
    Casual work
    Local area
    Remote work

    DerbySoft

    Dallas, TX
    4 days ago
  • $140.2k - $267.6k

     ...on the Fraud Team, you will be at the forefront of developing and executing an innovative strategy utilizing Artificial Intelligence (AI), Machine Learning (ML), Deep Learning (DL), Natural Language Processing (NLP), Large Language Models (LLM), and Multi-Agent Systems... 
    Temporary work

    Adobe

    Austin, TX
    3 days ago
  • TikTok is seeking a Head of Video Policy & Regional High Harm to lead its Video & Regional...  ...strategies for human moderation and AI-driven systems while managing enforcement...  ...will have extensive experience in Trust & Safety, strong leadership skills, and a deep understanding... 

    TikTok

    New York, NY
    3 days ago
  • Ethena, a fully remote-first compliance training company, is seeking a Head of Learning Experience to lead the LX team. You will report to the CTO and oversee strategy, people, budget, and tooling to deliver best-in-class training content. You will drive vision, manage... 
    Remote work

    Ethena

    New York, NY
    1 day ago
  •  ...growth. Role Description This is a full-time hybrid role for a Head of Artificial Intelligence, based in Boston, MA and Manhattan Beach...  ...lead the design, development, and implementation of advanced AI initiatives in alignment with CoAction’s strategic goals. Responsibilities... 
    Full time
    Remote work

    CoAction

    Boston, MA
    2 days ago
  • $165k - $239k

    CodePath seeks a Senior Director of Curriculum to shape the vision and execution of the curriculum portfolio, including AI-native software engineering and cybersecurity courses. This role involves leading curriculum strategy, maintaining course relevance, and managing... 
    Remote job
    Full time

    codepath-org

    New York, NY
    3 days ago
  •  ...committed to a secure and decarbonised future and provides tailored AI‑powered, end‑to‑end solutions for all industries. Atos Group is...  ...sustainably, in a safe and secure information space. Position: Head of AI Delivery Company Atos Location USA / Remote Job Type Fulltime... 
    Full time
    Remote work

    Atos

    Beaverton, OR
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Head of AI Safety. Be the first to apply!