Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Technical AI Policy Researcher, Frontier Risk - Trust and Safety

$108k - $208.8k

TikTok

Responsibilities The Trust & Safety (T&S) Responsible AI Policy team's mission is to ensure the development of GenAI models and applications are safe, fair and trustworthy. We do this by defining, measuring and mitigating safety and fairness AI model risks through policy frameworks, model risk assessments, and upstream policy solutions. The T&S Responsible AI Policy team sits within the T&S GenAI and Emerging Products pillar. We work closely with Trust & Safety teams (product policy, product, engineering, data science, operations, red teaming), business and model teams, and cross-functional stakeholders (comms, legal, public policy) across global markets. Success in this team requires strong policy acumen, judgment, creativity, analytical rigour, and the ability to translate Generative AI risk to different stakeholders effectively. As a Technical AI Policy Researcher on the T&S Responsible AI Policy team, you will champion the responsible development and deployment of our frontier AI models across multiple businesses with a specialty on technical research for frontier risks. You will accelerate technical policy research, incubate new research efforts, and drive end-to-end policy to evaluate workflows for your domain areas. Responsibilities Design and maintain multimodal GenAI policies across safety-relevant domains, including frontier risk in agentic modalities, loss of control or deceptive misalignment. Translate risk and harm models into clear behavioral specifications, evaluation criteria, grading guidance, and system-level safeguards. Define practical boundaries between beneficial uses of AI and assistance that could materially enable harm, exploitation, misuse, or unsafe outcomes. Build policy artifacts that support model training, evaluation, and deployment. Partner with safety researchers, engineers, product teams, and other stakeholders to operationalize policy into scalable model behavior and measurable safeguards. Design end-to-end policy to pre-launch evaluation to post-launch monitoring workflows across safety-relevant domains, including golden set construction, labeling guidance, calibration, adjudication, and eval coverage analysis, to ensure policies can be reliably measured and improved. Use red-teaming results, deployment data, model failures, over-refusals, under-refusals, and ambiguous edge cases to improve policy and evaluation quality over time. Identify emerging capability areas where frontier AI systems could create new safety, fairness or bias challenges or lower barriers to harm. Monitor post-launch model activity to identify gaps in our policy framework to capture unsafe model behaviour. Champion research to strengthen the defensibility and operability of policy positions, including working with Outreach and Partnerships to incorporate external expert input into relevant policy positions. Combine longer-horizon safety research with hands-on launch and deployment work. Contribute to risk reports, policy documentation, launch reviews, and AI governance reviews on the company's approach to building AI responsibly. Support regulatory teams as a subject matter expert on AI compliance related initiatives. Qualifications Minimum Qualifications: 5 years in Trust & Safety, AI Safety Research, AI Ethics, technical AI Governance, or equivalent experience. Advanced degree in Computer Science, Human-Computer Interaction, Engineering, Data Science or quantitative Social Sciences Direct experience in frontier risk research, AI evaluations, red-teaming, or AI governance work. Strong technical understanding of LLM, multimodel, or genmedia model behavior, model failure modes, and safety risks. Demonstrated experience working with external experts and stakeholders, including civil society, government, and academia. Demonstrated success working in a fast-paced technology company or research organization conducting AI impact, risk assessments or algorithmic audits, and/or data science or product development related experience. Ability to advocate for safety amongst a wide variety of business stakeholders including Product Policy, Engineering, Public Policy, Legal, Communications, and Data Science. Preferred Qualifications: Technical knowledge in high efficiency on device architectures, multimodal understanding, V&V of AI systems, or RAG is beneficial but not required. Experience working with governments, frontier AI companies, or AI Safety organizations. Experience working in non-western cultures, with a particular emphasis on the global south. Understanding of Trust & Safety positioning in the entertainment media technology sector, with comfort learning internal tools and product workflows. Trust & Safety Content that this role interacts with includes images, video, and text related to every-day life, but it can also include (but is not limited to) bullying; hate speech; child safety; depictions of harm to self and others, and harm to animals. Hence, it is possible that this role will be exposed to harmful content on a daily basis. TikTok recognises that keeping our platform safe for the TikTok communities is no ordinary job which can be both rewarding and psychologically demanding and emotionally taxing for some. This is why we are sharing the potential hazards, risks and implications in this unique line of work from the start, so our candidates are well informed before joining. We are committed to the wellbeing of all our employees and promise to provide comprehensive and evidence-based programs, to promote and support physical and mental wellbeing throughout each employee's journey with us. We believe that wellbeing is a relationship and that everyone has a part to play, so we work in collaboration and consultation with our employees and across our functions in order to ensure a truly person-centred, innovative and integrated approach. About TikTok TikTok is the leading destination for short-form mobile video. At TikTok, our mission is to inspire creativity and bring joy. TikTok's global headquarters are in Los Angeles and Singapore, and we also have offices in New York City, London, Dublin, Paris, Berlin, Dubai, Jakarta, Seoul, and Tokyo. Why Join Us Inspiring creativity is at the core of TikTok's mission. Our innovative product is built to help people authentically express themselves, discover and connect – and our global, diverse teams make that possible. Together, we create value for our communities, inspire creativity and bring joy - a mission we work towards every day. We stride to do great things with great people. We lead with curiosity, humility, and a desire to make impact in a rapidly growing tech company. Every challenge is an opportunity to learn and innovate as one team. We're resilient and embrace challenges as they come. By constantly iterating and fostering an "Always Day 1" mindset, we achieve meaningful breakthroughs for ourselves, our company, and our users. When we create and grow together, the possibilities are limitless. Join us. Diversity & Inclusion TikTok is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At TikTok, our mission is to inspire creativity and bring joy. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too. TikTok Accommodation TikTok is committed to providing reasonable accommodations in our recruitment processes for candidates with disabilities, pregnancy, sincerely held religious beliefs or other reasons protected by applicable laws. If you need assistance or a reasonable accommodation, please reach out to us at Job Information 【For Pay Transparency】Compensation Description (Annually) The base salary range for this position in the selected city is $108000 - $208800 annually. Compensation may vary outside of this range depending on a number of factors, including a candidate’s qualifications, skills, competencies and experience, and location. Base pay is one part of the Total Package that is provided to compensate and recognize employees for their work, and this role may be eligible for additional discretionary bonuses/incentives, and restricted stock units. Benefits may vary depending on the nature of employment and the country work location. Employees have day one access to medical, dental, and vision insurance, a 401(k) savings plan with company match, paid parental leave, short-term and long-term disability coverage, life insurance, wellbeing benefits, among others. Employees also receive 10 paid holidays per year, 10 paid sick days per year and 17 days of Paid Personal Time (prorated upon hire with increasing accruals by tenure). The Company reserves the right to modify or change these benefits programs at any time, with or without notice. For Los Angeles County (unincorporated) Candidates: Qualified applicants with arrest or conviction records will be considered for employment in accordance with all federal, state, and local laws including the Los Angeles County Fair Chance Ordinance for Employers and the California Fair Chance Act. Our company believes that criminal history may have a direct, adverse and negative relationship on the following job duties, potentially resulting in the withdrawal of the conditional offer of employment: Interacting and occasionally having unsupervised contact with internal/external clients and/or colleagues; Appropriately handling and managing confidential information including proprietary and trade secret information and access to information technology systems; and Exercising sound judgment. #J-18808-Ljbffr TikTok

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Technical AI Policy Researcher, Frontier Risk - Trust and Safety in San Francisco, CA vacancy
  • TikTok is seeking a Technical AI Policy Researcher within the Trust & Safety division to advance responsible frontier AI work across multiple teams in San Francisco. You will accelerate...  ...technical policy research, translate risk insights into measurable safeguards, and... 
    Risk

    TikTok

    San Francisco, CA
    1 day ago
  • $150k - $300k

     ...visual cortex for physical AI. Sight is one of the...  ..., accelerate cancer research, improve construction site safety, digitizing floor plans...  ...You'll Do As a Member of Technical Staff on our Frontier Data team, you’ll build...  ...staffing, timeline, and risks. Team Interview Phase:... 
    Risk
    Remote work
    Work from home
    Home office
    Relocation package

    Roboflow, Inc.

    San Francisco, CA
    1 day ago
  • $171.2k - $273k

     ...Salesforce is the #1 AI CRM, where humans with...  ...meets action. Tech meets trust. And innovation isn't a...  ..., CA | Bellevue, WA Technical Program Managers (TPMs)...  ...incidents and emerging risks, where communication, sound...  ...and maintains a policy of non-discrimination with... 
    Risk

    Salesforce

    San Francisco, CA
    3 days ago
  •  ...Description Job Description Senior Technical Governance Manager...  ...we responsibly innovate with AI. As a Senior Technical Governance...  ...our commitment to guest safety and trust. You’ll empower teams to innovate...  ...controls focused on managing risk in AI/ML systems, ensuring... 
    Risk

    Cypress HCM

    San Francisco, CA
    4 days ago
  • $150k

    Join Amazon's Frontier AI & Robotics team as a Member of Technical Staff, this Technical Program Manager...  ...programs that bridge AI research, software, hardware, and...  ...· Own program-level risk management, proactively...  ...local laws and Company policies. Criminal history may have... 
    Risk
    Local area
    Day shift

    Amazon

    San Francisco, CA
    1 day ago
  • $290k - $365k

     ...and steerable AI systems. We want...  ...of committed researchers, engineers, policy experts, and...  ...breakthrough in AI safety research and...  ...for training frontier models,...  ...things. As a Technical Program Manager...  ...identify technical risks, and add value...  ...and can build trust with both... 
    Risk
    Work at office
    Visa sponsorship
    Flexible hours

    Anthropic

    San Francisco, CA
    3 days ago
  •  ...in San Francisco is seeking exceptional research engineers to push the boundaries of frontier AI safety, shaping empirical understanding of risk and owning end-to-end threads within...  ...datasets, rubrics, and auditable artifacts trusted by leadership during high-stakes... 
    Risk

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  • $123.68k - $254.64k

     ...Possible. At Pinterest, AI isn't just a...  ...Identity & Product Safety programs Own the multi...  ..., product, and policy partners to design...  ...regulations. Provide technical and domain leadership...  ...—clarifying risks, trade‑offs, and options...  ...company Build durable, trust‑based relationships... 
    Risk
    Work at office
    Local area
    Remote work
    Relocation
    Relocation package

    Pinterest

    San Francisco, CA
    3 days ago
  • Mercor is seeking a Trust & Safety product leader to own the fraud detection...  ..., including the multi-agent risk scoring system, signals, and...  ...engineering teams. This is a technical, customer-facing role that...  ...our platform and clients. #J-18808-Ljbffr AI Chopping Block
    Risk

    AI Chopping Block

    San Francisco, CA
    20 hours ago
  • Anthropic is seeking an Analyst to support cyber product policy, enforcement guidance, and launch policy notes....  ...with threat intel and enforcement teams to uphold safety standards. The role involves translating technical risk assessments into policy guidance and engaging with... 
    Risk

    Anthropic

    San Francisco, CA
    2 days ago
  • We're partnering with a frontier AI research company on a search for a Member of Technical Staff focused on AI Safet y. The company is building...  ...model vulnerabilities, misuse risks, and alignment gaps Design and build scalable safety evaluation frameworks and automated... 
    Risk

    Xcede

    San Francisco, CA
    2 days ago
  • $200k - $245k

     ...For the first time ever, safety, operations and finance...  ...with industry leading AI, the Motive platform...  .... We are looking for a Technical Program Manager to join...  ...drive planning, execution, risk management, and...  ...Regulations. It is Motive's policy to require that employees... 
    Risk
    Temporary work
    Work experience placement
    Shift work

    Motive Technologies

    San Francisco, CA
    1 day ago
  • $121k - $148k

     ...Technical Program Manager (TPM) The demand for electricity...  ...than ever, driven by AI data centers and...  ...longer lifetime, better safety, and not limited by raw...  ...streams. Technical Risk Management: Identify and...  ...dependents) ~ Unlimited PTO policy ~40 hours of paid... 
    Risk
    Temporary work

    Inlyte Energy

    San Francisco, CA
    3 days ago
  • $150k

    Join our Frontier AI & Robotics team to lead the...  ...support breakthrough AI research and real-world...  ..., team KPIs, and risk communication to...  ...leadership. Serve as the technical escalation point...  ...coordination, and safety/regulatory...  ...laws and Company policies. Criminal history... 
    Risk
    Local area

    Amazon

    San Francisco, CA
    1 day ago
  • $120k - $180k

     ...About the job Technical Account Manager Technical...  ...Manager (Enterprise SaaS / AI Platform) Compensation...  ...Hiring Target: 1 Work Policy: 5 days/week in-office...  ...scams, abuse, and platform risk. Our unified API powers Trust & Safety teams at Fortune 500 companies... 
    Risk
    Work at office
    Visa sponsorship

    Jenn Nguyen and Friends

    San Francisco, CA
    1 day ago
  •  ...the TeamAt OpenAI, Trust & Safety Operations is...  ...Engineering, Legal, Policy and Go To Market teams...  ...identify emerging risks, build and mature...  ...and scalable, AI-first solutions. You...  ...Business Integrity, technical operations, data...  ...OpenAIOpenAI is an AI research and deployment... 
    Risk
    Flexible hours

    OpenAI

    San Francisco, CA
    1 day ago
  •  ...is in need of a Technical Program Manager...  ...including full safety governance and integration...  ...various safety research and mitigations...  ..., API, and any frontier models. This...  ..., legal and policy, and ensuring all the risks are effectively...  ...OpenAI OpenAI is an AI research and deployment... 
    Risk
    Work at office
    Relocation package

    Slope

    San Francisco, CA
    1 day ago
  •  ...increasingly capable AI systems requires scalable technical safeguards,...  ...ownership of emerging risks, rigorous...  ...coordination across research, engineering,...  ...operations, legal, policy, and external...  ...initiatives that turn safety commitments into...  ..., integrity, trust and safety,... 
    Risk

    Triwill Group

    San Francisco, CA
    1 day ago
  •  ...states. Our team of AI researchers and company...  ...team to translate safety findings into concrete...  ...to deployment policies. Validate that every...  ...meets the lab’s risk thresholds before...  ...AI Safety. Deep technical understanding of...  ...advancing the frontier of intelligence.... 
    Risk
    Relocation package

    B Capital

    San Francisco, CA
    1 day ago
  •  ...is a critical Safety Research team at the company...  ...on mitigating AI threats to...  ...capabilities of frontier AI systems. Mitigation...  ...other frontier-risk areas), and...  ...leadership can trust during high-...  ...and/or another technical domain applicable...  ...Employment Opportunity Policy Statement.... 
    Risk
    Permanent employment
    Temporary work

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  • $104.9k - $199.07k

     ...the healthcare industry’s most trusted solutions for healthcare...  ...healthcare financing, enterprise risk management and regulatory compliance...  ...: MedInsight is seeking a Technical Product Manager with a platform...  ...healthcare analytics and AI products. This role will help... 
    Risk
    Full time
    Work experience placement
    Remote work
    Flexible hours

    Milliman

    San Francisco, CA
    1 day ago
  • Obsidian is seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design...  ..., and assess safety across high-risk topics. Join a cutting-edge research environment, collaborate with AI researchers... 
    Risk

    Obsidian

    San Francisco, CA
    4 days ago
  • $275k - $375k

     ...Manager for the Research team at Anthropic...  ...products as we advance frontier, safe AI technology. You...  ...is an AI safety and research company...  ...across ML, physics, policy, business and...  ...existing products Technical background with experience...  ...about the risks and benefits of new... 
    Risk
    Work at office
    Home office
    Visa sponsorship
    Relocation package
    Flexible hours

    Anthropic

    San Francisco, CA
    1 day ago
  • $207k - $285k

    OpenAI is seeking a Technical Program Manager in San Francisco to lead initiatives that ensure the safety and robustness of its AI models. The role involves collaborating with diverse teams to turn risks into actionable plans. Ideal candidates will have experience in technical... 
    Risk

    OpenAI

    San Francisco, CA
    1 day ago
  •  ...in San Francisco, seeks an experienced Safety & Security Counsel to lead legal guidance on policy design, incident response, and regulatory...  ...manage escalations, and advise executives on risk, privacy, and compliance in a fast-moving AI environment. #J-18808-Ljbffr Menlo... 
    Risk

    Menlo Ventures

    San Francisco, CA
    4 days ago
  • $165k - $202.5k

     ...and life stage. Our AI-native platform helps...  ...Product Manager for AI Trust & Safety will own the roadmap and...  ...scope and prioritize technical work that enables AI safety...  ..., red teaming, or risk mitigation Strong...  ...competitive paid time off policies including vacation,... 
    Risk
    Work at office
    Worldwide
    Sleeping nights
    2 days per week
    3 days per week

    Springhealth66

    San Francisco, CA
    3 days ago
  •  ...Principal Technical Program Manager At Hayden AI, we are on a mission to harness the power of computer vision...  ...accelerate transit, enhance street safety, and drive toward a sustainable future...  ...each domain to surface the right risks and drive the right conversations.... 
    Risk

    Hayden AI

    San Francisco, CA
    4 days ago
  • $116k - $145k

     ...motivated and experienced Technical Program Manager II to...  ..., status updates, risk logs — so stakeholders...  ...circumstances change.Apply AI tools to streamline your...  ....Collaborative: Builds trust through consistent delivery...  ...information on our AI policy, please visit... 
    Risk
    Work at office

    Datadog

    San Francisco, CA
    1 day ago
  • $240k - $260k

     ...JobAs the Director of Technical Program Management, you...  ...software delivery models as AI rapidly transforms the...  ...professional growth, trust, and a welcoming...  ...in trust, psychological safety, mutual accountability,...  ...planning frameworks, and risk management standards across... 
    Risk
    Work at office
    Local area
    Flexible hours
    Shift work
    2 days per week

    OpenTable

    San Francisco, CA
    4 days ago
  • $120k - $180k

     ...Technical Account Manager (TAM) At Bland, we're building the most human AI phone agents in the world. We're a Series...  ...accounts, surface real risks, tighten the feedback...  ...advocacy - Builds trust quickly, navigates competing...  ..., elite team at the frontier of voice AI, with... 
    Risk
    Work at office
    Immediate start

    Bland AI

    San Francisco, CA
    20 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Technical AI Policy Researcher, Frontier Risk - Trust and Safety. Be the first to apply!