Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Frontier AI Risk Evaluation Scientist

Scale AI

An innovative tech company in New York is seeking a Research Scientist focused on Frontier Risk Evaluations. The ideal candidate will contribute to designing and creating evaluation measures for assessing AI risks and will have strong experience in machine learning and technical research. The role includes collaboration with public sector organizations and requires excellent communication skills. This position offers competitive compensation packages including salary, equity, and comprehensive benefits. #J-18808-Ljbffr Scale AI

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Frontier AI Risk Evaluation Scientist in San Francisco, CA vacancy
  • $216k - $270k

    Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities and safeguarding AI models and systems. Building on this expertise, Scale Labs... 
    Risk
    Full time

    Scale AI

    San Francisco, CA
    3 days ago
  • $216k - $270k

    Scale AI, Inc. is looking for a Research Scientist specializing in Frontier Risk Evaluations to develop measures for assessing risks of advanced AI systems. In this role, you will design testing harnesses, collaborate with agencies, and publish reports to inform policymakers... 
    Risk

    Scale AI, Inc.

    San Francisco, CA
    2 days ago
  •  ...is seeking exceptional research engineers to push the boundaries of frontier AI safety, shaping empirical understanding of risk and owning end-to-end threads within this effort. You’ll design evaluations of frontier models against real threat models, develop datasets,... 
    Risk

    United States Digital Space LLC

    San Francisco, CA
    4 days ago
  • $320k

     ...interpretable, and steerable AI systems. We want AI to...  .... About the Team The Frontier Red Team (FRT) is a...  ...surface. As a Research Scientist on FRT focusing on cyber...  ...experiments to elicit and evaluate autonomous AI cyber...  ...etc.) to characterize risks, defensive potential, and... 
    Risk
    Work at office
    Relocation
    Visa sponsorship
    Flexible hours

    Neura Market

    San Francisco, CA
    2 days ago
  • OpenAI is seeking a Researcher for Frontier Cybersecurity Risks to design and implement an end-to-end mitigation stack that reduces severe cyber...  ...design model-enabled safeguards across various surfaces, evaluate trade-offs, and lead testing and red-teaming efforts to adapt... 
    Risk

    Triwill Group

    San Francisco, CA
    3 days ago
  •  ...AI Biologist - Refusal As molecular data generation and frontier model intelligence grows, new approaches...  ...empowering over 5,000 scientists across 150+ R&D labs...  ...establish ground truth for evaluations. You will work as...  ...research from high-risk requests Pathogen... 
    Risk
    Full time
    Contract work
    Work at office
    Remote work
    Flexible hours
    Night shift

    LatchBio

    San Francisco, CA
    12 days ago
  •  ...molecular data generation and frontier model intelligence grows,...  ...analysis, empowering over 5,000 scientists across 150+ R&D labs to...  ...our technical approach for evaluating how AI agents reason through complex...  ...Experience with immunogenicity risk assessment Hands-on work... 
    Risk
    Full time
    Contract work
    Work at office
    Remote work
    Flexible hours
    Night shift

    LatchBio

    San Francisco, CA
    12 days ago
  • $120k - $180k

     ...the Role At LatchBio, we build the benchmarks that frontier AI labs use to evaluate and train models on biological reasoning....  ...from requests that present meaningful biosecurity risks. We're looking for scientists with deep expertise in areas such as biosecurity,... 
    Risk
    Remote work
    Visa sponsorship

    LatchBio

    San Francisco, CA
    1 day ago
  • $216k - $270k

    Research Scientist, AI Controls and Monitoring Scale Labs, Research Scientist - AI Controls...  ...Monitoring As the leading data and evaluation partner for frontier AI companies, Scale plays an...  ...informed, scientific decisions about AI risks and capabilities. Our research tackles... 
    Risk
    Full time

    Scale

    San Francisco, CA
    3 days ago
  • $180k - $260k

     ...About the RoleWe’re looking for an Applied Scientist, AI to turn messy, high-stakes healthcare...  ...build strong baselines, design honest evaluations, run careful error analysis, and...  ...notes, scheduling, utilization, quality, risk, or patient engagement dataYou have experience... 
    Risk
    Temporary work
    Work at office
    Monday to Friday
    Monday to Thursday

    Sprinter Health

    San Francisco, CA
    1 day ago
  • $300k - $405k

     ...reliable, interpretable, and steerable AI systems. We want AI to be safe and...  ...time: designing and running capability evaluations against frontier models, generating and curating training...  ...modeling experts to ground them in realistic risk Train, tune, and iterate on safety... 
    Risk
    Visa sponsorship
    Shift work

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  •  ...which is focused on mitigating AI threats to global security...  ...the evolving capabilities of frontier AI systems. Mitigation. Keeping...  ...frontier preparedness capability evaluations—designing new evals grounded in...  ...as cyber and other frontier-risk areas), and maintaining existing... 
    Risk
    Permanent employment
    Temporary work

    United States Digital Space LLC

    San Francisco, CA
    4 days ago
  •  ...which is focused on mitigating AI threats to global security...  ...the evolving capabilities of frontier AI systems. Mitigation. Keeping...  ...Researcher for cybersecurity risks, you will help design and implement...  ...and new model capabilities. Evaluate technical trade-offs within... 
    Risk

    United States Digital Space LLC

    San Francisco, CA
    4 days ago
  •  ...learning researchers who want to shape the frontier of generative models for the atomistic...  ...of what’s possible and try out new, high‑risk ideas. Machine learning researcher with professional...  ...working on frontier problems in physical AI to invent the blueprint for how they will... 
    Risk
    Work at office

    Achira

    San Francisco, CA
    1 day ago
  • Scale Labs in San Francisco is seeking a Research Scientist focused on Agent Robustness to advance safe, aligned AI. You will tackle fundamental challenges in evaluating agent capabilities, safety, and risk, and help design benchmarks and protocols. You will design harnesses... 
    Risk

    Scale AI, Inc.

    San Francisco, CA
    2 days ago
  •  ...the Role We’re looking for an Applied Scientist, AI to turn messy, high‑stakes healthcare problems...  ...build strong baselines, design honest evaluations, run careful error analysis, and...  ...notes, scheduling, utilization, quality, risk, or patient engagement data You have experience... 
    Risk
    Temporary work
    Work at office
    Relocation package
    Monday to Friday
    Monday to Thursday
    Flexible hours

    Sprinter Health

    San Francisco, CA
    4 days ago
  •  ...the Team Our Cyber team builds AI systems and products that help...  ...the safety and reliability of frontier models in security-sensitive...  ...engineering, model training, evaluations, safeguards, and deployment to...  ...understand cyber use cases, evaluate risk, and turn feedback into... 
    Risk
    Full time

    OpenAI

    San Francisco, CA
    19 hours ago
  • Carnaby Fox is seeking a Member of Technical Staff (AI Research) in San Francisco to help shape the research direction for frontier AI models. You will collaborate with world-class researchers to design experiments, evaluate LLMs, and improve data quality for high-stakes... 

    Carnaby Fox

    San Francisco, CA
    19 hours ago
  • Crucibl is building judgment at scale for enterprise AI. You will push the frontier of how models reason on high-stakes business decisions, and translate...  ...experiments that inform product direction. You’ll design evaluation frameworks for uncertainty, multi-step reasoning, and... 

    Crucibl

    San Francisco, CA
    2 days ago
  • $80 - $150 per hour

     ...consulting firm is seeking Senior Behavioral Health Experts to work part-time and remotely on frontier AI research projects. You will be responsible for designing evaluations and testing AI systems in critical mental health contexts. The ideal candidate has over 5 years... 
    Risk
    Hourly pay
    Part time
    Remote work

    Aligned Labs

    San Francisco, CA
    3 days ago
  • Traverse is a research data lab building reinforcement learning environments for frontier AI labs. As a Research Scientist, you will design and build RL environments that teach models to do work that has historically required years of human expertise. You’ll work at the... 

    Traverse

    San Francisco, CA
    3 days ago
  • Harnham is seeking a Research Scientist to tackle frontier problems in world models and RL at an on-site SF lab. This role emphasizes shipping research...  ...-driven approach and close collaboration with a stealth AI team. You will own end-to-end research from hypothesis to... 
    Relocation package

    Harnham

    San Francisco, CA
    1 day ago
  • Anthropic in San Francisco seeks a Research Scientist to measure recursive-self-improvement in large models. You will design evaluations, build models of capability growth, and...  ...RL, and policy teams to advance safe and reliable AI systems. #J-18808-Ljbffr Anthropic

    Anthropic

    San Francisco, CA
    4 days ago
  • A leading AI evaluation firm based in San Francisco seeks a Machine Learning Scientist to foster understanding of AI model performance. You'll engage in designing and analyzing comprehensive experiments while collaborating across teams. Applicants should possess a PhD... 

    Arena Intelligence, Inc.

    San Francisco, CA
    4 days ago
  •  ...Technologies is building quantum-accelerated AI servers to exponentially speed up...  ...profound global impact. About the Role Frontier AI is moving toward scientific reasoning...  ...next era. We are looking for a Research Scientist who can help define Quantum AI: not just... 
    Casual work
    Visa sponsorship

    Sygaldry

    San Francisco, CA
    19 hours ago
  • the company is seeking a Research Scientist to advance measurable recursive-self-improvement in large models. You will design evaluations, build models of capability growth, and interpret...  ..., and a track record in evaluating AI systems. #J-18808-Ljbffr United States... 

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  •  ...On-site Department Technical About the Role Generative AI is transforming what's computationally possible—but it's...  ...offers a path through these bottlenecks. As an ML Research Scientist, you'll work at the frontier of generative modeling and quantum acceleration,... 
    Full time
    Casual work
    Visa sponsorship

    Wheel the World

    San Francisco, CA
    19 hours ago
  • $140k - $200k

    Research Engineer & Scientist The Center for AI Safety (CAIS) is a leading research...  ...mitigating societal-scale risks from AI. We address the toughest...  ...AI safety institutes and frontier AI labs, and they have...  ...build the tooling to train and evaluate models at scale, and turn... 
    Risk
    Work at office
    Local area

    Center for Ai Safety

    San Francisco, CA
    4 days ago
  • OpenAI is seeking a cybersecurity risk researcher to design and implement an end-to-end mitigation stack to reduce severe misuse across OpenAI’s products. This role requires deep technical depth and close cross‑functional collaboration to ensure safeguards are enforceable... 
    Risk

    OpenAI

    San Francisco, CA
    3 days ago
  • $150k

    Join Amazon's Frontier AI & Robotics team as a Member of Technical Staff, this Technical Program Manager will become the driving force behind...  ...decisions are one-way or two-way doors· Own program-level risk management, proactively identifying technical, schedule, and resource... 
    Risk
    Local area
    Day shift

    Amazon

    San Francisco, CA
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Frontier AI Risk Evaluation Scientist. Be the first to apply!