Frontier AI Risk Evaluations Scientist
Gravity Engineering Services Pvt Ltd.
Gravity Engineering Services Pvt Ltd. is seeking a Research Scientist for Frontier Risk Evaluations to create evaluation measures for AI systems. Candidates should have a strong commitment to safe AI deployments, practical experience in ML research, and excellent communication skills. Key tasks include designing evaluation harnesses, collaborating with government agencies, and publishing methodologies. Ideal candidates will have a track record in machine learning, particularly generative AI, and at least three years of related experience. #J-18808-Ljbffr Gravity Engineering Services Pvt Ltd.
- About Scale Labs, Research Scientist — Frontier Risk Evaluations As the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities and safeguarding AI models and systems. Building on this expertise, Scale...Risk
$197.4k - $246.75k
Research Scientist, Frontier Risk Evaluations Scale Labs, Research Scientist — Frontier Risk Evaluations As the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the capabilities and safeguarding AI models and...RiskFull time$216k - $270k
Scale AI, Inc. is looking for a Research Scientist specializing in Frontier Risk Evaluations to develop measures for assessing risks of advanced AI systems. In this role, you will design testing harnesses, collaborate with agencies, and publish reports to inform policymakers...Risk- An innovative tech company in New York is seeking a Research Scientist focused on Frontier Risk Evaluations. The ideal candidate will contribute to designing and creating evaluation measures for assessing AI risks and will have strong experience in machine learning and...Risk
$350k
...interpretable, and steerable AI systems. We want AI to... .... About The Team The Frontier Red Team (FRT) is a... ...surface. As a Research Scientist on FRT focusing on cyber... ...experiments to elicit and evaluate autonomous AI cyber... ...etc.) to characterize risks, defensive potential, and...RiskWork at officeVisa sponsorshipFlexible hours$216k - $270k
Scale Labs, Research Scientist - AI Controls and Monitoring As the leading data and evaluation partner for frontier AI companies, Scale plays an integral role in understanding the... ...make informed, scientific decisions about AI risks and capabilities. Our research tackles...RiskFull time- ...learning researchers who want to shape the frontier of generative models for the atomistic... ...of what’s possible and try out new, high‑risk ideas. Machine learning researcher with professional... ...working on frontier problems in physical AI to invent the blueprint for how they will...RiskWork at office
$245k
...robustness, and reliability of AI models towards their... ...health outcomes. As a Research Scientist, you will contribute to the development... ..., scalable oversight, etc. Evaluate methods using health-related data... ...systems, identifying areas of risk. Work with cross‑team...RiskWork at officeRelocation package- Scale Labs in San Francisco is seeking a Research Scientist focused on Agent Robustness to advance safe, aligned AI. You will tackle fundamental challenges in evaluating agent capabilities, safety, and risk, and help design benchmarks and protocols. You will design harnesses...Risk
- ...owning a holistic approach to identifying and modeling risks from AI systems, ensuring robust evaluation frameworks. You will work closely with technical and... ...safety challenges. The ideal candidate understands frontier AI risks and has experience in threat modeling and...Risk
$200k - $370k
...Manufacturing Co in San Francisco is seeking exceptional research engineers to explore AI safety concerns. Candidates will work on identifying risks and designing evaluations for frontier AI models. The ideal candidate will have a strong technical background and the...Risk$350k
...interpretable, and steerable AI systems. We want AI to... ...yourself as both a scientist and an engineer. As a... ...safety, with a focus on risks from powerful future... ...Fine-Tuning, and the Frontier Red Team. Our blog [... ...coordination with third-party evaluators. * Safeguards Research...RiskWork at officeVisa sponsorshipFlexible hours$400 per month
Obsidian is collaborating with a leading AI research lab for a project involving Frontier Code Agents. Contributors will focus on evaluating and improving AI coding models relevant to fraud and risk engineering. This role requires 2+ years in relevant fields and the ability...Risk- OpenAI is looking for a Researcher for Frontier Evals & Environments to help develop model environments that drive progress towards safe... ...role, you will lead efforts to define research programs and evaluate results, influencing product development. You will collaborate...
- ...Technologies is building quantum-accelerated AI servers to exponentially speed up... ...profound global impact. About the Role Frontier AI is moving toward scientific reasoning... ...next era. We are looking for a Research Scientist who can help define Quantum AI: not just...Casual workVisa sponsorship
- A leading AI evaluation firm based in San Francisco seeks a Machine Learning Scientist to foster understanding of AI model performance. You'll engage in designing and analyzing comprehensive experiments while collaborating across teams. Applicants should possess a PhD...
- ...On-site Department Technical About the Role Generative AI is transforming what's computationally possible—but it's... ...offers a path through these bottlenecks. As an ML Research Scientist, you'll work at the frontier of generative modeling and quantum acceleration,...Full timeCasual workVisa sponsorship
$80 - $150 per hour
...consulting firm is seeking Senior Behavioral Health Experts to work part-time and remotely on frontier AI research projects. You will be responsible for designing evaluations and testing AI systems in critical mental health contexts. The ideal candidate has over 5 years...RiskHourly payPart timeRemote work- Anthropic is seeking a dedicated individual to build evaluation infrastructure that enhances trust in AI systems. You'll work on constructing real-world data sets and analyzing agent performance to inform risk management. The ideal candidate will have strong proficiency...Risk
- ...accepted until further notice. Meet the Team At Foundation AI, we are leading frontier AI research across Cisco. Our mission is to advance the... ...training and post-training methods, inference optimization, evaluation techniques, and data pipelines. Together, we are...
$70 - $100 per hour
Mercor is seeking a STEM Computational Scientific Software & Evaluation Design specialist to tackle complex computational problems remotely... ...scientific software libraries. You will design challenges to assess AI models and collaborate with research teams to refine these...Remote jobFlexible hours- A leading AI research organization is seeking a Researcher to focus on cybersecurity risks. The role involves designing and implementing mitigation strategies across products, ensuring effective safeguards as AI capabilities develop. Ideal candidates should have deep learning...Risk
$192.6k - $344.85k
## AI Research Manager/Scientist, Reinforcement LearningApplylocations: San Francisco... ...training workflows### ### Evaluation, Alignment & *Model*... ...explicit articulation of known risks and trade-offs### ###... ...in an AI research lab or frontier *model* organization* Background...RiskRemote work- OpenAI is seeking a Researcher for cybersecurity risks in San Francisco. You will design and implement a mitigation stack for reducing severe cyber misuse across OpenAI’s products. The role requires strong technical depth and close collaboration with cross-functional teams...Risk
- Member of Technical Staff - Research Scientist Patronus AI is a frontier lab developing simulation research... ...and most influential research in AI evaluation like FinanceBench , Lynx , SimpleSafetyTests... ...clearly and proactively, flagging risks, blockers, or timeline changes early...Risk
$150k - $300k
About Patronus AI Patronus AI is a frontier lab developing simulation research and... ...research in AI evaluation like FinanceBench, Lynx, SimpleSafetyTests... ...As a Research Scientist at Patronus AI, you will own... ...and proactively, flagging risks, blockers, or timeline changes...RiskWork at office$207k - $230k
...OpenAI’s Frontier Evals team designs and builds evaluations that measure the capabilities, limitations, and emerging behaviors... ...plans, crisp milestones, owners, risks, and decision points. Are... ...offered. About OpenAI OpenAI is an AI research and deployment company dedicated...RiskWork at officeRelocation package$76k - $125.3k
...issues, provide advice and assistance managing risks across tax compliance and/or advisory... ...and other business analysis Compile and evaluate moderately complex data, computations, documentation... ...in capital markets. Enabled by data, AI and advanced technology, EY teams help...RiskSummer holidayFlexible hours$295k
...which is focused on mitigating AI threats to global security... ...the evolving capabilities of frontier AI systems. Mitigation. Keeping... ...Researcher for cybersecurity risks, you will help design and implement... ...and new model capabilities. Evaluate technical trade‑offs within...Risk$170k - $220k
Introduction The Center for AI Safety is a research and field-building... ...catastrophic and existential risks from artificial intelligence... ...research. As a research scientist, you will pursue a variety of... ...large-scale transformers and evaluating them under different data domains...RiskWork at office
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Frontier AI Risk Evaluations Scientist. Be the first to apply!

