AI Security & Control Researcher
apolloresearch
THE OPPORTUNITY THE OPPORTUNITY Apollo Research works with most frontier AI companies (OpenAI, Anthropic, Google, Meta, Thinking Machines and others) to test their models before deployment and collaborate on fundamental scheming research. Our coding agent security product, Watcher, is deployed in production and monitors billions of agent tokens per month across engineering teams at agent-building scale-ups and enterprises. We are looking for a security & control expert to help us design better threat models and control protocols against AI adversaries, and improve the effectiveness and security of Watcher. This is truly a "start-up role" in the sense that you have significant say in shaping the direction of the role. This is an individual contributor role but could lead to management responsibilities eventually, if desired. ABOUT THE TEAM The product team consists of research scientists: Victor Gillioz, Monika Jotautaitė, Dmitrii Volkov; product engineers: Jeremy Neiman, Zak Walters, Zen van Riel, Srdjan Miletic and Gustavo Bicalho; and our GTM lead: Kyle Dai. Marius Hobbhahn (CEO) advises the team. Furthermore you will interact with our other SWEs and researchers, since we intend to be "our own customer" by using our products internally for our research work. You can find our full team here. ABOUT APOLLO RESEARCH The rapid rise in AI capabilities offers tremendous opportunities, but also presents significant risks. At Apollo Research, we're primarily concerned with risks from Loss of Control, i.e. risks coming from the model itself rather than e.g. humans misusing the AI. We're particularly concerned with deceptive alignment / scheming, a phenomenon where a model appears to be aligned but is, in fact, misaligned and capable of evading human oversight. We work on the science of scheming, detection of scheming (e.g. building evaluations), and scheming mitigations (e.g. anti-scheming). We also work on control and monitoring research (see our scalable monitoring agenda). We work closely with many frontier AI companies, such as OpenAI, Anthropic, Google, Meta, Thinking Machines and others, e.g. to test their models and collaborate on the science of scheming. At Apollo, we aim for a culture that emphasizes truth-seeking, being goal-oriented, giving and receiving constructive feedback, and being friendly and helpful. If you're interested in more details about what it's like working at Apollo, you can find more information here. We also build a coding agent security product called Watcher that secures agent deployments in companies. Our goal is to reduce the probability of catastrophic incidents by securing coding agents, learning about their real-world risks, and publishing our research on how to build these control systems most effectively. Equality Statement: Apollo Research is an Equal Opportunity Employer. We value diversity and are committed to providing equal opportunities to all, regardless of age, disability, gender reassignment, marriage and civil partnership, pregnancy and maternity, race, religion or belief, sex, or sexual orientation. About the interview process: Our multi-stage process includes a screening interview, a take-home test (3 hours), 3 technical interviews, and a final interview with Marius (CEO). There are no leetcode-style general coding interviews. You may use AI tools on the take-home; we judge the result the way we'd judge any contributor's work, so you are responsible for the quality of everything you submit. If you want to prepare, we suggest building simple monitors for coding agents and running them on your own Claude Code / Cursor / Codex / etc. traffic. Your Privacy and Fairness in Our Recruitment Process: We are committed to protecting your data, ensuring fairness, and adhering to workplace fairness principles in our recruitment process. To enhance hiring efficiency, we use AI-powered tools to assist with tasks such as resume screening. These tools are designed and deployed in compliance with internationally recognized AI governance frameworks. All resumes are screened by a human and final hiring decisions are made by our team. If you have questions about how your data is processed or wish to report concerns about fairness, please contact us at View email address on click.appcast.io. #J-18808-Ljbffr apolloresearch
- Apollo Research in the United States is seeking a security & control expert to help design threat models and strengthen Watcher against AI adversaries. This is an individual contributor role with potential path to leadership, depending on interest and performance. You will...Suggested
- AI agents are changing how enterprises operate. Companies want to move fast with... ...safely. CodeIntegrity is the platform security teams use to control MCPs and agents , built for companies... ...and turn that work into clear security research. Research agent security risks across...Suggested
$293k - $405k
...the team Preparedness is a critical Safety Research team at OpenAI, which is focused on mitigating AI threats to global security that could scale to an extreme level of severity... ...could compromise OpenAI. Design security controls - focusing on measures with long lead times...Suggested- 0Labs in San Francisco/London is seeking an AI security researcher/engineer to design and build highly realistic cybersecurity environments.... ...models and study rogue model deployments to mitigate loss-of-control scenarios. Join a small, fast-moving team committed to...SuggestedRemote workFlexible hours
- ...shift from human operators to agents. AI becomes the most capable attacker in the... ...misaligned AI models. Back to careers AI security researcher San Francisco / London Full-time... ...internal deployments that lead to loss-of-control scenarios. Who We Are Looking For We are...SuggestedFull timeRemote workFlexible hoursShift work
- CodeIntegrity is building the platform security layer to control MCPs and agents across enterprises. We are a small team... ...notable investors, focused on making enterprise AI safe to use. We seek a security-focused researcher who can read code, prototype, and analyze how...
- AegisAI is seeking a Senior Threat Intelligence Analyst to lead investigations into phishing, BEC, and malware campaigns, and publish high-quality analysis to advance the cybersecurity community. You will collaborate with engineering and data science to improve detection...
- Anthropic is seeking a Capabilities Researcher to join the team building Claude Security. You will identify security capabilities in frontier models, measure performance, and ensure usability for non-experts. The role blends research with product impact, offering latitude...
- OpenAI is seeking a Researcher for Frontier Cybersecurity Risks to design and implement an end-to-end mitigation stack that reduces severe cyber misuse across OpenAI’s products. This role requires deep technical depth and close collaboration with risk, policy, product,...
- Anthropic is seeking a Capabilities Researcher to join the Claude Security team in San Francisco. You will identify which security capabilities in frontier... ...role on a product team offers broad latitude to explore AI security capabilities, design rigorous evaluations, and...
$293k - $405k
OpenAI is seeking a Security Role focused on preparing for potential threats from advanced AI systems. This position requires a deep technical understanding of security measures and active engagement with stakeholders to design and evaluate defense systems. Ideal candidates...$150k - $250k
About Distyl AI Distyl is an applied AI technology company partnering with the world... ..., and global social organizations.We research and deploy technologies that power AI-native... ...and robustness, capability and controllability. Their work informs how Distyl leverages...Work at office3 days per week$295k
...evolving capabilities of frontier AI systems.Mitigation. Keeping... ..., alignment tools, and security measures on track to adequately... ...RoleWe are seeking exceptional researchers who can push the frontier of... ...domains like interpretability, control, and alignment to ensure the...Work at officeLocal areaFlexible hours$216k - $270k
Scale Labs, Research Scientist — AI Controls and MonitoringAs the leading data and evaluation partner for frontier AI companies, Scale plays an integral... ...you’d have:Commitment to our mission of promoting safe, secure, and trustworthy AI deployments in the industry as...Full time$192k - $260k
...426R282Databricks is building the world's best and most secure platform for data and AI. We innovate and deploy industry-leading solutions in security... ...protected characteristics.ComplianceIf access to export-controlled technology or source code is required for performance of...Local areaWorldwide- Verita AI works with leading AI companies to identify model gaps and build the human... ...About the Role We are hiring an Applied AI Researcher to work directly with clients on model... ..., rubrics, expert workflows, and quality controls needed to address them. You will then work...
$192k - $260k
...world's toughest problems, from security threat detection to cancer... ...push the boundaries of data and AI technology, while... ...Physics, Economics, Operational Research or Engineering)Pay Range TransparencyDatabricks... ...access to export-controlled technology or source code is...Work at officeLocal areaWorldwide$84.13 - $91.34 per hour
AI Researcher - Efficient AI (Contractor) Step into the innovative world of LG Electronics. As a global leader in technology, LG Electronics... ...of information unless such collection displays a valid OMB control number. This survey should take about 5 minutes to complete....Full timeContract workTemporary workFor contractorsLocal areaImmediate start- Morpheus Talent Solutions seeks a founding Applied AI Researcher focused on model evaluation and data strategy for an early-stage, profitable... ...failures, and defining data schemas, rubrics, and quality controls in collaboration with engineers to scale programs. You will...Remote job
- AI Researcher (Computer Vision/Multimodal/Generative AI) About the Role We are hiring ML Researchers to develop novel approaches that... ...models are not designed for photorealistic human representation, controllable try‑on, or real‑world deployment constraints. You will...
$200,000 - $350,000 per day
Applied AI Researcher - Model Evaluation & Data Strategy San Francisco (in-person preferred; open to remote across US, UK, Australia, and... ...define the datasets, rubrics, reward signals, and quality controls that move performance - then work with engineers to turn them...Remote workVisa sponsorship- OpenAI in San Francisco is seeking exceptional researchers to push the frontier of safety mitigations for deployed models... ...and applying methods from interpretability, control, and alignment to ensure safe AI systems. The role requires strong technical depth and...
- ...company that uses real neurons to improve AI models. We study how biological neural... ...biology, computational neuroscience, AI research and software engineering to develop new algorithms... ...foundations for policy learning, control, and real‑world deployment. This is a senior...
- ...revolutionizing software development with AI-powered formal verification. We've... ...intended while proactively identifying bugs and security vulnerabilities. Our novel foundation... ...About the role Join our team as an AI Researcher and help us push the boundaries of what'...Contract work
- OpenAI is seeking an experienced security researcher to help mitigate AI threats and safeguard systems as AI agents become more capable. The role focuses... ...will identify attack paths, design long-lead security controls, and stress-test defenses with AI agent evaluations and...
$216k - $270k
Research Scientist, AI Controls and Monitoring Scale Labs, Research Scientist - AI Controls and Monitoring As the leading data and evaluation partner... ...’d have: Commitment to our mission of promoting safe, secure, and trustworthy AI deployments in the industry as...Full time- ...predicting the future and identifying actions to alter it. We seek researchers to build the planning layer on top of the LPM, conditioning... ...role involves research and implementation across planning, control, and decision-making with a focus on interventional causality...
- ...company that uses real neurons to improve AI models. We study how biological neural... ...biology, computational neuroscience, AI research and software engineering to develop new algorithms... ...representations, stable rollouts, and control‑oriented prediction Improve long-horizon...
- Scale Labs is seeking a Research Scientist focused on AI Controls and Monitoring to design methods and experiments ensuring alignment of advanced AI models in high-stakes environments. You will build real-time monitoring, layered control, and red-team simulations while...
- Perplexity's Secure Intelligence Institute seeks researchers and engineers to advance security and privacy in frontier AI systems. You will conduct original research and translate findings into practical protections for millions of users. You will develop threat models...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Security & Control Researcher. Be the first to apply!
- field researcher San Francisco, CA
- product researcher San Francisco, CA
- security researcher San Francisco, CA
- data collection researcher San Francisco, CA
- machine learning researcher San Francisco, CA
- court researcher San Francisco, CA
- researcher San Francisco, CA
- senior researcher San Francisco, CA
- independent researcher San Francisco, CA
- remote researcher San Francisco, CA
