Anthropic Fellows Program, ML Systems & Reinforcement Learning
Anthropic
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. This page is specific to one of the Anthropic Fellows Workstreams, see also the main Anthropic Fellows posting . Anthropic Fellows Program overview The Anthropic Fellows Program is designed to foster AI research and engineering talent. We provide funding and mentorship to promising technical talent - regardless of previous experience. Fellows will primarily use external infrastructure (e.g. open-source models, public APIs) to work on an empirical project aligned with our research priorities, with the goal of producing a public output (e.g. a paper submission). In one of our earlier cohorts, over 80% of fellows produced papers.We run multiple cohorts of Fellows each year and review applications on a rolling basis. What to expect 4 months of full-time research Direct mentorship from Anthropic researchers Access to a shared workspace (in either Berkeley, California or London, UK) Connection to the broader AI safety and security research community Weekly stipend of 3,850 USD / 2,310 GBP / 4,300 CAD + benefits (these vary by country) Funding for compute (~$15k/month) and other research expenses Interview process The interview process will include an initial application & reference check, technical assessments & interviews, and a research discussion. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Compensation The expected base stipend for this role is 3,850 USD / 2,310 GBP / 4,300 CAD per week, with an expectation of 40 hours per week for 4 months (with possible extension). Fellows workstreams Due to the success of the Anthropic Fellows for AI Safety Research program, we are now expanding it across teams at Anthropic. We expect there to be significant overlap in the types of skills and responsibilities across the roles and will by default consider candidates for all the workstreams. Some of the workstreams may include unique assessment steps; we therefore ask you for workstream preferences in the application . You can see an overview of the current workstreams below: AI Safety Fellows AI Security Fellows ML Systems & Performance Fellows Reinforcement Learning Fellows Economics & Societal Impacts Fellows This page is specific to one of the Anthropic Fellows Workstreams, see also the main Anthropic Fellows posting . Across the workstreams, you may be a good fit if you: Are motivated by making sure AI is safe and beneficial for society as a whole Are excited to transition into empirical AI research and would be interested in a full-time role at Anthropic Have a strong technical background in computer science, mathematics, or physics Thrive in fast-paced, collaborative environments Can implement ideas quickly and communicate clearly Strong candidates may also have: Strong background in a discipline relevant to a specific Fellows workstream (e.g. economics, social sciences, or cybersecurity) Experience in areas of research or engineering related to their workstream Candidates must be: Fluent in Python programming Available to work full-time on the Fellows program ML Systems & Performance Fellows Mentors, research areas, & past projects Fellows will undergo a project selection & mentor matching process. Potential mentors include: Alwin Peng Zygi Straznickas Note: You may research mentors' prior work, but all applications must go through the official form, not the mentors. For a past example of an engineering-heavy project, see: AI agents find $4.6M in blockchain smart contract exploits Projects in this workstream may include: Building a CPU simulator for accelerator workloads Adding backends for different accelerators on an open source project Building on demand infrastructure for other infrastructure heavy fellows projects Building complex synthetic data or environment pipelines Unique candidate criteria You might be a particularly great fit for this workstream if you: Have strong software engineering skills with experience building complex ML systems Can balance research exploration with engineering rigor and operational reliability Enjoy collaborating across research and engineering disciplines Are comfortable working with large-scale distributed systems and high-performance computing (e.g. in trading) Have experience with training, fine-tuning, or evaluating large language models Are adept at analyzing and debugging model training processes Reinforcement Learning Fellows Mentors, research areas, & past projects Fellows will undergo a project selection & mentor matching process. Potential research areas and mentors include: Ruhua Jiang Kaidi Cao Sunny Duan David Brandfonbrener Colt Steele Dino Distefano Will Williams Projects in this workstream may include: Building model-based tools to better understand AI training data and improve training data quality A research project to better understand generalization Creating RL environments to improve Claude models at capabilities that are within your domain of expertise Building RL environments for safety-related tasks Conducting research and implementing solutions in areas such as RL algorithms Unique candidate criteria You might be a particularly great fit for this workstream if you: Have strong software engineering skills with experience building complex ML systems Can balance research exploration with engineering rigor and operational reliability Enjoy collaborating across research and engineering disciplines Are comfortable working with large-scale distributed systems and high-performance computing Have experience with training, fine-tuning, or evaluating large language models Are adept at analyzing and debugging model training processes Logistics Logistics Requirements: To participate in the Fellows program, you must have work authorization in the US, UK, or Canada and be located in that country during the program. Workspace Locations: We have designated shared workspaces in London and Berkeley where fellows will work from and mentors will visit. We are also open to remote fellows in the UK, US, or Canada . We will ask you about your availability to work from Berkeley or London (full- or part-time) during the program. Visa Sponsorship: We are not currently able to sponsor visas for fellows. To participate in the Fellows program, you need to have or independently obtain full-time work authorization in the UK, the US, or Canada. Program Duration: The program runs for 4 months, full-time. If you can't commit to the full duration, please still apply and note your constraints in the application. We review these requests on a case-by-case basis. Please note: We do not guarantee that we will make any full-time offers to fellows. However, strong performance during the program may indicate that a Fellow would be a good fit for full-time roles at Anthropic. In previous cohorts, 25-50% of fellows received a full-time offer, and we’ve supported many more to go on to do great work on AI safety and security at other organizations. Applications and interviews are managed by Constellation , our recruiting partner. We will remove application references in refinement. All applicants currently use the same application portal but we are working to separate applications for safety/security and capabilities focused projects in future rounds. Apply here The below are Anthropic's policies for full time roles. These do NOT apply to the Fellows Program. Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you View email address on click.appcast.io addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process. #J-18808-Ljbffr Anthropic
$15k
...We apply state-of-the-art AI/ML techniques to construct our... ...an experienced and creative reinforcement learning researcher to our growing ML... ...venues on machine learning, systems, and theory, and we meet regularly... ...the Voleon Referral Bonus Program.Equal Opportunity...SuggestedLocal areaImmediate startRelocationWork visa- ...Responsibilities As a senior Machine Learning Systems Engineer on the Search Platform team, you... ...search. Own end-to-end delivery of ML components from experimentation through... ...explainable, and competitive compensation programs. We follow consistent hiring practices and...SuggestedWork at officeLocal area
- ...Responsibilities As a Principal Machine Learning Systems Engineer on the Search Platform team,... ...priorities across the Search Platform, ML Platform, AI Gateway, and Rovo product teams to keep critical cross-functional programs on track.Agentic Search & Retrieval ArchitectureOwn...SuggestedWork at officeLocal areaShift work
- ...legal entity.ResponsibilitiesAs a Principal ML System Engineer in the Rovo & AI Engineering... ...understanding of Machine Learning projects lifecycleExperience leading and... ...explainable, and competitive compensation programs. To support this goal, the baseline of our...SuggestedWork at officeLocal area
$365k
About Anthropic Anthropic’s mission is to create reliable... ..., and steerable AI systems. We want AI to be safe... ...the Role As a Technical Program Manager for Product, you... ...engage meaningfully with ML researchers and engineers... ...in AI Safety, and Learning from Human Preferences....SuggestedWork at officeVisa sponsorshipFlexible hoursShift work$165k - $230k
...for exceptional Machine Learning Engineers focused on... ...generative AI, and advertising systems—building the models... ..., and production ML systems. You'll... ...preference optimization, reinforcement learning, or related approaches... ...’s stock option program and share in...Full timeWork at officeRemote workWorldwide3 days per week- Anthropic is seeking Fellows to advance AI safety and research in a full-time program. Fellows will work on empirical AI projects with mentorship from Anthropic researchers, using external infrastructure and公开 outputs. The program provides weekly stipends and access to...Full time
- ...network. The Drive Machine Learning team builds the prediction and intelligence systems that power this... ...prediction accuracy.Apply reinforcement learning and optimization... ...a diverse set of ML techniques—from neural... ...assistance, and a mental health program, among others.To learn...Hourly payWork at officeLocal areaRemote workRelocationFlexible hours
- ...progress, we are implementing reinforcement learning tasks where humans currently... ...training and testing our systems. Rigorously document your experiments... ...Qualifications PhD in AI, ML, computer science, cognitive... ...and building products (e.g. Anthropic). We are completely focused...Work experience placementWork at office2 days per week
$200k - $400k
...understanding and designing AI systems that people can trust.... ...Menlo, Lightspeed and Anthropic and work with customers... ...re looking for Machine Learning Engineers to help build... ...years of experience in ML infra, research engineering, or systems programming. ~ Comfort working...Full time- ...optimization, machine learning, and causal inference.... ...turning ideas to scalable systems. This role... ...of machine learning (ML) solutions for complex... ...more object-oriented programming languages (e.g. Python... ...techniques, including reinforcement learning (RL), Bayesian...Full timeWork at officeWork from homeWorldwideFlexible hours
$365k
About Anthropic Anthropic’s mission is to create reliable... ..., and steerable AI systems. We want AI to be safe... ...development. As a Technical Program Manager for Research,... ...You Have a background in ML research or engineering... ...in AI Safety, and Learning from Human Preferences....Work at officeVisa sponsorshipFlexible hoursShift work$15k
...state-of-the-art AI and machine learning techniques to real-world... ...at the frontier of applying AI/ML to investment management. We have... ...and ensure that the resulting systems are performant, reliable, and... ...review the Voleon Referral Bonus Program.Equal Opportunity EmployerThe...Work at officeLocal area$190.2k - $345.65k
...development using Machine Learning and Agentic AI systems to help design and scale the... ...engineering, applied ML, and AI agent orchestration... ...privacy). Experience with reinforcement learning, multi-agent systems... ..., comprehensive benefits programs, the stories we tell, the...Full timeTemporary workLocal areaWorldwide$229k - $343k
...themselves, live in the moment, learn about the world, and... ...retrieval, and ranking systems that deliver the most... ...for engineering and ML excellence, developing... ...build our culture faster, reinforce our values, and serve... ...mental health support programs, and compensation packages...Full timeLive inWork at officeLocal area$264.8k - $331k
...world. The Enterprise ML Research Lab works on the... ...technologies to optimize our ML system. Your customer will be other MLREs... ..., retirement benefits, a learning and development stipend, and generous... ...our internal policies and programs designed to protect personal data...Full time- ...developing next-generation AI systems designed to simplify... ...and implement machine learning algorithms to optimize... ..., and modify computer programs to apply machine... ...Familiarity with OpenAI, Anthropic, or Hugging Face Transformers... .... Experience with ML and AI algorithms to model...Remote work
$209k - $313k
...themselves, live in the moment, learn about the world, and... ...modern ML techniques to solve large... ...and applying evolving AI systems and tools to remain at... ...build our culture faster, reinforce our values, and serve our... ...mental health support programs, and compensation packages...Full timeLive inWork at officeLocal area$162.8k - $203.5k
...lives warrants modern ML utilizing peta-byte scale... ...motivated Machine Learning Engineers work on these... ...building reliable ML systems, and is excited about... ...Python, Golang, or other programming language ~ Excellent... ...systems, reinforcement learning, and multi-armed...Hourly payFull timeWork experience placementWork at officeLocal area3 days per week$165k - $230k
...Responsibilities Build and improve ML systems for advertising ranking,... ...experience building machine learning systems for advertising.... ...preference optimization, or reinforcement learning. Strong software... ...in the company stock option program. Comprehensive benefits...Full timeWork at officeRemote work3 days per week- ...the loop of how those systems improve. This is a... ...Google DeepMind, OpenAI, Anthropic, Meta... ...Microsoft AI. Machine Learning Engineer, Quality Intelligence... ...uses our datasets and reinforcement learning environments... ...Intelligence to build the ML systems behind how we...Full time
$264.8k - $331k
...Machine Learning Systems Research Engineer, Agent Post-training - Enterprise GenAI AI is becoming... ...enterprises around the world. The Enterprise ML Research Lab works on the front lines of... ...with our internal policies and programs designed to protect personal data. Please...Full timeContract workFor contractorsFor subcontractorWork at office- ...across hospital and health systems, pharmacies and payors... ...revenue to fund programs for their in-need patient... ...looking for a Machine Learning Engineer to design, build... ...deploy production-grade ML systems that power the... ...and LLM APIs (OpenAI, Anthropic, etc.) Why You'll Love...Full timeWork at officeRemote workFlexible hours2 days per week
$229k - $343k
...themselves, live in the moment, learn about the world, and... ...Search ranking systems. In this role, you will... ...mentor engineers working on ML ranking systemsStay... ...experimentationStrong programming skills in Python, C++,... ...build our culture faster, reinforce our values, and serve our...Full timeLive inWork at officeLocal area$250k - $300k
...infrastructure, proprietary software and systems and data science expertise.... ..., and revenue goals into ML system requirements.Mentor... ...in Data Science, Machine Learning Engineering, or Software Engineering... ...-time feature generation and reinforcement learning is a plus.Experience...Full timeWork at officeRemote workFlexible hours$137.1k - $201.6k
...artificial intelligence and advanced ML, deep learning techniques to power decision-making in... ...build, optimize and scale large-scale ML systems within the Ads Delivery funnel.... ...statistics, and data modeling. Strong programming skills in Python, Java, or C++, and experience...Hourly payWork at officeLocal areaRemote workFlexible hours$150k - $200k
...MS degree in Computer Science, Machine Learning, Robotics, or equivalent technical discipline... ...in machine learning fundamentals, reinforcement learning, and associated frameworks... ...track record developing and deploying ML systems from research through production implementation...Full time- ...pretraining and scaling , post-training and Reinforcement Learning , sandbox environments for evaluation... ...skills (writing robust, performant systems) Experience with training or serving... ...Owning end-to-end production ML systems with monitoring and reliability...Full timeFlexible hours
- ...thousands of customers — including Anthropic, Notion, Google, and Ramp — go... ...Glassdoor page! Machine Learning Engineer @ Clay Clay's... ...someone uses it. This means data, ML, and AI are at the heart of... ...at the heart of the product: systems that learn a customer's business...Full time
- Principal or Senior Principal, Anthropic AI SolutionsAI Systems & Platforms | Anthropic... ...challenges. We help clients learn from their data, create... ...Lead or co-lead AI enablement programs: designing training curricula... ...and delivering AI/ML or large language model-based...Temporary workWork at officeLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Anthropic Fellows Program, ML Systems & Reinforcement Learning. Be the first to apply!
- machine learning scientist Berkeley, CA
- machine learning remote Berkeley, CA
- machine learning researcher Berkeley, CA
- machine learning Berkeley, CA
- machine learning research scientist Berkeley, CA
- artificial intelligence - machine learning intern Berkeley, CA
- data engineer machine learning Berkeley, CA
- machine learning drug discovery
- machine learning scientist
- machine learning artificial intelligence



