Research Engineer, Pre-training [Remote]
$340k - $425kAnthropic
- Remote job
About Anthropic
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.
Anthropic is at the forefront of AI research, dedicated to developing safe, ethical, and powerful artificial intelligence. Our mission is to ensure that transformative AI systems are aligned with human interests. We are seeking a Research Engineer to join our Pre-training team, responsible for developing the next generation of large language models. In this role, you will work at the intersection of cutting-edge research and practical engineering, contributing to the development of safe, steerable, and trustworthy AI systems.
Key Responsibilities:
- Conduct research and implement solutions in areas such as model architecture, algorithms, data processing, and optimizer development
- Independently lead small research projects while collaborating with team members on larger initiatives
- Design, run, and analyze scientific experiments to advance our understanding of large language models
- Optimize and scale our training infrastructure to improve efficiency and reliability
- Develop and improve dev tooling to enhance team productivity
- Contribute to the entire stack, from low-level optimizations to high-level model design
Qualifications:
- Advanced degree (MS or PhD) in Computer Science, Machine Learning, or a related field
- Strong software engineering skills with a proven track record of building complex systems
- Expertise in Python and experience with deep learning frameworks (PyTorch preferred)
- Familiarity with large-scale machine learning, particularly in the context of language models
- Ability to balance research goals with practical engineering constraints
- Strong problem-solving skills and a results-oriented mindset
- Excellent communication skills and ability to work in a collaborative environment
- Care about the societal impacts of your work
Preferred Experience:
- Work on high-performance, large-scale ML systems
- Familiarity with GPUs, Kubernetes, and OS internals
- Experience with language modeling using transformer architectures
- Knowledge of reinforcement learning techniques
- Background in large-scale ETL processes
You'll thrive in this role if you:
- Have significant software engineering experience
- Are results-oriented with a bias towards flexibility and impact
- Willingly take on tasks outside your job description to support the team
- Enjoy pair programming and collaborative work
- Are eager to learn more about machine learning research
- Are enthusiastic to work at an organization that functions as a single, cohesive team pursuing large-scale AI research projects
- Are working to align state of the art models with human values and preferences, understand and interpret deep neural networks, or develop new models to support these areas of research
- View research and engineering as two sides of the same coin, and seek to understand all aspects of our research program as well as possible, to maximize the impact of your insights
- Have ambitious goals for AI safety and general progress in the next few years, and you’re working to create the best outcomes over the long-term.
Sample Projects:
- Optimizing the throughput of novel attention mechanisms
- Comparing compute efficiency of different Transformer variants
- Preparing large-scale datasets for efficient model consumption
- Scaling distributed training jobs to thousands of GPUs
- Designing fault tolerance strategies for our training infrastructure
- Creating interactive visualizations of model internals, such as attention patterns
At Anthropic, we are committed to fostering a diverse and inclusive workplace. We strongly encourage applications from candidates of all backgrounds, including those from underrepresented groups in tech.
If you're excited about pushing the boundaries of AI while prioritizing safety and ethics, we want to hear from you!
The expected salary range for this position is:
Annual Salary:
$340,000—$425,000 USD
Logistics
Education requirements: We require at least a Bachelor's degree in a related field or equivalent experience.
Location-based hybrid policy:
Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.
Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.
We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team.
How we're different
We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills.
The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Come work with us!
Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process
$350k
A leading AI research company is seeking a Pre-training Research Engineer to advance large language models. You will be engaged in research, implementing solutions, and enhancing training infrastructure while collaborating with other experts. Candidates must possess an...TrainingWork at office$350k
...for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. Pre-training Research Engineer Anthropic is at the forefront of AI research, dedicated...TrainingWork at officeVisa sponsorshipFlexible hours$180k - $200k
About the role We are seeking exceptional Senior Research Engineers to join our mission-critical team building the world's best oncology... ...inform our clinical-stage programs. You'll participate in both pre-training and post-training of our foundation models, requiring deep...TrainingWork at office3 days per week$111.5k - $207.5k
...national security. Job Title: Lead, Security Research Engineer Job Code: 35904 Job Location: Remote... ...or contractors of the company. Training, management and provision of guidance to... ...maintains a drug‑free workplace and performs pre‑employment substance abuse testing and...TrainingFor contractorsLocal areaRemote workFlexible hours- A cutting-edge tech company in New York is seeking a Research Engineer to build and scale systems for their AI video generation models. This role involves training and optimizing large-scale models, improving efficiency, and translating research into production-ready systems...Training
$165k - $260k
A leading financial technology company in New York seeks a Senior NLP Engineer to work on innovative AI-driven solutions. This role involves designing, training, and evaluating NLP models, collaborating across teams, and publishing findings. Ideal candidates will have...Training$146.7k - $214.8k
...in of Cisco's products. If you enjoy vulnerability research, crash analysis, reverse engineering, and researching new techniques and writing tools to... ...experience, qualifications, education, certifications, and/or training. The full salary range for certain locations is...TrainingFull timeTemporary workLocal areaRemote workFlexible hours- ...located in Union Square). About the Role Mirage is seeking a Research Engineer to build and scale the systems powering our video generation... ...ultra‑low latency, real‑time generation. Responsibilities Train and optimize large‑scale video and multimodal models Improve...TrainingFull timeLocal areaNight shift
- About the Role Mirage is seeking a Research Engineer to build and scale systems for training and deploying large language models for multimodal creative tasks, specifically focusing on video analysis pipelines. You’ll work closely with researchers to turn ideas into high...TrainingFull timeLocal areaNight shift
- Job Title: Research Engineer Contract Duration: 1 year, possible extension Work Arrangement: Remote, ideally EST, buy anywhere in NORAM... ...deep learning libraries that support large-scale distributed training, open sourcing high quality code and reproducible results for...TrainingContract workWork experience placementRemote work
- About Basis Basis is a nonprofit applied AI research organization with two goals: Understand... ...human value. About the Role Research engineers support Basis' mission by translating research... ...new techniques, building large-scale training systems, or spanning multiple levels of...TrainingFull timeFlexible hours
$300k - $405k
...a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build... ...and evaluations, delivering your work into production training runs, and collaborating with other researchers, engineers,...TrainingWork at officeVisa sponsorshipFlexible hours- Bloomberg’s Engineering AI department has 400+ AI practitioners building highly sought after... ...is built on technology that makes news, research, financial data, and analytics on over 3... ...methods for LLMs, efficient methods for training, multimodal models, learning from feedback...TrainingWork experience placement
- About Basis Basis is a nonprofit applied AI research organization with two mutually... ...values first. About the Role Research Engineers in Operations at Basis build the internal... ...scratch You build systems and workflows for training large models distributed across many machines...TrainingFull timeContract workWork at office
$174k - $252k
Research Engineer, Pretraining, DeepMind DeepMind • New York, NY, USA Requirements Bachelor’s degree in Computer Science, Mathematics, Applied... ...conducting applied research to improve the quality and training/serving efficiency of large transformer-based models....TrainingFull time- ...Achira, we are building a team of world-class scientists, ML researchers, and engineers to work together to move beyond the beaten path in drug... ...participating with research teams to scale our ML systems, train and evaluate models, and engineer scientific prototypes into...TrainingWork at office
$200k
...Optiver is a seeking a Machine Learning Research Engineer to join our team, focusing on a pivotal AI initiative. This role would offer the... ...significant impact across Machine Learning infrastructure, training, and inference challenges to advance our futures trading strategies...TrainingWork at office$189.6k - $237k
...About This Role Lead applied ML engineering on Scale's Applied ML team, powering data... ...behaviors, scale human expertise, and drive research into real-world agent reliability... ...performance, and relevant education or training. Scale employees in eligible roles are also...TrainingFull time$124k - $144k
...Position Summary The Senior Research Network Engineer plays a critical role in designing, implementing, managing, and monitoring the dedicated... ...of the position, candidate's work experience, education/training, key skills, internal peer equity, as well as, market and...TrainingWork experience placementRemote work$15k
...the most profitable market makers on the street. As a Senior Research Engineer embedded within our first quant strategy team, you will own... .... You will be expected to help optimize machine learning training and testing environments, architect and maintain data pipelines...TrainingLocal areaNight shift$250k - $350k
...Machine Learning Research Engineer, Agents - Enterprise GenAI AI is becoming vitally important in every function of our society. At Scale... ...we are doubling down on building out state of the art post-training algorithms to reach the performance necessary for complex...TrainingFull time$245k - $330k
...team at the NBA is seeking an experienced applied scientist / research engineer with a strong foundation in Computer Vision and Machine... ...system, from building sensing pipelines to scalable ML data, training, modeling and evaluation pipelines. This team sits...TrainingTemporary workLocal areaRemote work1 day per week- ...investors . I’ll share more once we meet. About the Role As an ML Research Engineer at Maple, you'll be a part of our core product team... ...devices, minimizing latency. Manage rapid experimentation, training, and highly optimized production inference. Lead evaluations...TrainingWork at officeLocal area
$160k - $200k
...founded in 2023 by leading machine learning researchers from MIT and Harvard Medical School. We... ...We’re hiring an exceptional ML Engineer to join our team (Boston or NYC office).... ...development of end‑to‑end ML systems (design, training, inference, deployment, and monitoring;...TrainingWork at office$189.6k - $237k
...distributed framework for large language model training and inference. The platform has been powering MLEs, researchers, data scientists and operators for fast and... ...-scale distributed ML systems Strong software engineering skills, proficient in frameworks and tools such...TrainingFull time$120k - $250k
...build an end-to-end platform for developing, training, and deploying AI systems—designed to take ideas from research to production with less friction. Through our... ...FOR We are seeking a highly skilled Research Engineer to help optimize training and inference workloads...TrainingFull timeWork at officeWork from homeFlexible hours2 days per week$100k - $125k
...employment, including but not limited to hiring, placement, promotion, termination, transfer, leave of absence, compensation, and training. All Blackstone employees, including but not limited to recruiting personnel and hiring managers, are required to abide by this policy...TrainingLocal areaFlexible hours$70k - $120k
...about the company at Job Summary The R&D Engineer I/II is responsible for supporting... ...work within a multidisciplinary team of researchers and engineers. The position offers diverse... ...gowning, including successful respirator training. Wear Appropriate Personal Protective Equipment...TrainingFull timeTemporary workApprenticeshipWork experience placementWork at officeLocal area- A leading AI development company located in New York is seeking a Software Engineer to develop groundbreaking infrastructure for AI model training. The role involves designing large-scale data pipelines and optimizing distributed systems. Candidates should have strong skills...Training
$70k - $90k
...Senior Investigator - Pre-Pay (Healthcare FWA) Job Location: US-Remote Overview As a Senior Investigator, you will investigate suspected... ...summary and/or presentation. Conducts investigation-related training. Supports legal proceedings as needed, including testifying in court...TrainingWork experience placementWork at officeRemote workWork from home
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Engineer, Pre-training [Remote]. Be the first to apply!
- engineering change analyst New York, NY
- senior research engineer New York, NY
- research engineer New York, NY
- junior machine learning research engineer New York, NY
- research programmer New York, NY
- engineering analyst New York, NY
- engineering business analyst New York, NY
- deep learning research engineer New York, NY
- cyber research engineer New York, NY
- ai research engineer New York, NY

