Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Engineer, Pre-training [Remote]

$340k - $425k

Anthropic

New York, NY
  • Remote job

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.

Anthropic is at the forefront of AI research, dedicated to developing safe, ethical, and powerful artificial intelligence. Our mission is to ensure that transformative AI systems are aligned with human interests. We are seeking a Research Engineer to join our Pre-training team, responsible for developing the next generation of large language models. In this role, you will work at the intersection of cutting-edge research and practical engineering, contributing to the development of safe, steerable, and trustworthy AI systems.

Key Responsibilities:

  • Conduct research and implement solutions in areas such as model architecture, algorithms, data processing, and optimizer development
  • Independently lead small research projects while collaborating with team members on larger initiatives
  • Design, run, and analyze scientific experiments to advance our understanding of large language models
  • Optimize and scale our training infrastructure to improve efficiency and reliability
  • Develop and improve dev tooling to enhance team productivity
  • Contribute to the entire stack, from low-level optimizations to high-level model design

Qualifications:

  • Advanced degree (MS or PhD) in Computer Science, Machine Learning, or a related field
  • Strong software engineering skills with a proven track record of building complex systems
  • Expertise in Python and experience with deep learning frameworks (PyTorch preferred)
  • Familiarity with large-scale machine learning, particularly in the context of language models
  • Ability to balance research goals with practical engineering constraints
  • Strong problem-solving skills and a results-oriented mindset
  • Excellent communication skills and ability to work in a collaborative environment
  • Care about the societal impacts of your work

Preferred Experience:

  • Work on high-performance, large-scale ML systems
  • Familiarity with GPUs, Kubernetes, and OS internals
  • Experience with language modeling using transformer architectures
  • Knowledge of reinforcement learning techniques
  • Background in large-scale ETL processes

You'll thrive in this role if you:

  • Have significant software engineering experience
  • Are results-oriented with a bias towards flexibility and impact
  • Willingly take on tasks outside your job description to support the team
  • Enjoy pair programming and collaborative work
  • Are eager to learn more about machine learning research
  • Are enthusiastic to work at an organization that functions as a single, cohesive team pursuing large-scale AI research projects
  • Are working to align state of the art models with human values and preferences, understand and interpret deep neural networks, or develop new models to support these areas of research
  • View research and engineering as two sides of the same coin, and seek to understand all aspects of our research program as well as possible, to maximize the impact of your insights
  • Have ambitious goals for AI safety and general progress in the next few years, and you’re working to create the best outcomes over the long-term.

Sample Projects:

  • Optimizing the throughput of novel attention mechanisms
  • Comparing compute efficiency of different Transformer variants
  • Preparing large-scale datasets for efficient model consumption
  • Scaling distributed training jobs to thousands of GPUs
  • Designing fault tolerance strategies for our training infrastructure
  • Creating interactive visualizations of model internals, such as attention patterns

At Anthropic, we are committed to fostering a diverse and inclusive workplace. We strongly encourage applications from candidates of all backgrounds, including those from underrepresented groups in tech.

If you're excited about pushing the boundaries of AI while prioritizing safety and ethics, we want to hear from you!

The expected salary range for this position is:

Annual Salary:

$340,000—$425,000 USD

Logistics

Education requirements: We require at least a Bachelor's degree in a related field or equivalent experience.

Location-based hybrid policy:
Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.

Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.

We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed.  Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team.

How we're different

We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills.

The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.

Come work with us!

Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about  our policy for using AI in our application process

Vacancy posted more than 2 months ago
Similar jobs that could be interesting for youBased on the Research Engineer, Pre-training [Remote] in New York, NY vacancy
  • $350k

    A leading AI research company is seeking a Pre-training Research Engineer to advance large language models. You will be engaged in research, implementing solutions, and enhancing training infrastructure while collaborating with other experts. Candidates must possess an... 
    Training
    Work at office

    Menlo Ventures

    New York, NY
    5 days ago
  • $350k

     ...for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. Pre-training Research Engineer Anthropic is at the forefront of AI research, dedicated... 
    Training
    Work at office
    Visa sponsorship
    Flexible hours

    Menlo Ventures

    New York, NY
    5 days ago
  • $180k - $200k

    About the role We are seeking exceptional Senior Research Engineers to join our mission-critical team building the world's best oncology...  ...inform our clinical-stage programs. You'll participate in both pre-training and post-training of our foundation models, requiring deep... 
    Training
    Work at office
    3 days per week

    Pathos Lab

    New York, NY
    5 days ago
  • $111.5k - $207.5k

     ...national security. Job Title: Lead, Security Research Engineer Job Code: 35904 Job Location: Remote...  ...or contractors of the company. Training, management and provision of guidance to...  ...maintains a drug‑free workplace and performs pre‑employment substance abuse testing and... 
    Training
    For contractors
    Local area
    Remote work
    Flexible hours

    L3Harris

    New York, NY
    4 days ago
  • A cutting-edge tech company in New York is seeking a Research Engineer to build and scale systems for their AI video generation models. This role involves training and optimizing large-scale models, improving efficiency, and translating research into production-ready systems... 
    Training

    Mirage

    New York, NY
    1 day ago
  • $165k - $260k

    A leading financial technology company in New York seeks a Senior NLP Engineer to work on innovative AI-driven solutions. This role involves designing, training, and evaluating NLP models, collaborating across teams, and publishing findings. Ideal candidates will have... 
    Training

    Bloomberg

    New York, NY
    2 days ago
  • $146.7k - $214.8k

     ...in of Cisco's products. If you enjoy vulnerability research, crash analysis, reverse engineering, and researching new techniques and writing tools to...  ...experience, qualifications, education, certifications, and/or training. The full salary range for certain locations is... 
    Training
    Full time
    Temporary work
    Local area
    Remote work
    Flexible hours

    Cisco

    New York, NY
    7 hours ago
  •  ...located in Union Square). About the Role Mirage is seeking a Research Engineer to build and scale the systems powering our video generation...  ...ultra‑low latency, real‑time generation. Responsibilities Train and optimize large‑scale video and multimodal models Improve... 
    Training
    Full time
    Local area
    Night shift

    Mirage

    New York, NY
    1 day ago
  • About the Role Mirage is seeking a Research Engineer to build and scale systems for training and deploying large language models for multimodal creative tasks, specifically focusing on video analysis pipelines. You’ll work closely with researchers to turn ideas into high... 
    Training
    Full time
    Local area
    Night shift

    Mirage

    New York, NY
    2 days ago
  • Job Title: Research Engineer Contract Duration: 1 year, possible extension Work Arrangement: Remote, ideally EST, buy anywhere in NORAM...  ...deep learning libraries that support large-scale distributed training, open sourcing high quality code and reproducible results for... 
    Training
    Contract work
    Work experience placement
    Remote work

    EPITEC

    New York, NY
    5 days ago
  • About Basis Basis is a nonprofit applied AI research organization with two goals: Understand...  ...human value. About the Role Research engineers support Basis' mission by translating research...  ...new techniques, building large-scale training systems, or spanning multiple levels of... 
    Training
    Full time
    Flexible hours

    Second Renaissance

    New York, NY
    3 days ago
  • $300k - $405k

     ...a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build...  ...and evaluations, delivering your work into production training runs, and collaborating with other researchers, engineers,... 
    Training
    Work at office
    Visa sponsorship
    Flexible hours

    Menlo Ventures

    New York, NY
    1 day ago
  • Bloomberg’s Engineering AI department has 400+ AI practitioners building highly sought after...  ...is built on technology that makes news, research, financial data, and analytics on over 3...  ...methods for LLMs, efficient methods for training, multimodal models, learning from feedback... 
    Training
    Work experience placement

    Bloomberg

    New York, NY
    1 day ago
  • About Basis Basis is a nonprofit applied AI research organization with two mutually...  ...values first. About the Role Research Engineers in Operations at Basis build the internal...  ...scratch You build systems and workflows for training large models distributed across many machines... 
    Training
    Full time
    Contract work
    Work at office

    Basis Research Institute

    New York, NY
    1 day ago
  • $174k - $252k

    Research Engineer, Pretraining, DeepMind DeepMind • New York, NY, USA Requirements Bachelor’s degree in Computer Science, Mathematics, Applied...  ...conducting applied research to improve the quality and training/serving efficiency of large transformer-based models.... 
    Training
    Full time

    Google Inc.

    New York, NY
    2 days ago
  •  ...Achira, we are building a team of world-class scientists, ML researchers, and engineers to work together to move beyond the beaten path in drug...  ...participating with research teams to scale our ML systems, train and evaluate models, and engineer scientific prototypes into... 
    Training
    Work at office

    Achira

    New York, NY
    3 days ago
  • $200k

     ...Optiver is a seeking a Machine Learning Research Engineer to join our team, focusing on a pivotal AI initiative. This role would offer the...  ...significant impact across Machine Learning infrastructure, training, and inference challenges to advance our futures trading strategies... 
    Training
    Work at office

    Optiver

    New York, NY
    1 day ago
  • $189.6k - $237k

     ...About This Role Lead applied ML engineering on Scale's Applied ML team, powering data...  ...behaviors, scale human expertise, and drive research into real-world agent reliability...  ...performance, and relevant education or training. Scale employees in eligible roles are also... 
    Training
    Full time

    Scale AI

    New York, NY
    7 days ago
  • $124k - $144k

     ...Position Summary The Senior Research Network Engineer plays a critical role in designing, implementing, managing, and monitoring the dedicated...  ...of the position, candidate's work experience, education/training, key skills, internal peer equity, as well as, market and... 
    Training
    Work experience placement
    Remote work

    NYU Rory Meyers College of Nursing

    New York, NY
    7 hours ago
  • $15k

     ...the most profitable market makers on the street. As a Senior Research Engineer embedded within our first quant strategy team, you will own...  .... You will be expected to help optimize machine learning training and testing environments, architect and maintain data pipelines... 
    Training
    Local area
    Night shift

    The-Voleon-Group

    New York, NY
    4 days ago
  • $250k - $350k

     ...Machine Learning Research Engineer, Agents - Enterprise GenAI AI is becoming vitally important in every function of our society. At Scale...  ...we are doubling down on building out state of the art post-training algorithms to reach the performance necessary for complex... 
    Training
    Full time

    Scale AI

    New York, NY
    3 days ago
  • $245k - $330k

     ...team at the NBA is seeking an experienced applied scientist / research engineer with a strong foundation in Computer Vision and Machine...  ...system, from building sensing pipelines to scalable ML data, training, modeling and evaluation pipelines. This team sits... 
    Training
    Temporary work
    Local area
    Remote work
    1 day per week

    The National Basketball Association

    New York, NY
    2 days ago
  •  ...investors . I’ll share more once we meet. About the Role As an ML Research Engineer at Maple, you'll be a part of our core product team...  ...devices, minimizing latency. Manage rapid experimentation, training, and highly optimized production inference. Lead evaluations... 
    Training
    Work at office
    Local area

    Maple

    New York, NY
    4 days ago
  • $160k - $200k

     ...founded in 2023 by leading machine learning researchers from MIT and Harvard Medical School. We...  ...We’re hiring an exceptional ML Engineer to join our team (Boston or NYC office)....  ...development of end‑to‑end ML systems (design, training, inference, deployment, and monitoring;... 
    Training
    Work at office

    Verana Health

    New York, NY
    1 day ago
  • $189.6k - $237k

     ...distributed framework for large language model training and inference. The platform has been powering MLEs, researchers, data scientists and operators for fast and...  ...-scale distributed ML systems Strong software engineering skills, proficient in frameworks and tools such... 
    Training
    Full time

    DiversityJobs Inc

    New York, NY
    11 days ago
  • $120k - $250k

     ...build an end-to-end platform for developing, training, and deploying AI systems—designed to take ideas from research to production with less friction. Through our...  ...FOR We are seeking a highly skilled Research Engineer to help optimize training and inference workloads... 
    Training
    Full time
    Work at office
    Work from home
    Flexible hours
    2 days per week

    Lightning AI

    New York, NY
    2 days ago
  • $100k - $125k

     ...employment, including but not limited to hiring, placement, promotion, termination, transfer, leave of absence, compensation, and training. All Blackstone employees, including but not limited to recruiting personnel and hiring managers, are required to abide by this policy... 
    Training
    Local area
    Flexible hours

    Blackstone Restaurant

    New York, NY
    7 hours ago
  • $70k - $120k

     ...about the company at Job Summary The R&D Engineer I/II is responsible for supporting...  ...work within a multidisciplinary team of researchers and engineers. The position offers diverse...  ...gowning, including successful respirator training. Wear Appropriate Personal Protective Equipment... 
    Training
    Full time
    Temporary work
    Apprenticeship
    Work experience placement
    Work at office
    Local area

    Cresilon, Inc

    New York, NY
    2 days ago
  • A leading AI development company located in New York is seeking a Software Engineer to develop groundbreaking infrastructure for AI model training. The role involves designing large-scale data pipelines and optimizing distributed systems. Candidates should have strong skills... 
    Training

    Reflection

    New York, NY
    1 day ago
  • $70k - $90k

     ...Senior Investigator - Pre-Pay (Healthcare FWA) Job Location: US-Remote Overview As a Senior Investigator, you will investigate suspected...  ...summary and/or presentation. Conducts investigation-related training. Supports legal proceedings as needed, including testifying in court... 
    Training
    Work experience placement
    Work at office
    Remote work
    Work from home

    Cotiviti

    New York, NY
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Engineer, Pre-training [Remote]. Be the first to apply!