Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Engineer, Pre-training [Remote]

$340k - $425k

Anthropic

New York, NY
  • Remote job

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.

Anthropic is at the forefront of AI research, dedicated to developing safe, ethical, and powerful artificial intelligence. Our mission is to ensure that transformative AI systems are aligned with human interests. We are seeking a Research Engineer to join our Pre-training team, responsible for developing the next generation of large language models. In this role, you will work at the intersection of cutting-edge research and practical engineering, contributing to the development of safe, steerable, and trustworthy AI systems.

Key Responsibilities:

  • Conduct research and implement solutions in areas such as model architecture, algorithms, data processing, and optimizer development
  • Independently lead small research projects while collaborating with team members on larger initiatives
  • Design, run, and analyze scientific experiments to advance our understanding of large language models
  • Optimize and scale our training infrastructure to improve efficiency and reliability
  • Develop and improve dev tooling to enhance team productivity
  • Contribute to the entire stack, from low-level optimizations to high-level model design

Qualifications:

  • Advanced degree (MS or PhD) in Computer Science, Machine Learning, or a related field
  • Strong software engineering skills with a proven track record of building complex systems
  • Expertise in Python and experience with deep learning frameworks (PyTorch preferred)
  • Familiarity with large-scale machine learning, particularly in the context of language models
  • Ability to balance research goals with practical engineering constraints
  • Strong problem-solving skills and a results-oriented mindset
  • Excellent communication skills and ability to work in a collaborative environment
  • Care about the societal impacts of your work

Preferred Experience:

  • Work on high-performance, large-scale ML systems
  • Familiarity with GPUs, Kubernetes, and OS internals
  • Experience with language modeling using transformer architectures
  • Knowledge of reinforcement learning techniques
  • Background in large-scale ETL processes

You'll thrive in this role if you:

  • Have significant software engineering experience
  • Are results-oriented with a bias towards flexibility and impact
  • Willingly take on tasks outside your job description to support the team
  • Enjoy pair programming and collaborative work
  • Are eager to learn more about machine learning research
  • Are enthusiastic to work at an organization that functions as a single, cohesive team pursuing large-scale AI research projects
  • Are working to align state of the art models with human values and preferences, understand and interpret deep neural networks, or develop new models to support these areas of research
  • View research and engineering as two sides of the same coin, and seek to understand all aspects of our research program as well as possible, to maximize the impact of your insights
  • Have ambitious goals for AI safety and general progress in the next few years, and you’re working to create the best outcomes over the long-term.

Sample Projects:

  • Optimizing the throughput of novel attention mechanisms
  • Comparing compute efficiency of different Transformer variants
  • Preparing large-scale datasets for efficient model consumption
  • Scaling distributed training jobs to thousands of GPUs
  • Designing fault tolerance strategies for our training infrastructure
  • Creating interactive visualizations of model internals, such as attention patterns

At Anthropic, we are committed to fostering a diverse and inclusive workplace. We strongly encourage applications from candidates of all backgrounds, including those from underrepresented groups in tech.

If you're excited about pushing the boundaries of AI while prioritizing safety and ethics, we want to hear from you!

The expected salary range for this position is:

Annual Salary:

$340,000—$425,000 USD

Logistics

Education requirements: We require at least a Bachelor's degree in a related field or equivalent experience.

Location-based hybrid policy:
Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.

Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.

We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed.  Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team.

How we're different

We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills.

The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.

Come work with us!

Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about  our policy for using AI in our application process

Vacancy posted more than 2 months ago
Similar jobs that could be interesting for youBased on the Research Engineer, Pre-training [Remote] in New York, NY vacancy
  • $350k

    Anthropic seeks a Research Engineer/Research Scientist to join our Pre-training team in a remote-friendly position. You will develop next-generation large language models and work on projects at the intersection of research and engineering. The role requires an advanced... 
    Training
    Remote work
    Flexible hours

    Anthropic

    New York, NY
    13 hours ago
  • $350k

    A leading AI research company is seeking a Pre-training Research Engineer to advance large language models. You will be engaged in research, implementing solutions, and enhancing training infrastructure while collaborating with other experts. Candidates must possess an... 
    Training
    Work at office

    Menlo Ventures

    New York, NY
    1 day ago
  • $350k

     ...for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. Pre-training Research Engineer Anthropic is at the forefront of AI research, dedicated... 
    Training
    Work at office
    Visa sponsorship
    Flexible hours

    Menlo Ventures

    New York, NY
    1 day ago
  •  ...and distribution. Higharc is hiring a Research Engineer to join our Special Projects team. In this...  ...can be built. Design end-to-end LLM training and fine-tuning pipelines tailored to...  ...construction domains, developing domain-specific pre-training and SFT datasets from... 
    Training
    Full time
    Temporary work
    Remote work
    Home office
    Flexible hours

    Higharc Inc.

    New York, NY
    13 hours ago
  • $180k - $200k

    About the role We are seeking exceptional Senior Research Engineers to join our mission-critical team building the world's best oncology...  ...inform our clinical-stage programs. You'll participate in both pre-training and post-training of our foundation models, requiring deep... 
    Training
    Work at office
    3 days per week

    Pathos

    New York, NY
    1 day ago
  •  ...Clinical Research Coordinator II - Cell Engineering & Therapy The Clinical Research Coordinator (CRC) will play...  ...other samples. Maintain updated training in human research protection rules...  ...support query resolution, resolve pre-and post-monitoring action items, and... 
    Training
    Work at office
    Local area

    Columbia University

    New York, NY
    2 days ago
  • $111.5k - $207.5k

     ...Title Senior Specialist, Security Software Research Engineer Job Code 36915 Job Location Remote Job...  ...deliverables for the organization. Training, management and provision of guidance to...  ...maintains a drug‑free workplace and performs pre‑employment substance abuse testing and... 
    Training
    Local area
    Immediate start
    Remote work
    Flexible hours

    L3Harris

    New York, NY
    13 hours ago
  • $165k - $260k

    A leading financial technology company in New York seeks a Senior NLP Engineer to work on innovative AI-driven solutions. This role involves designing, training, and evaluating NLP models, collaborating across teams, and publishing findings. Ideal candidates will have... 
    Training

    Bloomberg

    New York, NY
    3 days ago
  • A pioneering AI startup is looking for a Research Engineer to develop innovative reinforcement learning systems. The role requires experience in distributed RL using PyTorch and JAX, with a focus on building production-ready systems for complex optimization tasks. Candidates... 
    Training
    Remote job
    Full time

    Strativ Group

    New York, NY
    13 hours ago
  • $200k - $250k

     ...Build and maintain robust distributed training systems using PyTorch and JAX Build and...  ...training and evaluation. Drive innovation by researching and developing scalable reinforcement...  ...Full-time Job function Research and Engineering Software Development and Research... 
    Training
    Full time
    Immediate start
    Remote work

    Strativ Group

    New York, NY
    13 hours ago
  • A cutting-edge tech company in New York is seeking a Research Engineer to build and scale systems for their AI video generation models. This role involves training and optimizing large-scale models, improving efficiency, and translating research into production-ready systems... 
    Training

    Mirage

    New York, NY
    2 days ago
  • About Basis Basis is a nonprofit applied AI research organization with two mutually...  ...values first. About the Role Research Engineers in Operations at Basis build the internal...  ...scratch You build systems and workflows for training large models distributed across many machines... 
    Training
    Full time
    Contract work
    Work at office

    Basis Research Institute

    New York, NY
    2 days ago
  • About the Role Mirage is seeking a Research Engineer to build and scale systems for training and deploying large language models for multimodal creative tasks, specifically focusing on video analysis pipelines. You’ll work closely with researchers to turn ideas into high... 
    Training
    Full time
    Local area
    Night shift

    Mirage

    New York, NY
    3 days ago
  • $120k - $175k

     ...We are seeking a Security Research Engineer to operate as a hybrid Forward Deployed Engineer and offensive security researcher. You'll be...  ...company Professional development budget for conferences, training, and certifications Support for publishing research and presenting... 
    Training

    Pensar

    New York, NY
    7 days ago
  •  ...Achira, we are building a team of world-class scientists, ML researchers, and engineers to work together to move beyond the beaten path in drug...  ...participating with research teams to scale our ML systems, train and evaluate models, and engineer scientific prototypes into... 
    Training
    Work at office

    Achira

    New York, NY
    4 days ago
  • $200k

     ...Optiver is a seeking a Machine Learning Research Engineer to join our team, focusing on a pivotal AI initiative. This role would offer the...  ...significant impact across Machine Learning infrastructure, training, and inference challenges to advance our futures trading strategies... 
    Training
    Work at office

    Optiver

    New York, NY
    2 days ago
  • $200k - $400k

     ...team. About the Team Read more about the research team's work here: The Research team...  ...implement state‑of‑the‑art techniques in model training, prompting, orchestration, and...  ...scale. About the Role As a Staff Research Engineer, you’ll be responsible for building industry... 
    Training
    Full time
    Work at office
    Local area

    Decagon

    New York, NY
    2 days ago
  • $124k - $144k

     ...Senior Research Network Engineer Posting Number 2026-15621 Location : Location US-NY-New York Hybrid Remote Work Classification...  ...of the position, candidate's work experience, education/training, key skills, internal peer equity, as well as, market and... 
    Training
    Full time
    Work experience placement
    Remote work

    New York University

    New York, NY
    27 days ago
  • $315k - $340k

    [Expression of Interest] Research Scientist/Engineer, Honesty About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable...  ...data curation pipelines to identify, verify, and filter training data for accuracy given the model’s knowledge Develop... 
    Training
    Full time
    Work at office
    Visa sponsorship
    Flexible hours

    Aisafety

    New York, NY
    1 day ago
  •  ...dimensionality creates computational and storage costs that make training and deployment prohibitively expensive at world scale. We...  ...seeking a highly skilled and versatile Machine Learning Engineer to join our Research team. As a Member of the Research Staff, you will partner... 
    Training
    Home office
    Flexible hours

    Deepgram

    New York, NY
    13 hours ago
  • $160k - $200k

     ...founded in 2023 by leading machine learning researchers from MIT and Harvard Medical School. We...  ...We’re hiring an exceptional ML Engineer to join our team (Boston or NYC office)....  ...development of end‑to‑end ML systems (design, training, inference, deployment, and monitoring;... 
    Training
    Work at office

    Verana Health

    New York, NY
    2 days ago
  •  ...investors . I’ll share more once we meet. About the Role As an ML Research Engineer at Maple, you'll be a part of our core product team...  ...devices, minimizing latency. Manage rapid experimentation, training, and highly optimized production inference. Lead evaluations... 
    Training
    Work at office
    Local area

    Maple

    New York, NY
    13 hours ago
  • $174k - $252k

     ...PhD degree in Computer Science, Computer Engineering, Cybersecurity, a related quantitative...  ...language. 2 years of professional or academic research experience applying machine learning...  ...or curating large‑scale datasets for training or evaluating AI/ML models, particularly... 
    Training
    Full time

    Google DeepMind

    New York, NY
    4 days ago
  •  ...are designed to solve real-world business problems. We're training and deploying frontier models for enterprises who are...  ...the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft... 
    Training
    Work at office
    Remote work

    Cohere

    New York, NY
    2 days ago
  • $150k - $225k

     ...Quantitative Research Engineer London; New York We are looking for an experienced Quantitative Research Engineer to work on our data and...  ...and maintenance of documentation, user guides, and training materials related to research tools and processes Skill... 
    Training

    Numeus

    New York, NY
    4 days ago
  • $264.8k - $331k

     ...Staff Machine Learning Research Engineer, Agent Post-training - Enterprise GenAI San Francisco, CA; New York, NY AI is becoming vitally important in every function of our society. At Scale, our mission is to accelerate the development of AI applications. For 9 years... 
    Training
    Full time

    Scale AI

    New York, NY
    4 days ago
  •  ...Stripe), top-tier angels. About the Role You'll bridge research and engineering-rapidly implementing, experimenting with, and scaling new...  ...large, real-world data. Build tools and pipelines for training, evaluation, and analysis. Implement state-of-the-art research... 
    Training

    ATG intelligence

    New York, NY
    4 days ago
  •  ...International, Inc. (ATL) is hiring a Senior Engineer/Scientist - PhD/ScD epidemiologist (MACS-...  ...and health saving account options offer pre-tax savings for qualified medical, dental...  ...but not limited to, recruiting, hiring, training, promotion, compensation, benefits, and... 
    Training
    Full time
    Temporary work
    Local area
    Immediate start
    Flexible hours

    Advanced Technologies and Laboratories International, Inc.

    New York, NY
    4 days ago
  • $100k - $125k

     ...employment, including but not limited to hiring, placement, promotion, termination, transfer, leave of absence, compensation, and training. All Blackstone employees, including but not limited to recruiting personnel and hiring managers, are required to abide by this policy... 
    Training
    Local area
    Flexible hours

    Blackstone Restaurant

    New York, NY
    1 day ago
  • A leading AI development company located in New York is seeking a Software Engineer to develop groundbreaking infrastructure for AI model training. The role involves designing large-scale data pipelines and optimizing distributed systems. Candidates should have strong skills... 
    Training

    Reflection

    New York, NY
    2 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Engineer, Pre-training [Remote]. Be the first to apply!