Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Research Engineer -- AI Alignment & Evaluation

W3 Sourcing

Job Description

Job Description

Research Engineer — AI Alignment & Evaluation

AI Safety / Research Engineering | San Francisco, CA | Hybrid / In-Person

About the Company

We are representing a high-growth AI research organization working at the intersection of frontier model evaluation, AI safety, and security.

The team develops sophisticated evaluation environments designed to surface undesirable or misaligned model behavior and help leading AI organizations better understand how advanced systems behave under complex, long-horizon conditions.

This is a technically rigorous environment for engineers who are interested in AI alignment, agent behavior, model evaluation, and building systems that help make increasingly capable AI more reliable and controllable.

The Role

This is an opportunity to join a small, highly technical team as a Research Engineer with significant end-to-end ownership.

You will independently design and build evaluation environments that test frontier AI systems for subtle forms of undesirable behavior. You will own the full lifecycle of each environment, from initial concept and failure-mode identification through implementation, grader development, testing, measurement, and refinement.

A significant part of the role involves working directly with advanced LLM agents: prompting them to perform technical tasks, reviewing their output, identifying subtle errors, and making judgment calls where current models still fall short.

The role is ideal for a strong software engineer or technical researcher who enjoys ambiguous problems, learns new domains quickly, and is deeply interested in AI alignment and security.

What You'll Do
  • Design and build complex evaluation environments for frontier AI models.
  • Own evaluation projects end to end, including ideation, implementation, testing, grading, measurement, and iteration.
  • Investigate potential model failure modes and identify ways advanced agents may exploit or circumvent intended constraints.
  • Develop and improve software infrastructure used to isolate, reproduce, and evaluate model behavior.
  • Work extensively with LLM-based agents to accelerate implementation and research workflows.
  • Review agent-generated work critically and identify subtle technical or conceptual errors.
  • Build long-horizon tasks that operate near the edge of current model capabilities.
  • Apply strong qualitative judgment when evaluating behavior that cannot be captured through simple automated metrics.
  • Rapidly learn unfamiliar technical domains as required by individual evaluation environments.
  • Share findings, lessons, and technical context with a highly collaborative research and engineering team.
What We're Looking For
  • 1+ years of experience in software engineering, machine learning engineering, technical research, or a closely related field.
  • Strong traditional software engineering fundamentals.
  • Proficiency with Python .
  • Strong interest in AI alignment, AI safety, or AI security.
  • Ability to reason carefully about complex systems and ambiguous failure modes.
  • Strong conceptual judgment and the ability to think through how an autonomous agent may interpret or exploit a task.
  • Ability to learn new technical domains quickly.
  • Experience using LLMs or AI agents effectively as part of technical workflows.
  • Strong ability to assess whether agent-generated work is correct, including when errors are subtle.
  • Comfortable taking full ownership of technically demanding projects with limited oversight.
  • High standards for quality, execution, and accountability.
Nice to Have
  • Experience building evaluation frameworks, benchmarks, simulation environments, or agent-based systems.
  • Exposure to frontier language models or autonomous agent workflows.
  • Background in AI safety, alignment research, adversarial testing, or security.
  • Experience designing tasks that require multi-step or long-horizon reasoning.
  • Research experience involving model behavior, reward hacking, robustness, or control mechanisms.
Why This Role Is Exciting
  • Own technically challenging research environments from concept through final evaluation.
  • Work directly with state-of-the-art AI systems and agentic workflows.
  • Tackle problems at the frontier of AI safety, model behavior, and alignment.
  • Join a small technical team where individual work has meaningful visibility and impact.
  • Operate with substantial autonomy while receiving frequent technical feedback.
  • Build expertise across a wide range of domains rather than working within a narrow product surface.
  • Contribute to work focused on understanding and mitigating undesirable AI behavior rather than simply increasing model capabilities.
Work Model
  • Full-time position.
  • San Francisco-based role with regular in-office collaboration expected.
  • Flexibility around hybrid working arrangements.
  • Open to candidates willing to relocate.
  • Visa transfers and new visa sponsorship may be available.
  • Work is highly ownership-driven, with emphasis on the quality of what you ship.

Confidential details removed: salary, client name, founder names, exact address, company links, investor names, funding details, exact team size, founding year, and highly identifiable wording.

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Research Engineer -- AI Alignment & Evaluation in San Francisco, CA vacancy
  • A leading AI research organization in San Francisco is seeking a Research Engineer / Research Scientist to ensure AI systems align with human intent amid complex scenarios. Your responsibilities...  ...for AI alignment, developing evaluation methodologies, and building tools... 
    Suggested
    Work at office
    Relocation package

    OpenAI

    San Francisco, CA
    1 day ago
  • $174k - $252k

    Research new alignment methods, studying alignment failures, and applying AGI-scalable alignment techniques...  ...techniques to understand what AI systems are thinking.Work with product...  ...Computer Science, a related Software Engineering field, or equivalent practical experience... 
    Suggested

    Google

    San Francisco, CA
    1 day ago
  • United States Digital Space LLC is seeking a bio safety researcher to design and run capability evaluations for frontier models in biology, curate training data for safety classifiers, and collaborate with ML engineers to ensure robustness and practical deployment. The... 
    Suggested

    United States Digital Space LLC

    San Francisco, CA
    17 hours ago
  • $110.7k - $379.2k

    Position Summary Research Engineer — Post-Training & Small Language Models (SLMs), Healthcare AI Three hundred fifty million Americans rely on a healthcare system...  ...-training team, you will design, train, evaluate, and align the models that reason about healthcare —... 
    Suggested
    Local area
    Visa sponsorship

    Deloitte

    San Francisco, CA
    3 days ago
  •  ...mission is general causal intelligence; AI that is capable of (1) predicting the...  ...Insitro, Nabla Bio, and CERN. We look for research engineers who are excited to tackle unsolved...  .... Your mission is to build the central evaluation framework that the entire research organization... 
    Suggested

    Causal Labs

    San Francisco, CA
    2 days ago
  • $315k

    We are looking for Research Engineers to build “gold standard” evaluations for catastrophic risks, in order to understand what AI Safety Level (ASL) to assign to models. Research leads on this team collaborate with engineers in one of our focus areas: CBRN, Cyber, Autonomy... 
    Currently hiring
    Work at office
    Immediate start
    Home office
    Visa sponsorship
    Relocation package

    Anthropic

    San Francisco, CA
    3 days ago
  •  ...Research Engineer Lotus Health is a groundbreaking primary care app that integrates your medical records, AI, and real doctors to provide free, personalized healthcare...  ..., model training and evaluation, retrieval and tool use, safety and alignment, and putting... 

    Lotus Health

    San Francisco, CA
    3 days ago
  •  ...Archive Human Archive is a research lab backed by Y...  ...Opportunity As a Research Engineer, you'll work on multimodal...  ...research for embodied AI and robotics. This role...  ...to design experiments, evaluate new sensing stacks, and...  ...fusion, and multimodal alignment Prototype quickly... 
    Shift work

    Human Archive

    San Francisco, CA
    4 days ago
  • Acceler8 Talent seeks a Research Engineer to bring formal methods and rigorous reasoning to frontier AI alignment research. You will apply ideas from formal verification, program analysis, and compilers to understand and constrain model behavior, building practical research... 

    Acceler8 Talent

    San Francisco, CA
    1 day ago
  • Anthropic in San Francisco seeks a bio safety researcher to design and run capability evaluations for biology-focused models, build and curate datasets for...  ...classifiers, and iterate on those classifiers with ML engineers. You will work at the intersection of applied ML and... 

    Anthropic

    San Francisco, CA
    3 days ago
  • HeyMilo AI is hiring a Research Engineer to join our Applied AI team in San Francisco. You’ll design and build reinforcement learning environments...  ...use cases, creating simulators, reward functions, and evaluation harnesses to measure model performance. The role blends... 

    HeyMilo AI

    San Francisco, CA
    4 days ago
  • $227.2k - $284k

     ...is to develop reliable AI systems for the world’...  ..., combining rigorous evaluation with full-stack deployment...  ...with applied ML research, design, and evaluation...  ...Machine Learning Research Engineer, you will operate...  ...training methods, LLM alignment, or applied MLExperience... 
    Full time

    Scale AI

    San Francisco, CA
    4 days ago
  •  ...world reasoning. Build and operate end-to-end LLM evaluation systems, including runs, scoring, dashboards, and...  ...augmentation, and curation. Collaborate with AI researchers, applied AI teams, and data producers to align evaluations with training objectives. Own... 
    Full time
    Work at office
    Relocation package

    Mercor

    San Francisco, CA
    9 days ago
  • A leading AI technology firm in San Francisco is looking for a Research Engineer, Post-Training. This role involves optimizing AI models and developing algorithms to enhance data efficiency. Candidates should have a PhD or equivalent and experience with machine learning... 

    Character.AI

    San Francisco, CA
    17 hours ago
  • $315k

    As a Research Engineer or Research Scientist in Applied Finetuning, you will...  ...to the public via Claude.AI and our API. In this role, you...  ...on data mixes, design evaluations, and improve our production...  ...machine learning, NLP, or AI alignment or similar industry experience... 
    Work at office
    Home office
    Visa sponsorship
    Relocation package

    Anthropic

    San Francisco, CA
    3 days ago
  • Senior Experimental Research Engineer, Electromagnetics Senior Experimental...  ...Note: this does not include evaluating PCB performance using...  ...Research Scientist/Engineer, Alignment Finetuning Research Engineer...  ...article, started with the help of AI. #J-18808-Ljbffr Gridware
    Full time

    Gridware

    San Francisco, CA
    17 hours ago
  •  ...Space LLC in San Francisco is seeking exceptional research engineers to push the boundaries of frontier AI safety, shaping empirical understanding of risk and...  ...-to-end threads within this effort. You’ll design evaluations of frontier models against real threat models,... 

    United States Digital Space LLC

    San Francisco, CA
    1 day ago
  •  ...learning, post-training, evaluations, harnessing, and...  ...deployment—and connect that research to the patients,...  ...developing frontier biomedical AI capabilities. Prior...  ..., ownership, and alignment with the mission are most...  ...with researchers, engineers, clinicians, and product... 
    Work at office
    Relocation package

    OpenAI

    San Francisco, CA
    6 days ago
  • $250k - $350k

    Research Engineer / Scientist (Robot Learning) About World Labs: We build...  ...physical worlds — unlocking AI's full potential through spatial...  ...identification, real-sim alignment. Collaborate with simulation...  ...end-to-end training and evaluation workflows for robot policies... 

    World Labs Inc.

    San Francisco, CA
    1 day ago
  •  ...empowering and governing autonomous AI agents across industries....  ...Role We are seeking a Staff Research Engineer, AI/ML & Cybersecurity to...  ...Develop model evaluation, benchmarking, and stress-testing...  ...distribution Contribute to SOC2-aligned engineering practices Qualifications... 

    Ephapsys

    San Francisco, CA
    4 days ago
  • $140k - $200k

    Research Engineer & Scientist The Center for AI Safety (CAIS) is a leading research and advocacy organization focused...  ..., build the tooling to train and evaluate models at scale, and turn results...  ..., state, and local laws. In alignment with the San Francisco Fair Chance... 
    Work at office
    Local area

    Center for Ai Safety

    San Francisco, CA
    1 day ago
  • $197.3k - $313.7k

     ...SalesforceSalesforce is the #1 AI CRM, where humans with agents...  ...software and platform engineers to embed in our AI team to bridge...  ...directly enable world-class research and products used by millions...  ...services, like APIs, UIs, agentic evaluators, and more.Roll out scalable... 
    Full time

    Salesforce

    San Francisco, CA
    3 days ago
  • Factory is seeking innovative Research Engineers to design and integrate advanced AI and ML capabilities that revolutionize productivity and accelerate innovation...  ..., focusing on retrieval systems, code generation evaluation, agentic user experience development, and agent... 
    Work at office

    The San Francisco AI Factory

    San Francisco, CA
    4 days ago
  • $164.6k - $313.3k

     ...OpportunityAdobe's Sound Design AI group (SODA) is looking for a driven Data/ML engineer to push the boundaries of audio...  ...small, collaborative and efficient research team looking for highly...  ...finetune models on pipeline outputs, evaluate their behavior, and use those findings... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    2 days ago
  •  ...iterate on, and innovate on the AI brains behind The Path's AI Therapist. Combine research, data science, and engineering to create models, orchestration, and evaluation systems that make therapy...  ...new ideas. Improve Safety, Alignment, and Clinical Guardrails Work... 

    The Path

    San Francisco, CA
    4 days ago
  • $180k - $280k

     ...Research Engineer SuperAnnotate helps the world's leading AI teams build responsible, next-generation models powered by high-quality human data. We're a fast...  ...hands-on implementation, including annotating, evaluating, or sourcing data. Turn research directions into... 
    Full time

    SuperAnnotate AI

    San Francisco, CA
    3 days ago
  • $210k - $275k

     ...Research Engineer You'll build the evaluation systems that tell us whether Firecrawl actually works. That sounds simple. It isn't. Our core promise, convert...  ...is the easiest way to turn the web into data AI agents can use. One API call converts any URL into clean... 
    Full time
    Temporary work
    For contractors
    Remote work
    Visa sponsorship
    Flexible hours

    firecrawl

    San Francisco, CA
    1 day ago
  • $180k - $340k

     ...Research Engineer You'll own the quality of AI across everything Gamma creates. As our Research Engineer, you'll design evaluation frameworks that measure AI output quality, systematically improve production prompts, and fine-tune models to ensure millions of users... 
    Full time
    Work at office
    Work from home

    Gamma

    San Francisco, CA
    2 days ago
  • $200k - $400k

     ...Research Engineer As a Research Engineer, you'll be responsible for building industry-leading conversational AI models that power Decagon's agent, and taking them all the way from idea...  ...implement frontier approaches for training, evaluation, and orchestration across the system... 
    Work at office
    Local area

    decagon

    San Francisco, CA
    1 day ago
  • About Sable Sable built Aidan, the first AI employee who can lead customer calls...  ...graph), and the verifiers (how we can keep evaluating Aidan's performance in real scenarios)....  ...Making that feel human is one of the hardest engineering problems in AI, and our engineering team... 

    Sable AI, Inc

    San Francisco, CA
    17 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Research Engineer -- AI Alignment & Evaluation. Be the first to apply!