Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Agent Post-Training, Artifacts Research

United States Digital Space LLC

About the Team The Agent Post-Training team creates the frontier agents the company ships to the world. We are training the models behind our agents in Codex, ChatGPT, the API, and other frontier products: persistent, proactive intelligence that can operate computers, collaborate with people and other agents, and expand what people and organizations can imagine, attempt, and achieve. We define what the next generation of agents should be able to do, build the training signal that teaches those abilities, and run the experiments that make them real. Our work spans coding, tool use, computer use, multi-agent coordination, long-horizon execution, factuality, instruction following, calibrated reasoning, and taste. Our team is where new model capabilities get made. We build the data, environments, graders, training methods, and feedback loops that shape what the company's next agents can do, then carry those capabilities through major training runs and into the products people use. About the Role As a member of Agent Post-Training, Artifacts, you will train frontier models to create polished, useful work products: documents, spreadsheets, slide decks, dashboards, reports, analyses, and other interactive or editable artifacts. You will help teach our models to move from a vague user goal to a finished artifact with strong structure, visual taste, domain judgment, correctness, and low latency. This work will require owning improvements across our post-training stack, including RL, data pipelines, graders, reward signals, evals, and behavioral analysis. You will work with researchers, engineers, product teams, infrastructure teams, and safety/alignment partners to decide what should go into major model runs, measure whether it worked, and ship improvements into products used by real people. This is a high-agency role for people who want their work to land directly in frontier models. Design and run experiments that improve agentic model behavior for complex software and plugins. Own end-to-end improvements to the post-training stack, including RL, data pipelines, graders, reward signals, evals, diagnostics, and model-behavior analysis. Build evals and environments that expose the next set of model failures, then turn those failures into training data, product fixes, or new research directions. Partner with Codex and ChatGPT product teams to understand what users need and translate product signal into model improvements. Work on early-training and alignment interventions, including data mixtures, objectives, synthetic data, and eval loops that shape downstream agent behavior. Help decide which integrations, capabilities, and fixes are ready for inclusion in major model runs. Improve the machinery for large-scale training and launch: experiment velocity, reliability, observability, reproducibility, cost, latency, and production readiness. Take on cross-functional projects that touch model training, product infrastructure, and the production agent harness, such as multi-agent systems or training directly against production-like environments. Debug hard failures in shipped or near-shipped models and turn messy qualitative behavior into concrete hypotheses, experiments, and fixes. You might thrive in this role if you: Have strong technical fundamentals in machine learning, software engineering, systems, statistics, or a related field, and can learn quickly across the parts you have not worked in before. Have hands-on experience with LLMs, RL, RLHF/RLAIF, post-training, evals, graders, synthetic data, model training, coding agents, tool-using agents, or production ML systems. Are excited by open-ended problems where the path is unclear, the signal is noisy, and the right answer requires both research taste and engineering execution. Care about product impact and model behavior, not just benchmark movement. You have opinions about what makes an agent useful, reliable, honest, tasteful, and easy to work with. Can move from a vague behavioral problem to a concrete experiment: define the hypothesis, build the pipeline, run the model, analyze the result, and decide what to do next. Are comfortable working across research, product, infrastructure, data, evals, and safety boundaries, and can communicate clearly with each group. Like building load-bearing systems and processes when that is what the team needs, even if the work is not glamorous. Want to train and ship the models that make agents genuinely useful for developers, enterprises, researchers, and everyday users. Have some prior background in consulting, finance, marketing, operations, or data science. About the company the company is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. #J-18808-Ljbffr United States Digital Space LLC

Vacancy posted 11 hours ago
Similar jobs that could be interesting for youBased on the Agent Post-Training, Artifacts Research in San Francisco, CA vacancy
  • $147k - $210k

    Drive projects by defining key research questions.Design, implement, and evaluate experiments...  ...AI, reinforcement learning, multimodal agents, computational social science, or...  ...qualifications:3 years of experience with training, evaluating, and interpreting large language... 
    Training

    Google

    San Francisco, CA
    6 hours ago
  • United States Digital Space LLC is looking for a Context Researcher to scale compute on context in the Agent Post-Training team. This role involves collaborating with researchers and engineers on model training and improving product interfaces. Your responsibilities will... 
    Training

    United States Digital Space LLC

    San Francisco, CA
    11 hours ago
  • cursor is hiring a Research Scientist to drive research in reinforcement learning at their New York office. Candidates...  ...of RL, improving data quality for model training, and executing realtime RL for coding agents. This position offers significant scope and autonomy... 
    Training
    Work at office

    cursor

    San Francisco, CA
    5 days ago
  •  ...Intelligence Co. in San Francisco/New York is seeking a Research Scientist to define the research agenda for models...  ...successfully. You will work across reasoning, post-training, continual learning, evals, and agent reliability, with the goal of improving Eico’s systems... 
    Training

    Economic Intelligence Co.

    San Francisco, CA
    4 days ago
  • Scale Labs is seeking a Research Scientist focused on Agent Robustness to advance safe and aligned AI agents. You will contribute to evaluating risks...  ...and environments. The role emphasizes collaboration, post-training techniques, and publishing results to shape policy and... 
    Training

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  • $25k

     ...agencies to safeguard the public through information sharing, training, research, and technology. Duties & Responsibilities This is an Open...  ...Announcement (OCA) to fill multiple Criminal Investigator (Special Agent) vacancies for the Homeland Security Task Force. Eligible... 
    Training
    Work at office
    Local area
    Immediate start
    Relocation package

    ATF

    San Francisco, CA
    2 days ago
  • $216k - $270k

    Scale Labs, Research Scientist - Agent Robustness As the leading data and evaluation partner for frontier AI companies, Scale plays an integral...  ...literature into working prototypes. Experience with post-training and RL techniques such as RLHF, DPO, GRPO, and similar approaches... 
    Training
    Full time

    Scale AI, Inc.

    San Francisco, CA
    5 days ago
  • $380k

    Research Scientist - Multimodal Agent, Consumer Devices | OpenAI Careers Research Scientist - Multimodal Agent, Consumer Devices Consumer Products...  ...Future of Computing Research team to work on RLHF and post‑training for personalized, multimodal AI systems. This role... 
    Training
    Work at office
    Immediate start
    Relocation package

    OpenAI

    San Francisco, CA
    4 days ago
  • $264.8k - $331k

     ...from Meta, we are doubling down on building out state of the art post-training algorithms to reach the performance necessary for complex agents in enterprises around the world. The Enterprise ML Research Lab works on the front lines of this AI revolution. We are working... 
    Training
    Full time

    Scale AI

    San Francisco, CA
    11 hours ago
  • $151.5k - $222.2k

     ...automation with foundation models, multi-agent systems, and robotics to make scientific...  ...and act against them. Responsibilities: Research & Innovation Partner with scientists to...  ...functions from noisy scientific signal. Post-train domain models (SFT, DPO/GRPO/PPO, reward... 
    Training
    Full time
    Flexible hours

    Initial Therapeutics, Inc.

    San Francisco, CA
    2 days ago
  • $218.4k - $273k

     ...evaluations.About the ACE team The Agent Capabilities & Environments (ACE) team, part of Scale’s Research organization, brings together...  ...range displayed on each job posting reflects the minimum and...  ...performance, and relevant education or training. Scale employees in eligible... 
    Training
    Full time

    Scale AI

    San Francisco, CA
    11 hours ago
  • $190k - $270k

     ...About the TeamThe Databricks AI Research organization is pushing the...  ...we're building the models and agents that unlock it. Our work spans the full stack, from model training to advanced multi-agent systems...  ...primary technical pillars include post-training enhancements, harness... 
    Training
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    2 days ago
  • $264.8k - $331k

     ...capabilities.About the General Agents TeamThe General Agents team,...  ...problem spaces, balancing research-driven approaches with pragmatic...  ...range displayed on each job posting reflects the minimum and maximum...  ..., and relevant education or training. Scale employees in eligible... 
    Training
    Full time

    Scale AI

    San Francisco, CA
    11 hours ago
  • $250k - $300k

     ...breakthrough AI models at leading research labs and enterprises. Since 20...  ...teams to produce high-quality training data at scale Frontier Data...  ...and evaluate autonomous agent capabilities. Design agent-focused...  ...journals, conferences, and blog posts. What You Bring Ph.D. or... 
    Training
    Work at office
    Flexible hours
    2 days per week

    Labelbox

    San Francisco, CA
    4 days ago
  • Polymath Labs is hiring a Founding Member of Technical Staff - Research to push the frontier of autonomous agents. You will work on long-horizon evaluation, agent post-training, and environment design, wearing multiple hats from building benchmarks to running rigorous experiments... 
    Training

    Polymath Labs

    San Francisco, CA
    11 hours ago
  •  ...intersection of cutting‑edge AI research and practical application,...  ...for building state‑of‑the‑art agents, such as browser and SWE...  ...range displayed on each job posting reflects the minimum and maximum...  ...performance, and relevant education or training. Scale employees in eligible... 
    Training

    Scale AI

    San Francisco, CA
    11 hours ago
  • Scale is hiring a Staff Machine Learning Research Engineer, Post-training, for Enterprise GenAI in San Francisco. The role focuses on building an Agent RL training platform and integrating cutting-edge research into the training stack to deploy high‑performing enterprise... 
    Training

    Scale

    San Francisco, CA
    3 days ago
  • Scale is hiring a Staff Agent Post-Training MLRE in New York to build and scale an Agent RL training platform for enterprise use-cases. You...  ...train state-of-the-art models on both internal and community research and deploy them to customers. Ideal candidates have 5+ years... 
    Training

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  • Scale is seeking a Machine Learning Research Engineer, Agents - Enterprise GenAI to advance state-of-the-art Agent RL training for enterprise datasets. You will train models, iterate...  ...with LLMs in production, experience with post-training methods (RLHF/RLVR, PPO/GRPO), and... 
    Training

    Scale

    San Francisco, CA
    1 day ago
  •  ...States Digital Space LLC in New York is seeking researchers to advance AI through production-grade LLM work, synthetic data pipelines and agent-building technologies. You will develop data for reinforcement learning post-training and build agents from production traces to... 
    Training

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  • $264.8k - $331k

    Machine Learning Systems Research Engineer, Agent Post-training - Enterprise GenAI AI is becoming vitally important in every function of our society. At Scale, our mission is to accelerate the development of AI applications. For 9 years, Scale has been the leading AI data... 
    Training
    Full time
    Contract work
    For contractors
    For subcontractor
    Work at office

    Scale LLP

    San Francisco, CA
    1 day ago
  • Scale is hiring a Machine Learning Research Engineer, Agent Data Foundation - Enterprise GenAI in San Francisco. You’ll work on synthetic data pipelines, agent-building frameworks, and post-training RL approaches to advance enterprise GenAI capabilities. You should have... 
    Training

    Scale

    San Francisco, CA
    3 days ago
  •  ...seeking a skilled individual to join their Agent Post-Training team. This role focuses on training advanced AI models to create useful artifacts such as reports and dashboards. It involves significant collaboration with researchers, engineers, and product teams to refine... 
    Training

    United States Digital Space LLC

    San Francisco, CA
    1 day ago
  • $158k - $269k

     ...Waabi develops cutting-edge simulation agents and scenario generation algorithms for Waabi...  ...World, our simulation platform. As a Research Scientist on the Behaviors team, you will...  ...system to its limits, generate the training signal that makes them better, and form... 
    Training
    Full time
    Work at office
    Work from home
    Flexible hours

    Waabi

    San Francisco, CA
    7 days ago
  •  ...A federal law enforcement agency is seeking Special Agents dedicated to criminal investigations and protection of high-level officials...  ...physical qualifications, including visual acuity. Intensive training is required and opportunities for career mobility exist nationally... 
    Training

    US Secret Service

    San Francisco, CA
    1 day ago
  •  ...Secret Service. During the course of their careers, special agents carry out assignments in both investigations and protection and...  ...while you occupy the position. Complete 13 weeks of intensive training at the Federal Law Enforcement Training Center(FLETC) in Glynco... 
    Training
    Overseas

    Disability Solutions

    Daly City, CA
    4 days ago
  •  ...Joining the United States Secret Service as a Special Agent offers a unique and rewarding career dedicated to serving the nation through...  ...Preferences are not guaranteed. Complete 13 weeks of intensive training at the Federal Law Enforcement Training Center (FLETC) in... 
    Training
    Overseas

    US Secret Service

    San Francisco, CA
    1 day ago
  •  ...scheduling that supports work-life balance Pre-qualified client consultations provided — no cold outreach required Comprehensive training and licensing support Weekly direct deposit Monthly and quarterly performance bonuses Equity opportunities for qualifying team members... 
    Training
    Remote work
    Flexible hours

    AO Globe Life

    San Francisco, CA
    4 days ago
  • $248.4k - $310.5k

     ...production, paired with applied ML research, design, and evaluation to...  ...Engineering Manager for Agent Oversight, you'll lead the team...  ...range displayed on each job posting reflects the minimum and maximum...  ..., and relevant education or training. Scale employees in eligible... 
    Training
    Full time

    Scale AI

    San Francisco, CA
    11 hours ago
  •  ...Contact Center Agent We are seeking a Contact Center Agent to join our team at Independent Living Systems (ILS). ILS, along with...  ...ethic, initiative, organization, and ongoing participation in training and professional development. Perform other duties as assigned... 
    Training
    Work at office
    Flexible hours
    Shift work

    Independent Living Systems LLC

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Agent Post-Training, Artifacts Research. Be the first to apply!