Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Staff ML Engineer, Agent Training & Environments

$250k - $280k
Full-time

Labelbox

Shape the Future of AI


At Labelbox, we're building the critical infrastructure that powers breakthrough AI models at leading research labs and enterprises. Since 2018, we've been pioneering data-centric approaches that are fundamental to AI development, and our work becomes even more essential as AI capabilities expand exponentially.

About Labelbox


We're the only company offering three integrated solutions for frontier AI development:


  1. Enterprise Platform & Tools : Advanced annotation tools, workflow automation, and quality control systems that enable teams to produce high-quality training data at scale

  2. Frontier Data Labeling Service : Specialized data labeling through Alignerr, leveraging subject matter experts for next-generation AI models

  3. Expert Marketplace : Connecting AI teams with highly skilled annotators and domain experts for flexible scaling

Why Join Us



  • High-Impact Environment : We operate like an early-stage startup, focusing on impact over process. You'll take on expanded responsibilities quickly, with career growth directly tied to your contributions.

  • Technical Excellence : Work at the cutting edge of AI development, collaborating with industry leaders and shaping the future of artificial intelligence.

  • Innovation at Speed : We celebrate those who take ownership, move fast, and deliver impact. Our environment rewards high agency and rapid execution.

  • Continuous Growth : Every role requires continuous learning and evolution. You'll be surrounded by curious minds solving complex problems at the frontier of AI.

  • Clear Ownership : You'll know exactly what you're responsible for and have the autonomy to execute. We empower people to drive results through clear ownership and metrics.

Role Overview


Labelbox is the RL data factory for advancing frontier agent capabilities. We build the data, environments, and evaluations that frontier labs use to train and judge their agents.

This role sits where training meets infrastructure. You will run the experiments and build the systems that run them: environments agents act in, verifiers that decide whether they succeeded, and the fine-tuning pipelines that turn that signal into a better model. We're looking for someone who does both halves — the engineering throughput of a strong platform engineer, and real depth in post-training agents.

The bar is high: engineers with strong judgment who set technical direction, turn prototypes into reliable systems fast, and are at the frontier of agent-first engineering practice.

 

What you'll work on



  • RL environments for agentic tasks: task definitions, tool surfaces, state and reset semantics, reward design — and the harness that runs thousands of them in parallel.

  • Verifiers and graders: programmatic checks, LLM judges, rubric pipelines, View email address on us.fitly.work scoring. Deciding what "the agent succeeded" means, and making that judgment trustworthy at scale.

  • Fine-tuning pipelines that turn evaluation signals into measurable agent improvements — SFT and RL, from data collection through training to checkpoint evaluation.

  • Eval systems that run millions of agent trajectories to measure model and product quality.

  • Training and serving infrastructure that scales to the throughput frontier labs need: multi-launcher orchestration, long-running job fault tolerance, cost accounting.

What we're looking for


As an engineer


  • A 3+ year track record of shipping systems that customers and other engineers still rely on.

  • Exceptional throughput, without the quality tax. You ship a lot, you review a lot, and the v1 you ship becomes the foundation the rest of the team builds on.

  • Strong system and API design judgment. Hard architecture calls land with you: you make them, defend them under pressure, and update fast when someone else is right.

  • You ship production code with coding agents daily. You know where they break and what it takes to make them reliable, and you use that to move the whole team faster.

  • You build the substrate other people's work runs on — tooling, CI, harnesses, libraries — and you treat that as the job, not a distraction from it.

  • You move fast in ambiguous, startup-pace environments, with influence over authority.

  • Deep proficiency in Python, and comfort across the rest of the stack.

As an RL post-training practitioner


  • You have fine-tuned models for agentic tasks and made them measurably better. SFT plus at least one RL method (GRPO, PPO, DPO, or similar) in production.

  • You have built environments agents operate in, and you know why reward and task design is where most of the difficulty actually lives.

  • You have designed verifiers or graders for open-ended work, and you know how they get gamed.

  • You debug training runs forensically and methodically.

  • You reason about compute-economics. You know what an experiment costs, when a run is not worth finishing, and how to get the same signal for a tenth of the spend.

  • You write up what you learned so it changes what the team does next.

 

Nice to have



  • Experience with agent harnesses and coding agents as subjects of training and evaluation.

  • Multi-tenancy and isolation for untrusted agent execution: sandboxing, egress control, credential handling.

  • Background in production distributed systems, ML infrastructure, or data systems at scale.

  • Experience working directly with frontier labs or other highly technical customers.

 

Our Technology Stack


Our engineering team works with a modern tech stack designed for scalability, performance, and developer efficiency:


  • Frontend: React.js with Redux, TypeScript

  • Backend: Node.js, TypeScript, Python, some Java & Kotlin

  • APIs: GraphQL

  • Cloud & Infrastructure: Google Cloud Platform (GCP), Kubernetes

  • Databases: MySQL, Spanner, PostgreSQL

  • Queueing / Streaming: Kafka, PubSub

Labelbox strives to ensure pay parity across the organization and discuss compensation transparently. The expected annual base salary range for United States-based candidates   is below. This range is not inclusive of any potential equity packages or additional benefits. Exact compensation varies based on a variety of factors, including skills and competencies, experience, and geographical location.

Annual base salary range

$250,000 - $280,000 USD

Life at Labelbox



  • Location : Join our dedicated tech hub in San Francisco

  • Work Style : Hybrid model with 3 days per week in office, combining collaboration and flexibility

  • Environment : Fast-paced and high-intensity, perfect for ambitious individuals who thrive on ownership and quick decision-making

  • Growth : Career advancement opportunities directly tied to your impact

  • Vision : Be part of building the foundation for humanity's most transformative technology

Our Vision


We believe data will remain crucial in achieving artificial general intelligence. As AI models become more sophisticated, the need for high-quality, specialized training data will only grow. Join us in developing new products and services that enable the next generation of AI breakthroughs.

Labelbox is backed by leading investors including SoftBank, Andreessen Horowitz, B Capital, Gradient Ventures, Databricks Ventures, and Kleiner Perkins. Our customers include Fortune 500 enterprises and leading AI labs.

Your Personal Data Privacy : Any personal information you provide Labelbox as a part of your application will be processed in accordance with Labelbox’s Job Applicant Privacy notice .

Any emails from Labelbox team members will originate from a @labelbox.com email address. If you encounter anything that raises suspicions during your interactions, we encourage you to exercise caution and suspend or discontinue communications.

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Staff ML Engineer, Agent Training & Environments in Remote vacancy
  • Staff Machine Learning Engineer, Agent Memory & Reasoning (University) Location Employment Type Full time Location...  ...least the next six months: no model training, no fine tuning, no deep GPU or CUDA...  ...fundamentals to go with your ML and agent experience. Practical fluency... 
    Training
    Full time
    Remote work
    Shift work

    Wand Inc.

    Brooklyn, NY
    3 days ago
  • $207k - $300k

     ...generated content through advanced context engineering and agentic feedback loops to identify...  ...(AIGC) to fulfill user interests. As a Staff Machine Learning Engineer focusing on Search...  ..., experience, and relevant education or training. US: $207000 - $300000 (USD) + 20% bonus... 
    Training

    Google

    Mountain View, CA
    5 days ago
  • $189.3k - $320.7k

     ...behavior across real-world scenarios.As a Staff ML Engineer on the Prometheus team within the...  ...of autonomous vehicle development—from training and validation to testing and safety....  ...Experience deploying ML models into production environments and understanding end-to-end... 
    Training
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, TX
    5 days ago
  •  ...-connected world.Role OverviewAs our Staff Software Engineer, ML infra Engineer for Search & Discovery...  ...structured and unstructured data needed to train complex ML models and efficiently...  ...committed to providing a safe work environment for its employees and its consumers.If... 
    Training
    Temporary work

    Coupang

    Mountain View, CA
    1 day ago
  • $189k - $300k

     ...works on and delivers ML models to the product...  ...foundation model pre-training and fine-tuning with data...  ...-impact team of AI/ML engineers, data scientists and...  ...autonomous vehicles. As a Staff AI/ML Engineer in the...  ...into production environments and understanding end-... 
    Training
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    5 days ago
  •  ...with Moveworks’ Reasoning Engine and natural language...  ...Role Moveworks' AI agents don't just generate text...  ...judge good enough to train against . The same calibrated...  ...role. It's applied ML at a point where the...  ...multi-tenant enterprise environments.   What you get to... 
    Training
    Work at office
    Remote work
    Flexible hours

    ServiceNow

    Mountain View, CA
    8 days ago
  • $252k - $315k

     ...build production AI agents that automate complex...  ...one of the hardest engineering challenges.As a Staff Frontier Agent Engineer...  ....Unlike traditional ML roles that focus on...  ...in high-stakes environments.Collaborate with infrastructure...  ...education or training. Scale employees in... 
    Training
    Full time

    Scale AI

    San Francisco, CA
    3 days ago
  • $189.3k - $290.7k

     ...behavior across real-world scenarios.As a Staff ML Infra Engineer, you will drive the development of...  ...enable rapid dataset generation, training, evaluation, and iteration of our most...  ...providing an inclusive workplace creates an environment in which our employees can thrive and... 
    Training
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, CA
    5 days ago
  • $171.7k - $303.9k

     ...the world!The Data Labeling Engineering team designs, builds, and operates...  ..., data engineering, and AI/ML, defining the strategies,...  ...controls that create reliable training data at scale. Our tools and...  ...inclusive workplace creates an environment in which our employees can thrive... 
    Training
    Full time
    Local area
    Remote work
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, TX
    2 days ago
  • Wand is seeking a Staff Machine Learning Engineer to join University, a new team focused on agent memory, evolution, and reasoning. You will design and build memory systems, ensure secure, scalable deployment in the cloud, and work with a world-class team to push the frontier... 
    Remote job

    Wand Inc.

    Brooklyn, NY
    3 days ago
  •  ...understand, and navigate any complex environment, enhancing the usability and safety of...  ...Gaia is Wayve’s video world model: trained on large-scale driving video, it predicts...  ...rare or safety-critical events. As a Staff ML Engineer on Gaia, you’ll own and drive work on... 
    Training
    Full time
    Work at office
    Work from home

    Wayve

    United Kingdom
    7 days ago
  • Role Description As Staff ML Engineer, you will lead the Scene Generation team, setting its technical...  ...perception simulation and AV 3.0 training. ~Own the neural rendering...  ...integration of the framework in a cloud environment and automate the pipeline to scale target... 
    Training
    Full time
    Work experience placement
    Immediate start

    Torc Robotics

    Remote
    a month ago
  • $251k - $310k

     ...Conduct comprehensive experimentation to train and deploy state-of-the-art Multimodal LLMs...  ...and Radar.. Partner effectively with engineering and research teams across Waymo to deploy...  ...(e.g., xprof), and debugging of ML models. You have: PhD or Masters... 
    Training
    Full time
    Temporary work
    Remote work

    Waymo

    New York, NY
    3 days ago
  • $220k - $247k

     ...Do As a  Senior Staff Machine Learning Engineer , you will operate...  ...design of large-scale ML systems and shared...  ...capabilities across agents, text, image, audio,...  ...scalable ML platforms (training, evaluation,...  ...high-growth startup environments   Location This... 
    Training
    Full time
    Work at office
    Immediate start
    Flexible hours
    3 days per week

    Typeface

    Remote
    3 days ago
  •  ...Description The Consumer Agent team builds Shopify'...  ...buy. As a Senior Staff Applied Machine Learning Engineer, you'll be the...  ...orchestration, model training and distillation, search...  ...model labs and ML teams across Shopify...  ...paced, cross-functional environment. ~Able to work with... 
    Training
    Full time
    Shift work

    Shopify

    Remote
    22 days ago
  • $281k - $356k

     ...machine learning models to deliver training and evaluation data for...  ...for researchers and software engineers who are passionate about developing...  ...in Python and standard ML frameworks (e.g., JAX, TensorFlow...  ..., or complex simulation environments ~ Deep understanding of state... 
    Training
    Full time

    Waymo

    Remote
    3 days ago
  • $251k - $310k

     ...simulations of realistic environments for testing, training, and validation of the Waymo...  ...of machine learning (ML) engineers, software engineers, and ML...  ...world, encompassing realistic agents, roads, traffic systems,...  ...you will report to a Senior Staff Engineering Manager. You... 
    Training
    Full time
    Remote work

    Waymo

    Remote
    a month ago
  • $189.3k - $320.7k

     ...behavior across real-world scenarios.As a Staff AI/ML Future Sensing Engineer in the Embodied AI organization,...  ...efforts spanning data curation, training, validation, performance...  ...preparing ML models for production environments and understanding end-to-end deployment... 
    Training
    Full time
    Local area
    Remote work
    Work from home
    Relocation
    Relocation package
    Flexible hours

    General Motors

    Sunnyvale, TX
    1 day ago
  • $218.8k - $335.3k

    Job DescriptionStaff AI/ML Engineer, AV ML Infra We’re General Motors...  ...optimizes large-scale ML training and inference across cloud and...  ...commercialization.Position Overview: As a Staff AI/ML Engineer, you will be...  ...workplace creates an environment in which our employees can... 
    Training
    Full time
    Local area
    Work from home
    Flexible hours

    General Motors

    Austin, TX
    1 day ago
  •  ...Quilter, we are helping electrical engineers save time and accomplish more...  ..., electromagnetic simulation, ML/AI, and high-performance...  ...We’re looking for a Senior or Staff ML Systems Engineer to join Quilter...  ...the full ML lifecycle: training pipelines, data generation and... 
    Training
    Full time
    Remote work

    Quilter

    Remote
    10 days ago
  • YO AI Labs seeks experienced Senior Software Engineers to support an AI training project by creating reinforcement learning environments that evaluate AI models on complex software engineering tasks using Model Context Protocol (MCP) tools. You will design reproducible... 
    Training
    Remote job

    YO AI Labs

    San Francisco, CA
    5 days ago
  • YO AI Labs is seeking a Senior Software Engineer to support an AI training project by creating reinforcement learning environments that evaluate AI models on complex software engineering tasks using Model Context Protocol (MCP) tools. You will design reproducible environments... 
    Training
    Remote job
    Contract work

    YO AI Labs

    Los Angeles, CA
    5 days ago
  •  ...Perplexity is seeking experienced ML engineers to design, build, and optimize the recommendation...  ...Experience with large-scale ranking and training infrastructure (multi-stage retrieval...  ...something and hope. They must have armies of agents and workers who can constantly work in... 
    Training
    Full time

    Perplexity®️

    San Francisco, CA
    3 days ago
  • $195k - $298k

     ...Technical Center - Cole Engineering Center Podium or...  ...assistance. About the Team:The ML Compute Platform is...  ...platform supports the training and deployment of...  ...Role:We are seeking a Staff ML Engineer to help build...  ...dynamic, multi-tasking environment with ever-evolving... 
    Training
    Full time
    Local area
    Work from home
    Relocation package
    Flexible hours

    General Motors

    Warren, MI
    4 days ago
  • $262k - $361k

     ...energy, AI, software, engineering, and product to build...  ...development of production ML/AI systems. You will...  ...mentoring senior and staff-level engineers, establishing...  ...in production environments.Establish scalable architectural...  ...experience building, training, and deploying large-... 
    Training
    Full time
    Remote work
    Flexible hours

    X Company

    Mountain View, CA
    1 day ago
  • $244.14k - $413.16k

     ...global driving. As a Senior Staff Machine Learning Engineer, you will architect the...  ...simulations for closed-loop training and evaluation.Policy Evolution...  ...planning and complex multi-agent interactions.Scaling & Data...  ...resources to support your ML model development/research.... 
    Training
    Full time
    Overseas

    XPENG Motors

    Santa Clara, CA
    1 day ago
  •  ...specialized intelligence trained on their own knowledge and...  ...are seeking an experienced Staff Machine Learning Engineer with a strong background in...  ...teams to integrate ML models into our platform.Conduct...  ...collaboratively in a team environment.Excellent problem-solving... 
    Training
    Live out
    Work at office
    Local area
    Remote work

    webAI

    Austin, TX
    5 days ago
  • $254k - $350k

     ...efficient navigation in complex environments through sophisticated detection,...  ...and tracking systems.As a software engineer on the perception mapping team, you...  ...initiative. You will design, train, validate, and integrate into the stack ML models that detect semantic map elements... 
    Training
    Full time
    Temporary work
    Relocation package

    Zoox

    Foster, CA
    1 day ago
  • $245k - $319k

     ...We are looking for a Senior Staff Software Engineer, Machine Learning to be a...  ...the next generation of AI/ML models and systems that connect...  ...right buyers.Our technical environment is evolving rapidly. We are...  ...on both large-scale model training and serving.A foundational... 
    Training
    Full time
    Work at office
    Local area
    Remote work
    Visa sponsorship
    Flexible hours

    Etsy

    Brooklyn, NY
    4 days ago
  • $214.67k - $322k

     ...friendly, rewarding, and diverse environment for every OK-er.OKX is part...  ...different from conventional ML engineering. The data spans on-chain...  ...investigations, and intelligent review agents are part of the team’s daily...  ...feature pipelines, training workflows, model serving,... 
    Training

    OKX

    San Jose, CA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Staff ML Engineer, Agent Training & Environments. Be the first to apply!