Head of Robot Learning
Anvil Robotics
San Francisco, CA or Hong Kong / Taipei (on-site preferred) | Full-time You’ll be Anvil’s first robot learning hire — the person who takes the papers everyone is retweeting, gets them running on our hardware in weeks, and ships them as demos and guides polished enough that the whole field notices. Anvil is building the platform layer for Physical AI — robotics hardware and software that’s radically more accessible than legacy industrial solutions. In our first 12 months we built and shipped 200+ robots (OpenARM and OpenYAM manipulators, Linux Devboxes, teleop kits, and a UMI-style handheld data-collection device) to customers in 60+ countries, doing over $2.1M in revenue on our own Taipei manufacturing line. And here’s what makes this seat unusual: not how much data we have — how much leverage you’d have over what gets collected. Most robot learning engineers work with whatever data someone else decided to collect, on rigs they can’t change. Here, the entire collection system bends to your judgment — and there’s a factory behind it. We design and build our own UMI-style handheld collector in‑house, on our own Taipei manufacturing line, sitting inside the Asia supply chain: if the data would be better with a different camera, a different mount, a new sync scheme, or a custom fixture that doesn’t exist yet, you prescribe it and it gets built — in weeks, not procurement quarters. We have factory access most teams can’t get — our own facility and our investors’ and partners’ plants — meaning differentiated, real‑industrial‑task data rather than the same recycled public datasets. And we have people who can do the collection work if you write the protocol: you prescribe what good data looks like, they collect it. Let’s be honest about scale, because it matters: this is not a foundation‑model data operation, and we’re not pretending it is. It’s the setup to reach LeRobot‑shirt‑folding scale — hundreds of high‑quality demonstrations of the right task, on tooling shaped to your spec — with more control and less friction than almost anyone in the field gets. The situation you’re walking into: We have 200+ robots in the field and zero dedicated ML function. Model training happens in the gaps between the founders’ and controls engineers’ actual jobs. You are hire #1 for the entire robot learning function, and for the foreseeable future the team is you. We ship a UMI‑style handheld data‑collection device — and nobody has yet closed the loop of training a policy purely from its data and running it on our arms. Validating that pipeline end to end is a product decision waiting on you, and it’s one of your first deliverables. Robot learning is compounding weekly — VLAs, diffusion policies, ACT, world models. At our stage the highest‑leverage move is not novel research; it’s replicating the best public work on our hardware fast, and publishing it. Think the LeRobot shirt‑folding project — data to deployment, in the open — running on Anvil arms, with our name on the guide. Demos are not vanity here. A flagship replication is simultaneously marketing to the exact community that buys us and the enablement guide our customers follow. Which means the last 10% — the clean repo, the honest success rates, the written guide, the good video — is where most of the value lives. We need a builder with a real knack for polish before calling a project done. The collection system above is built but undirected. The UMI exists and can be revised to your spec, the operators exist, the factory access exists — but nobody with ML judgment decides what to collect, how, and what “good” looks like. The leverage is sitting there; the person who prescribes it doesn’t exist yet. What you’ll own: Training pipelines, end to end: teleop and UMI data ingestion, dataset formats and quality triage, training jobs, and an eval harness with honest success‑rate protocols — built so that eventually someone who isn’t you can train a model. Paper replication > published demos: picking the highest‑leverage public work (folding‑class manipulation, VLA fine‑tunes, diffusion policies), getting it running on Anvil hardware in weeks, and shipping it as a public demo video + reproducible guide. UMI pipeline validation: being the person who proves — or fixes — the path from our handheld data collector to a working policy on an OpenARM. The UMI is built in‑house, so your findings don’t end as feedback — they become hardware revisions you prescribe. The data flywheel: DAgger / human‑in‑the‑loop correction workflows on our teleop stack, so policies improve from intervention data instead of plateauing after the first training run — and, as your protocols mature, scaling collection beyond yourself: designing what dedicated data‑collection operators record in our factory and, over time, in partner facilities. The polish bar: nothing you ship stops at “works on my machine.” Every project ends with the video, the guide, and the repo someone else can run. What the first 100 days look like: By day 30: your rig is set up, and a first policy (ACT or Diffusion Policy class) trained on Anvil‑collected data is running on real hardware. You have a written map of where our stack fights you. By day 60: the UMI pipeline is validated end to end — a policy trained purely from handheld‑collected data, running on an OpenARM — with a clear‑eyed writeup of what’s broken in the pipeline and what to fix. By day 100: your first flagship replication (folding‑class) is live on Anvil hardware and published — demo video, honest success rates, reproducible guide. Who you are: You’ve personally trained and deployed imitation‑learning policies — ACT, Diffusion Policy, VLA fine‑tunes — on real robot arms. Not cloud benchmarks: real motors, real cameras, real failures. You’re fluent in the layer under the model, because that’s where deployments die: action chunking and temporal ensembling, inference latency versus control‑loop frequency, camera synchronization and timestamp alignment, joint‑space versus cartesian command interfaces. You replicate papers in weeks. Given a paper and a repo, you know what will transfer, what won’t, and where the unreported gotchas hide. You have data instincts: you can scrub through teleop demonstrations and see what’s wrong — inconsistent grasps, occlusions, timing skew — before wasting a training run on it. You’re a finisher. Your projects end with a guide, a video, and a repo with a README — and you know the difference between a demo that worked once on camera and one with a measured success rate you’d defend. You’re honest about results. You report n, the eval protocol, and the failure modes — especially in public. You’re comfortable being the only ML person in a fast, lean, founder‑led company, setting your own agenda and shipping on a weeks‑not‑quarters cadence. Bonus: LeRobot or similar open‑source contributions; DAgger / interactive imitation learning experience; public demos that got real reach; RL fine‑tuning on real hardware. Based in or willing to relocate to San Francisco, Hong Kong, or Taipei. This role needs to sit with robots, cameras, and a GPU box, wherever that is — Taipei puts you next to the hardware team, SF next to the demo room and customers. Flexible hours for cross‑timezone collaboration either way. Education & experience: Master’s or Bachelor’s in CS, robotics, or a related field — a PhD is explicitly not required or expected, and a publication record is not the bar. Your portfolio is: policies you personally trained running on real hardware, with the repos and videos to prove it. Years matter less to us than trajectory. A typical req for a "Head of Robot Learning" would ask for a PhD plus 5–8 years; we’re looking for 2–4 unusually fast‑growing years — or 1–2 on a steep curve — spent as the research engineer beside a strong robot learning lead: the person who made the lab’s or team’s work actually run on hardware, closed a growing share of the hard problems personally, but never owned the direction because someone above them did. This role is the first time the wheel is yours. What this role is not: Not a research role. No publication mandate, no novel architectures for their own sake. You’re making the frontier run on our hardware, not extending it. Not a cloud ML role. If your experience is training and eval with no physical robot in the loop, this isn’t the seat. Not a management seat. "Head of" means you own the function; the team is you, possibly for a year or more. Not a solutions‑architect role. The deliverable is never a plan or a memo — it’s a policy running on an arm and the published artifact around it. Not a role for 90%-done builders. If your projects historically end at "it worked, moving on" — before the writeup, the video, the reproducible repo — our polish bar will chafe daily. What We Offer Health and Wellness Compensation and Support #J-18808-Ljbffr Anvil Robotics
- BridgeBio, a biopharma company, seeks a Senior Director of Training and Learning Excellence (Commercial) in San Francisco/Palo Alto. You will architect a scalable learning program, lead cross‑functional teams, and ensure training supports product launches and market access...Suggested
$230k - $400k
...about each other and our customers, we'd love to meet you. About the Role We are looking for an engineering leader to lead machine learning efforts across Hightouch. While hundreds of companies use Hightouch today to sync data into their SaaS systems to automate and...SuggestedRemote jobImmediate start- ...in San Francisco, CA, with a global presence across the United States, EMEA, and APAC. Role Overview We are looking for a Head of Machine Learning to lead the development of the next generation of our machine learning platform. This role combines technical leadership,...Suggested
$350k
...Head of AI Type: Full-time Base Salary: $350,000 - $450,000 (Plus Bonus & Equity) Base Pay Range: $350,000.00/yr - $450,00... ...experiment, and engineer in the real world. The work spans robotics, machine learning, autonomous experimentation, and applied science, with a focus...SuggestedFull time- Role Overview Anvil Robotics is building the Physical AI platform for robotics builders—modular... ..., and data tooling. We're hiring a Head of Applied ML to eliminate the biggest bottleneck... ...informed about what's happening in robot learning and what it means for our product and...Suggested
$130k - $250k
...leader to join our marketing organization as Head of Demand Generation & ABM. This role... ...outcomes. Implement a continuous test‑and‑learn approach across messaging, channels, and... ...Knowledge of AI, automation, and robotics industries strongly preferred Knowledge...$190k - $240k
.... Partner with Autonomy and Platform Software teams to ensure robotic system capabilities are surfaced clearly and safely through web... ...engineering leaders and teams; create a culture of high standards, fast learning loops, product ownership, and customer accountability. Partner...Full timeTemporary workFlexible hours- ...and engineers weeks of manual effort. We're looking for a Head of Machine Learning to build and grow the organization that turns our data into... ...Mach9 — remote sensing, geomatics, autonomous driving, or robotics. Experience leveraging large unstructured datasets, especially...Work experience placementRemote work
$150k - $250k
Head of Machine Learning We are currently seeking a Head of Machine Learning to lead our machine learning team. You will own the machine learning lifecycle from research to deployment, driving the computer vision capabilities that power our category‑defining app. The...$250k - $350k
...based on your skills and experience — talk with your recruiter to learn more. Base pay range $250,000.00/yr - $350,000.00/yr Direct... ...from Harnham Senior Recruitment Consultant (Phoenix) at Harnham HEAD OF SOLUTIONS ENGINEERING SAN FRANCISCO, BAY AREA ONSITE $250,000...Full timeFlexible hours- ...Job Title: Head of AI Engineering Location: NYC or SF (On-site, 5 days a week) About Highlight AI We're a small, senior team building... ...and evals Research, prototype and deploy state-of-the-art machine learning text and OCR models (e.g., transformer architectures, computer...Work at officeRelocationRelocation packageFlexible hours
$203k - $399k
...leading cross-functional organizations of 100+ people. Product-Led Mindset: A product-hacker mentality. Prioritizes shipping and learning over perfection, but understands the reliability requirements of enterprise-grade software. Strategic Communication: Ability to distill...Local areaFlexible hours$260k - $300k
...support Employer-paid basic life and disability coverage Annual learning and development stipend to fuel your professional growth Daily meals... ...in and out of the US. About the Role We’re looking for a Global Head of Benefits to lead and improve our benefits programs, policies...Full timeWork at officeLocal areaWork from homeRelocation packageFlexible hours- ...a role currently held by the co-founder). Collaborate with the Head of CS and Head of Sales to define and refine the end-to-end process... ...insights from prospects and customers — translating field learnings into product direction. Accountable for three outcomes: deals move...Contract workLocal areaFlexible hours
$140k - $165k
...Head of Provider Success (Confidential HealthTech SaaS) Tagged: Customer Success , Remote , Start-up We’re working with a high... ...impact—this might be the only role that makes sense. Apply now to learn more. Compensation $140-165k base Posted Friday,...Work at officeRemote workFlexible hours$165k - $190k
...insurance Wellness allowance Company-sponsored personal and professional development Partnerships with Ethena and monthly Lunch & Learns Wellbeing: access to many wellbeing perks, including Peloton, Fetch, OneMedical, Headspace care+, etc. Caregiver Support: company seed...Work at office3 days per week$4,500 - $6,000 per month
..., Collaboration, and Curiosity. Job Duties and Expectations All head coaches and assistant coaches report directly to the Athletic Director... ...student‑athletes by using positive methods to create an ideal learning environment therefore helping them through their life journey,...Immediate startWeekend workAfternoon shiftWeekday work$238k - $322k
Amer Head Of Channel Partnerships As the AMER Head of Channel Partnerships, you will shape and execute our Channel Partnerships Program... ...about AI tools and emerging technologies, with a willingness to learn and leverage them to enhance productivity, collaboration, or...Temporary workWork at officeLocal areaWork from homeWorldwide$300k
Base pay range $300,000.00/yr - $500,000.00/yr Head of Autonomous Systems We are partnered with a highly technical DeepTech Research... ...technical engineers and researchers from backgrounds including robotics, AI systems, control, distributed infrastructure, and applied ML...Full timeImmediate start- ...by acquiring other local firms, which they call “tuck-ins.” The Head of Integration will architect and scale the process of how their... ...professionals with a track record of excellence, a commitment to continuous learning, and a burning desire to win. Meaningful Impact: Build a...Local areaImmediate start
$208k - $260k
...is precise, proactive, and scalable. We are seeking a Director, Head of License & Exam Management to own this pillar end-to-end. This... ...cause analysis of findings, corrective action tracking, lessons learned documentation, and procedure updates. Develop pre‑exam...Full timeLocal areaRelocation packageFlexible hours2 days per week3 days per week- The opportunity We are seeking a Head of Lab Platform to join our team working at the interface of generative AI and synthetic biology... ...closes the loop between biological experimentation and machine learning at scale. You will oversee a team of scientists, set the strategic...Flexible hours
$245k - $337k
...At Faire, we’re using the power of technology, data, and machine learning to connect this thriving community of entrepreneurs across the... ...to life across every customer touchpoint. We are looking for a Head of Creative to lead Faire’s creative team and help take the brand...For contractorsWork experience placementWork at officeLocal areaRemote workMonday to FridayFlexible hoursAfternoon shift3 days per week$197.6k - $272k
...to grow its Government Relations function, we are looking for a Head of Global Policy to support our engagements with Congress, Government... ...health, dental and vision coverage, retirement benefits, a learning and development stipend, and generous PTO. Additionally, this role...- ...infrastructure to every enterprise that runs the real economy. Learn more about our vision in our manifesto. About the Role This is... ...Growth team, based in San Francisco, reporting directly to the Head of Growth and working closely with founders to shape how HappyRobot...Worldwide
- Head of Content (Video Producer + Social MediaManager) Location: San Francisco, CA (On-site at REK HQ) Company: REK Inc. REK is creating a new global sport VR-controlled humanoid robot fighting. We blend robotics, gaming, and live entertainment into cinematic, high-energy...Local area
$197.6k - $272k
...to grow its Government Relations function, we are looking for a Head of Global Policy to support our engagements with Congress, Government... ...health, dental, and vision coverage; retirement benefits; a learning and development stipend; generous PTO; and potentially a...Full time$238k - $294k
...track record in a high-growth tech environment, with demonstrated experience distilling complex technical subject matter (e.g. AI, robotics, or engineering) into clear, engaging content for diverse audiences. Experience supporting C-Suite executives, including...Full timeImmediate startRemote workShift work- A leading mobile advertising company is seeking a Head of Machine Learning to oversee the development of their machine learning platform. The ideal candidate will have over 7 years of experience in production machine learning systems, strong programming skills in Python...
- ...just tech Twitter. You can be smart without being inaccessible. You're data‑informed but not data‑paralyzed: You track what works, learn from what doesn't, and use metrics to improve - but you also trust your instincts and don't need a dashboard to tell you a piece is...Temporary workWork at officeRelocation package
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Head of Robot Learning. Be the first to apply!
- head of rewards San Francisco, CA
- head of seo San Francisco, CA
- head credit administration San Francisco, CA
- head coach San Francisco, CA
- head San Francisco, CA
- head of portfolio management San Francisco, CA
- head golf professional San Francisco, CA
- head of architecture San Francisco, CA
- hilton head hospital
- head of rewards

