Member of Technical Staff, Post-Training
$200k - $350kHandshake
About Handshake
Handshake was founded on a simple belief that everyone deserves a path to a great career, regardless of where they went to school or who they know. Today, we power 25 million job seekers, 1 million+ employers, and 1,600 educational institutions.
In 2025, we started Handshake AI and built the fastest-growing AI data business in history. We work directly with frontier AI lab researchers to create evaluations, publish benchmarks, and push the boundary of data. We've grown from $0 to ~$1B run rate and pay ~$60M to over 30K individuals every month.
About Handshake Labs
Handshake Labs is building external AI products, research platforms, and customer-facing AI systems. We are evolving work that is often custom-built for an individual partner into reusable products and platforms that improve with every deployment.
Our work spans the full post-training loop: designing evaluations and training environments, building high-quality data and feedback systems, running experiments, and turning what works into durable infrastructure. For example, we are developing agents that can analyze long, complex coding-agent sessions in days rather than weeks—with expert review and calibration built into the system.
The Role
We are hiring a Member of Technical Staff, Post-Training to help define and build this new organization. This is a broad, high-ownership role for researchers who build. You may come from research science, research engineering, machine learning engineering, or a closely related background; what matters is the ability to reason deeply about model improvement and turn that reasoning into reliable systems.
You will partner with researchers, domain experts, and customers to turn ambiguous post-training questions into experiments, evaluation frameworks, data pipelines, and products. Early members of the team will have unusual influence over our technical direction, operating culture, and the reusable systems we build.
We care more about demonstrated research capability, technical judgment, and a builder’s mindset than a specific title, degree, or career path.
Location: San Francisco & Mountain View preferred; we are open to exceptional candidates in other locations.
What you’ll do
Design post-training systems and methodologies for frontier models, including supervised fine-tuning, reinforcement learning, preference optimization, reward modeling, and related approaches.
Translate open-ended research or partner needs into clear hypotheses, experiments, evaluation plans, and production-quality implementations.
Build and improve evaluation frameworks, benchmarks, training environments, data-processing pipelines, and quality-control systems.
Run fast, rigorous iteration loops: prototype, evaluate, interpret results, and turn learnings into the next system or product.
Partner directly with AI researchers and domain experts to develop high-signal data, feedback, and evaluation methods.
Identify repeatable patterns across engagements and productize them into reusable software and platforms.
Raise the technical bar through strong design judgment, clear communication, code quality, and mentorship.
Contribute to the field through benchmarks, open-source tools, research, and technical writing where it creates leverage.
What we’re looking for
3+ years of demonstrated strength in post-training, fine-tuning, or model-evaluation work. Relevant experience may include RL, SFT, LoRA/PEFT, full fine-tuning, RLHF, DPO, PPO, reward modeling, or training environments.
Strong Python skills and the ability to write clean, efficient, scalable software.
Hands-on experience with modern ML tooling, particularly PyTorch and large-scale data, training, or evaluation workflows.
Sound experimental judgment: you can form hypotheses, choose meaningful metrics, diagnose failures, and distinguish signal from noise.
Experience designing systems—not only implementing specifications—including the ability to make tradeoffs around quality, scale, reliability, and reuse.
Comfort operating in an ambiguous, fast-moving environment with substantial ownership.
Collaborative, low-ego communication and the ability to work effectively with researchers, engineers, domain experts, and customers.
Especially compelling experience
Building or operating large-scale ML training, inference, data, or evaluation systems.
Developing LLM/agent benchmarks, evaluation methodologies, annotation systems, or data-quality frameworks.
Research or applied work on reinforcement learning, alignment, model behavior, synthetic data, or human-in-the-loop systems.
Published research, meaningful open-source contributions, or evidence of technical leadership in ML systems or AI research.
Experience productizing research or repeated customer work into robust, reusable platforms.
Why join
Work on problems at the center of how frontier AI systems improve, alongside leading labs and domain experts.
Help build an early technical organization where your work shapes the roadmap, standards, and culture.
Move fluidly from research insight to real-world systems, with the resources and customer context to see those systems matter.
Join a company building durable infrastructure for careers in the AI economy.
Perks
Handshake delivers benefits that help you feel supported—and thrive at work and in life.
The below benefits are for full-time US employees.
Ownership: Equity in a fast-growing company
Financial Wellness: 401(k) match, competitive compensation, financial coaching
Family Support: Paid parental leave, fertility benefits, parental coaching
Wellbeing: Medical, dental, and vision, mental health support, $500 wellness stipend
Growth: $2,000 learning stipend, ongoing development
Office: Commuting support, free lunch, and gym in our SF office
Time Off: Flexible PTO, 15 holidays + 2 flex days
Connection: Team outings & referral bonuses
$150k - $300k
...and more. Roboflow Labs is our applied research unit. We train our own open models (RF-DETR). We build our own benchmarks... ...traverses the physical economy. What You'll Do As a Member of Technical Staff on our Frontier Data team, you'll build the environments, evaluations...TrainingWork at officeRemote work$160k - $190k
...Job Description Member of Technical Staff, Machine Learning, Artificial Intelligence (AI) Required, Work From Home As a Member of Technical... ...: - Build and improve ML components across data, training, evaluation, and inference. - Fine-tune and adapt models...TrainingFull timeRemote workWork from home$240k - $290k
...model – which checkpoint to keep training, what to ship to millions of users... ...within research teams as a member of ML Platform, and you'll set the technical direction for evals across the company... ...expectation for the function as posted, but we are also open to considering...TrainingFull time$402.05k
...likely alongside 1-4 other METR staff. Between exercises, you'll... ...tradecraft analysis, and post-incident reporting. Cloud and... ...how frontier models are trained and deployed (RL post-training... ...matters. Hybrid Preferred: Our technical team members are in our office in...TrainingH1bWork at officeWork from homeHome officeRelocation packageFlexible hours3 days per week- We are hiring one Member of Technical Staff to build the ML systems that make sustained AI research possible. This is a full-time position based... ...devices and nodes. Build distributed execution paths for training, post-training, and agent rollouts; connect research workflows...TrainingFull time
- ...As a Member of Technical Staff on Machine Learning, you will: Contribute to the entire development cycle of our cutting-edge large deep... ...Prepare datasets, design architectures, implement solutions, train and evaluate models to improve our products. Collaborate...TrainingFull timeH1bRemote workVisa sponsorship
- ...financial investors, distils our deep technical research and knowledge into key... ...We are looking for a highly motivated member of technical staff to join our engineering team to work... ...chip projects across both frontier LLM training & inference models Implement modern...TrainingFull timeWork at officeRemote workWorldwide
$150k - $220k
# Founding Member of Technical Staff (MTS)Bay Area, CAFull-time$150k-$220k + equity## About UsVizopsAI... ...You'll Do* •Build backend services for training, evals, telemetry, and online policy... ...Llama Factory, Agent Lightning* •LLM post-training exposure (preference data collection...Training- ...Collaborate with our world model team to build state-of-the-art training and simulation platform for robotics. Build the “World”:... ...around the world. Our founding team, along with many of our team members, has contributed to many of the breakthroughs in AI over the...TrainingFull timeRemote workRelocationVisa sponsorship
$240k - $350k
...Job Title: Member of Technical Staff, Frontier AI Job Type: Full time Location: Remote The... ...environments, simulators, or feedback-driven training systems. Experience improving... ...information contained in this job posting, including but not limited to role responsibilities...TrainingFull timeLocal areaRemote work$220k - $250k
...company focused on personalizing patient care, is hiring a Member of Technical Staff to join their team remotely. The successful candidate will... ...and user experience. ~Create robust evaluation and training frameworks, including automated testing and human-in-the-loop...TrainingFull timeRemote work$200k - $260k
...Job Title: Member of Technical Staff, Coding Research Job Type: Full-time Location: Remote... ...findings into actionable improvements for training and evaluation. Build tooling and infrastructure... ...tool use, reinforcement learning, or post-training methodologies....TrainingFull timeLocal areaRemote work$250k - $350k
Member of Technical Staff — RL Research (New PhD Grad) Seattle, Washington About Nuance Labs Nuance Labs is building photorealistic, real-time... ...deeply technical Member of Technical Staff to own RL and post‑training for large-scale omni models. This posting is aimed at...TrainingInternshipH1bWork at officeVisa sponsorshipShift work- ...Python and large-scale model training. In this role, you will design... ...and directly contribute to technical decisions that optimize performance... .... You will also work on post-training processes, including... ...along with many of our team members, has contributed to numerous...TrainingFull timeH1bRemote workVisa sponsorship
- ...Physical Superintelligence Job Post Physical Superintelligence is a startup with roots... ...AI researchers on agent architecture, training, and evaluation. Write production code that... .... How We Work We hold a high technical bar and give people full ownership of their...TrainingRemote work
- ...field or equivalent practical experience Strong candidates may also have: Experience building or operating AI inference, training, HPC, or neocloud infrastructure Experience with bare-metal provisioning, PXE/iPXE, image pipelines, BIOS/firmware management,...Training
- ...platform for specialized intelligence, enabling companies to build, train, and serve AI models tailored to their own data, workflows, and... ...'ll architect scalable, resilient backend infrastructure, lead technical design discussions, mentor engineers, and establish best...Training
$200k - $300k
...Job Title: Member of Technical Staff, Research Engineering Job Type: Full-time Location: Remote... ...(RL), developing novel environments, training pipelines, and evaluation systems... ...The information contained in this job posting, including but not limited to role responsibilities...TrainingFull timeLocal areaRemote work- ...startup environment with high ownership and ambiguity. Strong candidates may also have Experience building or operating AI inference, training, HPC, or neocloud infrastructure. Experience with bare‑metal provisioning, PXE/iPXE, image pipelines, BIOS/firmware management,...Training
- ...Member of Technical Staff, Frontier AI is a remote review track for evaluating AI outputs across member of technical staff frontier ai specialist... ...and document the right next step so the modeling team can train on it. Why this role matters Member of Technical Staff...TrainingRemote jobHourly payFor contractorsWork experience placement10 hours per week
- ...Role Overview Lead technical ownership at the intersection of research, data, and deployed AI systems to improve model and agent performance... ...converts diverse subject matter expertise into high-quality training data, evaluations, and feedback loops for frontier models and...TrainingFull timeRemote work
$250k - $300k
...organisation delivers the infrastructure powering next-generation AI training and inference at scale. This role offers the opportunity to... ...Skills / Must Have: ~ A track record of impressive technical work you can speak to in depth, the years matter less than the...TrainingFull timeRemote work- Position Summary You'll be a senior technical leader for the pre-training of the omni-models at the core of our stack—single models trained across a wide... ...lunch, and other benefits. The pay ranges noted on our posts are for salary only. Walden Robotics is an equal...TrainingWork from homeFlexible hours
- ...real world. Our team works across model post-training, reinforcement-learning infrastructure,... ...research, open-source contributions, technically ambitious independent projects, or production... ...hybrid role. We currently expect all staff to work from one of our offices at...TrainingWork at officeVisa sponsorshipFlexible hours3 days per week
- ...policy and infra teams to move models from training into deployment. Required... ...other benefits. The pay ranges noted on our posts are for salary only. Walden Robotics is... ...exceptional professionals who combine world-class technical skills with creative vision, grounded in...TrainingWork from homeFlexible hours
- ...Member of Technical Staff, Economics Research is a remote review track for evaluating AI outputs across member of technical staff economics research... ..., and document the correct method so the modeling team can train on it. Why this role matters Member of Technical Staff...TrainingRemote jobHourly payFor contractors10 hours per week
- Position Summary You'll be a senior technical leader for how our foundation models become capable, aligned robot policies. Working alongside... ...on real robots in the real world. Core Responsibilities Post-Training Strategy: Help shape the post-training recipe—SFT, preference...TrainingWork from homeFlexible hours
- ...reinforcement learning and classical control: you'll train policies in large-scale simulation, blend... ...benefits The pay ranges noted on our posts are for salary only. Walden Robotics is... ...professionals who combine world-class technical skills with creative vision, grounded in...TrainingWork from homeFlexible hours
- ...environments to translate mission needs into technical systems. Build and maintain large-... ..., and evaluation pipelines for model training and improvement. Develop infrastructure... ...The information contained in this job posting, including but not limited to role responsibilities...TrainingFull timeLocal areaRemote work
$150k - $300k
...research prototypes. ~ BS or higher in CS, ML, math, or a related technical field. ~ Production-level Python, PyTorch, and the Transformers ecosystem. ~ Experience with LLM evaluation, training, or prompt engineering. ~ Able to articulate failed experiments and...TrainingFull timeLive inRemote workVisa sponsorshipFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Member of Technical Staff, Post-Training. Be the first to apply!
- senior IT support technician Remote
- remote support technician Remote
- tech aide Remote
- senior technical associate Remote
- IT help desk technician Remote
- customer support analyst Remote
- work from home technical support specialist Remote
- product support technician Remote
- mri tech aide Remote
- support technician Remote


