Post-Training Researcher: AI Evaluation & Scaling
Thinking Machines
Thinking Machines is seeking post-training researchers in San Francisco to bridge theory and practice, writing high‑performance code and evaluating post‑training recipes that scale across datasets and models. You’ll blend fundamental research with engineering, iterating on evaluations, debugging training configurations, and publishing insights to move the field forward. Evergreen role with visa sponsorship and competitive compensation. #J-18808-Ljbffr Thinking Machines
$150k - $250k
About Distyl AI Distyl is an applied AI technology... ...organizations.We research and deploy technologies... ...Key ResponsibilitiesThe Post-Training team focuses on... ...Researchers develop and evaluate techniques such as supervised... ..., effectively, and at scale across industriesWhat...TrainingWork at office3 days per week$197.3k - $313.7k
...SalesforceSalesforce is the #1 AI CRM, where humans... ..., diverse team of researchers at Agentforce... ...by the global scale and trust of... ...DoingDesign, implement, and train novel deep... ...developing, deploying, and evaluating machine learning... ...opt out options.Posting...TrainingFull timeImmediate startRemote work$216.3k - $280.8k
...close on: 08/03/2026Job posting may be removed earlier... ...the TeamAt Foundation AI, we are leading frontier AI research across Cisco. Our mission... ...systems, scalable training algorithms, evaluation science, inference optimization... ...challenging enterprise-scale problems.Your ImpactAs...TrainingFull timeTemporary workLocal areaFlexible hours- ...creating trustworthy and reliable AI systems, changing banking... ...touches every aspect of the research life cycle, from partnering... ..., from design through training, evaluation, validation, and implementation... ...record of delivering models at scale both in terms of training...TrainingFlexible hours
$262.5k - $299.6k
...Applied Researcher II (AI Foundations, LLM Core and Agentic AI) Overview... ...development, from design through training, evaluation, validation, and... ...record of delivering models at scale both in terms of training data... ...to pay at the time of this posting. Salaries for part‑time...TrainingFull timePart timeLocal areaFlexible hours$262.5k - $299.6k
Applied Researcher II (AI Foundations) Overview: At Capital One, we are... ...development, from design through training, evaluation, validation, and... ...record of delivering models at scale both in terms of training data... ...to pay at the time of this posting. Salaries for part-time...TrainingFull timePart timeLocal areaFlexible hours$200k - $280k
...architectures, engines) and post-training / RL systems. We build... ...can run at production scale. Our mandate is to... ...stack. Have a solid research foundation in your area... ...collection and evaluation cheaper. Use these pipelines... ...engineering. About Together AI Together AI is a...TrainingFull time- We are seeking an Edge AI Research Scientist to develop next-generation... ...architectures Efficient training and inference techniques Contribute... ...frameworks. Performance Evaluation Develop rigorous benchmarking... ...distilling, or evaluating large-scale foundation models using distributed...Training
- AI Researcher (Computer Vision/Multimodal/Generative AI) About the Role... ...architectures, algorithms, and training strategies that improve... ..., research innovations must scale to real‑world deployment environments... ...product differentiation. Evaluate new model paradigms for scalability...Training
- ...real neurons to improve AI models. We study how... ...computational neuroscience, AI research and software... ...platform. You will design and scale models that serve as... ...experimentation and system-level evaluation. You will make... ...evaluation Identify modeling, training, and scaling risks...Training
- ...mission is to make custom AI affordable for every... ...DeepMind, xAI, Microsoft Research, etc.), where we built large-scale training infrastructure powering... ...eval benchmarks: Build evaluation frameworks that capture... ...from data curation through post‑training optimization....TrainingFull timeWork at office
- ...San Francisco, CA, USA Posted on Jul 31, 2026 About... ...real neurons to improve AI models. We study how... ...computational neuroscience, AI research and software... ...anticipate modeling and scaling risks, and partner... ...including core modeling, training, evaluation, and deployment decisions...Training
$196k - $230k
Who We AreNotion is the collaborative AI workspace where teams and agents think together. We're building one place... ...’s work.About the Role:We’re seeking an experienced UX Researcher to define and scale how we evaluate Notion’s AI-powered experiences—focusing on what “good”...Local areaShift work$204k - $300k
...Technology Group (ATG) is the research division of the company. ATG’s... ...electrical engineering, such as AI/ML, algorithms, digital signal... ...deploying AI systems at scale in production environments• Demonstrated... ..., and relevant education or training. Your recruiter can share more...TrainingFull timeLocal areaWorldwideFlexible hours- Thinking Machines Lab is seeking a Post‑Training Researcher in San Francisco to bridge raw model intelligence... ..., and metrics that drive human-aligned AI behavior using data-driven methods and... ...research, data operations, and engineering to scale human‑AI #J-18808-Ljbffr DoistTraining
- ...025, we started Handshake AI and built the fastest-growing... ...with frontier AI lab researchers to create evaluations, publish benchmarks, and push... ...the AI economy, at global scale, with impact your friends,... ...various data-intensive post-training techniques. We believe that...TrainingFull timeWork at officeRemote workFlexible hours
$170k - $216k
...destinations. We conduct research to address real-world... ...models and techniques at scale. Our mission is to... ...critical automation and evaluation frameworks that... ...experience in industrial AI applications involving... ..., experience, relevant training and education, and skill...TrainingFull timeRemote work- ...center of high-impact multimodal AI. The Chat and Multimodal Safety... ...safe, scalable models and evaluations for text, vision, and audio tasks. As a Researcher on the Chat and Multimodal Safety... ...evaluation frameworks, and advance post-training and safety testing. The role is...TrainingWork at officeRelocation package
- AI Researcher Location: San Francisco About Hum.ai is building planetary superintelligence... ...us at the cutting edge, where we’re scaling generative transformer diffusion models... ...experience implementing a wide range of pre‑training and post‑training models, including large...TrainingRemote work
- ...a critical Safety Research team at the company... ...on mitigating AI threats to global... ...security that could scale to an extreme level... ...model capabilities. Evaluate technical trade-... ...with methods for training and fine‑tuning large... ...believe this job posting is non‑compliant,...Training
$84.13 - $91.34 per hour
AI Researcher - Efficient AI (Contractor) Step into the innovative world of LG Electronics. As a global... ...multimodal models, and agentic workloads across post-training, inference, and deployment workflows. • Propose and evaluate novel compression methods (PTQ, QAT, pruning,...TrainingFull timeContract workTemporary workFor contractorsLocal areaImmediate start- ...impact multimodal work in AI. ChatGPT serves a massive... ...experiences. We develop the research, training methods, and evaluations needed to make these... ...encoders, developed multimodal post-training or evaluations, or... ...to cross-modal reasoning, scaling, and inference tradeoffs....TrainingWork at officeRelocation package
$200k - $300k
Unsiloed AI — Founding ML Researcher Type: Full-time | On-site | San Francisco, CA Compensation: $200,000-$300,000 + 0.1%-1% equity Hiring... ...: Own the full ML lifecycle — research experimentation training evaluation production deployment — with full autonomy to drive systems...TrainingFull timeH1bWork at officeVisa sponsorshipFlexible hoursWeekend work$136.45k - $231.65k
Principal UX Researcher - senior research lead on the Product... ...frameworks that scale across the product organization... ...research operations: evaluate and select research... ...practices for responsible AI usage in research... ...absence, compensation, and training. Privacy We value your...TrainingSummer holidayLocal areaFlexible hours- Drata is seeking an Applied AI Engineer in San Francisco, California to drive the effectiveness of AI systems through rigorous research and experimentation. You will optimize retrieval strategies and build evaluation frameworks, ensuring our AI delivers accurate results...Flexible hours
- Carnaby Fox is seeking a Member of Technical Staff (AI Research) in San Francisco to help shape the research direction for frontier AI models... ...with world-class researchers to design experiments, evaluate LLMs, and improve data quality for high-stakes AI benchmarks....
- ...security-first enterprise AI company, is hiring a... ...in Data Analysis and Evaluation. You will design data-... .... Collaborate with researchers and engineers to improve... ...including distributed training of LLMs. The role emphasizes... ...AI systems that scale across diverse #J-188...Training
- CLERA in San Francisco, CA is seeking a Medical AI Researcher to bridge benchmark results with real-world reliability. You will own customer engagements, define evaluation questions, and deliver evidence to support FDA submissions. The role blends ML rigor with clinical...
$238k - $302k
...states. The Large Model Evaluation team is at the nexus of Waymo’s AI ambition . With... ...quantitatively-minded engineers to research and propose new ways to... ...based on large-scale simulations. Conduct... ...location, experience, relevant training and education, and skill...TrainingFull timeRemote work$203.5k - $299.3k
...DoorDash is building an AI Research org from the ground up... ...operating at massive scale, with millions of... ...High compute budgets for training and inference, sized to... ...model pre-training and post-training, RL training runs, and large-scale evaluation sweeps Full research...TrainingHourly payWork at officeLocal areaRemote workFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Post-Training Researcher: AI Evaluation & Scaling. Be the first to apply!
- senior researcher San Francisco, CA
- machine learning researcher San Francisco, CA
- researcher San Francisco, CA
- senior design researcher San Francisco, CA
- design researcher San Francisco, CA
- qualitative researcher San Francisco, CA
- data collection researcher San Francisco, CA
- product researcher San Francisco, CA
- survey researcher San Francisco, CA
- legal researcher San Francisco, CA



