Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Post-Training Researcher: AI Evaluation & Scaling

Thinking Machines

Thinking Machines is seeking post-training researchers in San Francisco to bridge theory and practice, writing high‑performance code and evaluating post‑training recipes that scale across datasets and models. You’ll blend fundamental research with engineering, iterating on evaluations, debugging training configurations, and publishing insights to move the field forward. Evergreen role with visa sponsorship and competitive compensation. #J-18808-Ljbffr Thinking Machines

Vacancy posted 1 day ago
Similar jobs that could be interesting for youBased on the Post-Training Researcher: AI Evaluation & Scaling in San Francisco, CA vacancy
  • $150k - $250k

    About Distyl AI Distyl is an applied AI technology...  ...organizations.We research and deploy technologies...  ...Key ResponsibilitiesThe Post-Training team focuses on...  ...Researchers develop and evaluate techniques such as supervised...  ..., effectively, and at scale across industriesWhat... 
    Training
    Work at office
    3 days per week

    Distyl AI

    San Francisco, CA
    1 day ago
  • $197.3k - $313.7k

     ...SalesforceSalesforce is the #1 AI CRM, where humans...  ..., diverse team of researchers at Agentforce...  ...by the global scale and trust of...  ...DoingDesign, implement, and train novel deep...  ...developing, deploying, and evaluating machine learning...  ...opt out options.Posting... 
    Training
    Full time
    Immediate start
    Remote work

    Salesforce

    San Francisco, CA
    4 days ago
  • $216.3k - $280.8k

     ...close on: 08/03/2026Job posting may be removed earlier...  ...the TeamAt Foundation AI, we are leading frontier AI research across Cisco. Our mission...  ...systems, scalable training algorithms, evaluation science, inference optimization...  ...challenging enterprise-scale problems.Your ImpactAs... 
    Training
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Francisco, CA
    4 days ago
  •  ...creating trustworthy and reliable AI systems, changing banking...  ...touches every aspect of the research life cycle, from partnering...  ..., from design through training, evaluation, validation, and implementation...  ...record of delivering models at scale both in terms of training... 
    Training
    Flexible hours

    Capital One

    San Francisco, CA
    3 days ago
  • $262.5k - $299.6k

     ...Applied Researcher II (AI Foundations, LLM Core and Agentic AI) Overview...  ...development, from design through training, evaluation, validation, and...  ...record of delivering models at scale both in terms of training data...  ...to pay at the time of this posting. Salaries for part‑time... 
    Training
    Full time
    Part time
    Local area
    Flexible hours

    Capital One

    San Francisco, CA
    18 hours ago
  • $262.5k - $299.6k

    Applied Researcher II (AI Foundations) Overview: At Capital One, we are...  ...development, from design through training, evaluation, validation, and...  ...record of delivering models at scale both in terms of training data...  ...to pay at the time of this posting. Salaries for part-time... 
    Training
    Full time
    Part time
    Local area
    Flexible hours

    Capital One Financial Corp

    San Francisco, CA
    3 days ago
  • $200k - $280k

     ...architectures, engines) and post-training / RL systems. We build...  ...can run at production scale. Our mandate is to...  ...stack. Have a solid research foundation in your area...  ...collection and evaluation cheaper. Use these pipelines...  ...engineering. About Together AI Together AI is a... 
    Training
    Full time

    Together

    San Francisco, CA
    2 days ago
  • We are seeking an Edge AI Research Scientist to develop next-generation...  ...architectures Efficient training and inference techniques Contribute...  ...frameworks. Performance Evaluation Develop rigorous benchmarking...  ...distilling, or evaluating large-scale foundation models using distributed... 
    Training

    Huxley

    San Francisco, CA
    1 day ago
  • AI Researcher (Computer Vision/Multimodal/Generative AI) About the Role...  ...architectures, algorithms, and training strategies that improve...  ..., research innovations must scale to real‑world deployment environments...  ...product differentiation. Evaluate new model paradigms for scalability... 
    Training

    SpreeAI

    San Francisco, CA
    3 days ago
  •  ...real neurons to improve AI models. We study how...  ...computational neuroscience, AI research and software...  ...platform. You will design and scale models that serve as...  ...experimentation and system-level evaluation. You will make...  ...evaluation Identify modeling, training, and scaling risks... 
    Training

    The Biological Computing Co. (TBC)

    San Francisco, CA
    2 days ago
  •  ...mission is to make custom AI affordable for every...  ...DeepMind, xAI, Microsoft Research, etc.), where we built large-scale training infrastructure powering...  ...eval benchmarks: Build evaluation frameworks that capture...  ...from data curation through post‑training optimization.... 
    Training
    Full time
    Work at office

    Goaly

    San Francisco, CA
    2 days ago
  •  ...San Francisco, CA, USA Posted on Jul 31, 2026 About...  ...real neurons to improve AI models. We study how...  ...computational neuroscience, AI research and software...  ...anticipate modeling and scaling risks, and partner...  ...including core modeling, training, evaluation, and deployment decisions... 
    Training

    Change Order

    San Francisco, CA
    4 days ago
  • $196k - $230k

    Who We AreNotion is the collaborative AI workspace where teams and agents think together. We're building one place...  ...’s work.About the Role:We’re seeking an experienced UX Researcher to define and scale how we evaluate Notion’s AI-powered experiences—focusing on what “good”... 
    Local area
    Shift work

    Notion Labs

    San Francisco, CA
    1 day ago
  • $204k - $300k

     ...Technology Group (ATG) is the research division of the company. ATG’s...  ...electrical engineering, such as AI/ML, algorithms, digital signal...  ...deploying AI systems at scale in production environments• Demonstrated...  ..., and relevant education or training. Your recruiter can share more... 
    Training
    Full time
    Local area
    Worldwide
    Flexible hours

    Dolby

    San Francisco, CA
    3 days ago
  • Thinking Machines Lab is seeking a Post‑Training Researcher in San Francisco to bridge raw model intelligence...  ..., and metrics that drive human-aligned AI behavior using data-driven methods and...  ...research, data operations, and engineering to scale human‑AI #J-18808-Ljbffr Doist
    Training

    Doist

    San Francisco, CA
    2 days ago
  •  ...025, we started Handshake AI and built the fastest-growing...  ...with frontier AI lab researchers to create evaluations, publish benchmarks, and push...  ...the AI economy, at global scale, with impact your friends,...  ...various data-intensive post-training techniques. We believe that... 
    Training
    Full time
    Work at office
    Remote work
    Flexible hours

    Handshake

    San Francisco, CA
    1 day ago
  • $170k - $216k

     ...destinations. We conduct research to address real-world...  ...models and techniques at scale. Our mission is to...  ...critical automation and evaluation frameworks that...  ...experience in industrial AI applications involving...  ..., experience, relevant training and education, and skill... 
    Training
    Full time
    Remote work

    Waymo

    San Francisco, CA
    18 hours ago
  •  ...center of high-impact multimodal AI. The Chat and Multimodal Safety...  ...safe, scalable models and evaluations for text, vision, and audio tasks. As a Researcher on the Chat and Multimodal Safety...  ...evaluation frameworks, and advance post-training and safety testing. The role is... 
    Training
    Work at office
    Relocation package

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  • AI Researcher Location: San Francisco About Hum.ai is building planetary superintelligence...  ...us at the cutting edge, where we’re scaling generative transformer diffusion models...  ...experience implementing a wide range of pre‑training and post‑training models, including large... 
    Training
    Remote work

    hum.ai

    San Francisco, CA
    3 days ago
  •  ...a critical Safety Research team at the company...  ...on mitigating AI threats to global...  ...security that could scale to an extreme level...  ...model capabilities. Evaluate technical trade-...  ...with methods for training and fine‑tuning large...  ...believe this job posting is non‑compliant,... 
    Training

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  • $84.13 - $91.34 per hour

    AI Researcher - Efficient AI (Contractor) Step into the innovative world of LG Electronics. As a global...  ...multimodal models, and agentic workloads across post-training, inference, and deployment workflows. • Propose and evaluate novel compression methods (PTQ, QAT, pruning,... 
    Training
    Full time
    Contract work
    Temporary work
    For contractors
    Local area
    Immediate start

    LG Electronics

    San Francisco, CA
    3 days ago
  •  ...impact multimodal work in AI. ChatGPT serves a massive...  ...experiences. We develop the research, training methods, and evaluations needed to make these...  ...encoders, developed multimodal post-training or evaluations, or...  ...to cross-modal reasoning, scaling, and inference tradeoffs.... 
    Training
    Work at office
    Relocation package

    United States Digital Space LLC

    San Francisco, CA
    3 days ago
  • $200k - $300k

    Unsiloed AI — Founding ML Researcher Type: Full-time | On-site | San Francisco, CA Compensation: $200,000-$300,000 + 0.1%-1% equity Hiring...  ...: Own the full ML lifecycle — research experimentation training evaluation production deployment — with full autonomy to drive systems... 
    Training
    Full time
    H1b
    Work at office
    Visa sponsorship
    Flexible hours
    Weekend work

    davidjoseph-co

    San Francisco, CA
    9 hours ago
  • $136.45k - $231.65k

    Principal UX Researcher - senior research lead on the Product...  ...frameworks that scale across the product organization...  ...research operations: evaluate and select research...  ...practices for responsible AI usage in research...  ...absence, compensation, and training. Privacy We value your... 
    Training
    Summer holiday
    Local area
    Flexible hours

    BetterUp

    San Francisco, CA
    2 days ago
  • Drata is seeking an Applied AI Engineer in San Francisco, California to drive the effectiveness of AI systems through rigorous research and experimentation. You will optimize retrieval strategies and build evaluation frameworks, ensuring our AI delivers accurate results... 
    Flexible hours

    Drata

    San Francisco, CA
    9 hours ago
  • Carnaby Fox is seeking a Member of Technical Staff (AI Research) in San Francisco to help shape the research direction for frontier AI models...  ...with world-class researchers to design experiments, evaluate LLMs, and improve data quality for high-stakes AI benchmarks.... 

    Carnaby Fox

    San Francisco, CA
    4 days ago
  •  ...security-first enterprise AI company, is hiring a...  ...in Data Analysis and Evaluation. You will design data-...  .... Collaborate with researchers and engineers to improve...  ...including distributed training of LLMs. The role emphasizes...  ...AI systems that scale across diverse #J-188... 
    Training

    Cohere

    San Francisco, CA
    4 days ago
  • CLERA in San Francisco, CA is seeking a Medical AI Researcher to bridge benchmark results with real-world reliability. You will own customer engagements, define evaluation questions, and deliver evidence to support FDA submissions. The role blends ML rigor with clinical... 

    CLERA

    San Francisco, CA
    18 hours ago
  • $238k - $302k

     ...states. The Large Model Evaluation team is at the nexus of Waymo’s AI ambition . With...  ...quantitatively-minded engineers to research and propose new ways to...  ...based on large-scale simulations. Conduct...  ...location, experience, relevant training and education, and skill... 
    Training
    Full time
    Remote work

    Waymo

    San Francisco, CA
    18 hours ago
  • $203.5k - $299.3k

     ...DoorDash is building an AI Research org from the ground up...  ...operating at massive scale, with millions of...  ...High compute budgets for training and inference, sized to...  ...model pre-training and post-training, RL training runs, and large-scale evaluation sweeps Full research... 
    Training
    Hourly pay
    Work at office
    Local area
    Remote work
    Flexible hours

    Doordashusa

    San Francisco, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Post-Training Researcher: AI Evaluation & Scaling. Be the first to apply!