Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

ML Researcher — Frontier Reasoning & Evaluation

Crucibl

Crucibl is building judgment at scale for enterprise AI. You will push the frontier of how models reason on high-stakes business decisions, and translate findings into testable experiments that inform product direction. You’ll design evaluation frameworks for uncertainty, multi-step reasoning, and real-world failure modes, partnering with founders to shape technical vision and roadmap. Hybrid work and meaningful equity offered. #J-18808-Ljbffr Crucibl

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the ML Researcher — Frontier Reasoning & Evaluation in San Francisco, CA vacancy
  •  ...foundation model enables scalable, precise reasoning for formally verifiable code across Rust,...  ...About the role Join our team as an AI Researcher and help us push the boundaries of what'...  ...LLMs Build effective and efficient ML pipelines Collaborate with other teams... 
    Suggested
    Contract work

    Logical Intelligence

    San Francisco, CA
    16 days ago
  • $216k - $270k

    Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner for frontier AI companies, Scale plays...  ...comfortable building and instrumenting ML pipelines, writing evaluation...  ...working with and providing reasonable accommodations to applicants... 
    Suggested
    Full time

    Scale AI

    San Francisco, CA
    3 days ago
  • $196k - $230k

     ...Role:We’re seeking an experienced UX Researcher to define and scale how we evaluate Notion’s AI-powered experiences—...  ...hands-on with AI products, and can reason about how model behavior, uncertainty...  ...) and working with Data Science/ML partners on measurement strategy and... 
    Suggested
    Local area
    Shift work

    Notion Labs

    San Francisco, CA
    2 days ago
  •  ...then use that data to advance our own ML research, while also collaborating with leading AI labs to improve frontier models’ ability to reason over Axiom’s data inside Axiom’s agent...  ...representation learning, model training, evaluation, inference, deployment, and customer-... 
    Suggested

    Axiom

    San Francisco, CA
    3 days ago
  • # Researcher, Frontier Cybersecurity RisksOn-siteSan FranciscoAll jobsOn-site jobsCybersecurity Jobs##...  ...usage and new model capabilities.- Evaluate technical trade-offs within the...  ...unincorporated Los Angeles County workers: we reasonably believe that criminal history may have... 
    Suggested

    Triwill Group

    San Francisco, CA
    3 days ago
  • AI Researcher (Computer Vision/Multimodal/Generative AI...  ...the Role We are hiring ML Researchers to develop...  ...that advance the frontier of multimodal vision AI...  ...consistency multimodal reasoning and generative pipelines...  ...differentiation. Evaluate new model paradigms for... 

    SpreeAI

    San Francisco, CA
    4 days ago
  •  ...San Francisco, CA is seeking a Medical AI Researcher to bridge benchmark results with real-world...  ...You will own customer engagements, define evaluation questions, and deliver evidence to support FDA submissions. The role blends ML rigor with clinical context to assess... 

    CLERA

    San Francisco, CA
    1 day ago
  • $216.3k - $280.8k

     ....Meet the TeamAt Foundation AI, we are leading frontier AI research across Cisco. Our mission is to advance the state...  ...models, agentic AI, multimodal learning, reasoning systems, scalable training algorithms, evaluation science, inference optimization, and AI systems... 
    Full time
    Temporary work
    Local area
    Flexible hours

    CISCO Systems

    San Francisco, CA
    8 hours ago
  • $141.1k - $262.1k

     ...development. Roche’s Research and Early Development...  ...edge machine learning (ML) techniques. We are seeking...  ...to our internal reasoning Large Language Models...  ...training signals, and evaluation criteria.Evaluation &...  ...passion for applying frontier AI to drug discovery.Relocation... 
    Full time
    Work experience placement
    Local area
    Worldwide
    Relocation package

    Genentech

    South San Francisco, CA
    2 days ago
  • $167.4k - $310.8k

     ...development. Roche’s Research and Early Development...  ...edge machine learning (ML) techniques. We are seeking...  ...of our internal reasoning Large Language Models...  ...training strategies, and evaluation methodologies.Model Capability...  ...passion for applying frontier AI to drug discovery.... 
    Full time
    Local area
    Worldwide
    Relocation package

    Genentech

    South San Francisco, CA
    2 days ago
  • $197.3k - $313.7k

     ...collaborative, diverse team of researchers at Agentforce Operations. The...  ...field with an AI/ML research focusYou possess experience...  ...experience developing, deploying, and evaluating machine learning models in...  ....AccommodationsIf you need a reasonable accommodation during the... 
    Full time
    Immediate start
    Remote work

    Salesforce

    San Francisco, CA
    8 hours ago
  • $262.5k - $299.6k

     ...Applied Researcher II (AI Foundations, LLM Core and Agentic AI) Overview...  ..., our applications of AI & ML are bringing humanity and...  ...from design through training, evaluation, validation, and implementation...  ...extent required to provide needed reasonable accommodations. For technical... 
    Full time
    Part time
    Local area
    Flexible hours

    Capital One

    San Francisco, CA
    1 day ago
  • $262.5k - $299.6k

    Applied Researcher II (AI Foundations) Overview: At Capital One, we...  ...time, our applications of AI & ML are bringing humanity and...  ...from design through training, evaluation, validation, and implementation...  ...extent required to provide needed reasonable accommodations. For... 
    Full time
    Part time
    Local area
    Flexible hours

    Capital One Financial Corp

    San Francisco, CA
    4 days ago
  • Carnaby Fox is seeking a Member of Technical Staff (AI Research) in San Francisco to help shape the research direction for frontier AI models. You will collaborate with world-class researchers to design experiments, evaluate LLMs, and improve data quality for high-stakes AI... 

    Carnaby Fox

    San Francisco, CA
    5 days ago
  •  ...customers almost immediately. No speculative research track here. If you want your work to hit...  ..., and getting LLMs to understand and reason over audio directly. You'll take ideas...  ...across ASR, TTS, neural codecs, and the frontier of LLM-audio understanding & speech-to-... 
    Permanent employment
    Full time
    Immediate start

    DeepRec.ai

    San Francisco, CA
    2 days ago
  • $200k - $280k

     .... Our mandate is to push the frontier of efficient inference and RL...  ...high-performance computing for ML. Are comfortable working from...  ...training stack. Have a solid research foundation in your area(s) of...  ...scale rollout collection and evaluation cheaper. Use these pipelines... 
    Full time

    Together

    San Francisco, CA
    3 days ago
  • $100k - $150k

     ...your work firsthand. Push the frontier of neuroscience. You'll work...  ...We are a team of passionate researchers and engineers developing technology...  .... About the Role As an ML Researcher, you will help...  ...preprocessing, augmentation, training, evaluation, and deployment Develop... 
    Work at office

    AXION

    San Francisco, CA
    3 days ago
  • OpenAI is seeking a Researcher for Frontier Cybersecurity Risks to design and implement an end-to-end mitigation stack that reduces severe cyber...  ...security, policy, product, and engineering teams. You will evaluate trade-offs in coverage, latency, model utility, and user... 

    Applied Methods Ltd

    San Francisco, CA
    5 days ago
  • $200k - $300k

    Unsiloed AI — Founding ML Researcher Type: Full-time | On-site | San Francisco, CA Compensation...  ...— research experimentation training evaluation production deployment — with full autonomy...  ...shown on role page Top research labs / frontier AI — Google DeepMind, OpenAI, DeepSeek,... 
    Full time
    H1b
    Work at office
    Visa sponsorship
    Flexible hours
    Weekend work

    davidjoseph-co

    San Francisco, CA
    23 hours ago
  • $250k

     ...Transluce is a fast-moving nonprofit research lab building the public tech stack for AI evaluation and oversight. We are pioneering...  .... You don't need to be a pure ML engineer, but you should be...  ...directly with leading AI researchers, frontier AI labs, and prominent child... 
    Visa sponsorship

    Transluce

    San Francisco, CA
    23 hours ago
  • $84.13 - $91.34 per hour

    AI Researcher - Efficient AI (Contractor) Step into the innovative world...  ..., efficient inference, reasoning optimization, and next-generation...  ...workflows. • Propose and evaluate novel compression methods (PTQ...  ...or engineering experience in ML, efficient AI, model optimization... 
    Full time
    Contract work
    Temporary work
    For contractors
    Local area
    Immediate start

    LG Electronics

    San Francisco, CA
    4 days ago
  •  ...profound global impact. About the Role Frontier AI is moving toward scientific reasoning and design: molecules, materials,...  ...next era. We are looking for a Research Scientist who can help define...  ...at the intersection of frontier AI/ML, quantum algorithms, scientific machine... 
    Casual work
    Visa sponsorship

    Sygaldry

    San Francisco, CA
    5 days ago
  • $218.4k - $273k

     ...accelerating the abundance of frontier data to pave the road...  ...upon our prior model evaluation work with enterprise...  ...team, part of Scale’s Research organization, brings...  ...research in top ML venues (e.g., ACL, EMNLP...  ...Familiarity with agentic reasoning methods such as STaR... 
    Full time

    Scale AI

    San Francisco, CA
    2 days ago
  •  ...think deeply about how machines reason about the physical world;...  ...a track record of exceptional research or engineering achievement; move...  ...pipeline from data generation to evaluation to product integration....  ...Are fluent in Python and modern ML stack such as PyTorch or JAX,... 
    Relocation package

    P-1 Ai

    San Francisco, CA
    5 days ago
  •  ...Google DeepMind, xAI, Microsoft Research, etc.), where we built large-...  ...capable of robust multi‑step reasoning, tool use, and long‑horizon...  ...eval benchmarks: Build evaluation frameworks that capture real‑...  ...Hugging Face or competitive ML achievements (Kaggle medals,... 
    Full time
    Work at office

    Goaly

    San Francisco, CA
    3 days ago
  • $150k - $250k

     ...rearchitect critical operations for the frontier of AI. Our customers include the...  ..., and global social organizations.We research and deploy technologies that power AI...  ...many paradigms—retrieval pipelines, reasoning agents, evaluation harnesses, multimodal integrations, or... 
    Work at office
    3 days per week

    Distyl AI

    San Francisco, CA
    2 days ago
  •  ...questions in real time, our applications of AI & ML bring humanity and simplicity to banking....  .... Our work touches every aspect of the research life cycle, from partnering with academia...  ..., from design through training, evaluation, validation, and implementation. Engage in... 
    Flexible hours

    Capital One

    San Francisco, CA
    4 days ago
  •  ...accelerating the abundance of frontier data to pave the road...  ...upon our prior model evaluation work with enterprise...  ...of cutting‑edge AI research and practical...  ...published research in top ML venues (e.g., ACL, EMNLP...  ...Familiarity with agentic reasoning methods such as STaR... 

    Scale AI

    San Francisco, CA
    1 day ago
  •  ...LLC is seeking accomplished biology and biophysics researchers to contribute to AI systems for scientific reasoning. This position is placed at a leading AI lab as...  ...workforce in San Francisco. You will review and evaluate papers for rigor, author and review challenging problems... 
    Part time

    Mercor

    San Francisco, CA
    5 days ago
  • $234.3k - $349k

     ...AI. About the roleAI research at WRITER isn't just about...  ...models, agentic reasoning, and the system-level...  ...through model training, evaluation, and production deploymentDesign...  ...WRITER at the frontier of the field and contributing...  ...7+ years of hands-on ML research experience,... 
    Full time
    Work at office
    Local area

    Writer

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to ML Researcher — Frontier Reasoning & Evaluation. Be the first to apply!