ML Researcher — Frontier Reasoning & Evaluation
Crucibl
Crucibl is building judgment at scale for enterprise AI. You will push the frontier of how models reason on high-stakes business decisions, and translate findings into testable experiments that inform product direction. You’ll design evaluation frameworks for uncertainty, multi-step reasoning, and real-world failure modes, partnering with founders to shape technical vision and roadmap. Hybrid work and meaningful equity offered. #J-18808-Ljbffr Crucibl
- ...foundation model enables scalable, precise reasoning for formally verifiable code across Rust,... ...About the role Join our team as an AI Researcher and help us push the boundaries of what'... ...LLMs Build effective and efficient ML pipelines Collaborate with other teams...SuggestedContract work
$216k - $270k
Scale Labs, Research Scientist — Frontier Risk EvaluationsAs the leading data and evaluation partner for frontier AI companies, Scale plays... ...comfortable building and instrumenting ML pipelines, writing evaluation... ...working with and providing reasonable accommodations to applicants...SuggestedFull time$196k - $230k
...Role:We’re seeking an experienced UX Researcher to define and scale how we evaluate Notion’s AI-powered experiences—... ...hands-on with AI products, and can reason about how model behavior, uncertainty... ...) and working with Data Science/ML partners on measurement strategy and...SuggestedLocal areaShift work- ...then use that data to advance our own ML research, while also collaborating with leading AI labs to improve frontier models’ ability to reason over Axiom’s data inside Axiom’s agent... ...representation learning, model training, evaluation, inference, deployment, and customer-...Suggested
- # Researcher, Frontier Cybersecurity RisksOn-siteSan FranciscoAll jobsOn-site jobsCybersecurity Jobs##... ...usage and new model capabilities.- Evaluate technical trade-offs within the... ...unincorporated Los Angeles County workers: we reasonably believe that criminal history may have...Suggested
- AI Researcher (Computer Vision/Multimodal/Generative AI... ...the Role We are hiring ML Researchers to develop... ...that advance the frontier of multimodal vision AI... ...consistency multimodal reasoning and generative pipelines... ...differentiation. Evaluate new model paradigms for...
- ...San Francisco, CA is seeking a Medical AI Researcher to bridge benchmark results with real-world... ...You will own customer engagements, define evaluation questions, and deliver evidence to support FDA submissions. The role blends ML rigor with clinical context to assess...
$216.3k - $280.8k
....Meet the TeamAt Foundation AI, we are leading frontier AI research across Cisco. Our mission is to advance the state... ...models, agentic AI, multimodal learning, reasoning systems, scalable training algorithms, evaluation science, inference optimization, and AI systems...Full timeTemporary workLocal areaFlexible hours$141.1k - $262.1k
...development. Roche’s Research and Early Development... ...edge machine learning (ML) techniques. We are seeking... ...to our internal reasoning Large Language Models... ...training signals, and evaluation criteria.Evaluation &... ...passion for applying frontier AI to drug discovery.Relocation...Full timeWork experience placementLocal areaWorldwideRelocation package$167.4k - $310.8k
...development. Roche’s Research and Early Development... ...edge machine learning (ML) techniques. We are seeking... ...of our internal reasoning Large Language Models... ...training strategies, and evaluation methodologies.Model Capability... ...passion for applying frontier AI to drug discovery....Full timeLocal areaWorldwideRelocation package$197.3k - $313.7k
...collaborative, diverse team of researchers at Agentforce Operations. The... ...field with an AI/ML research focusYou possess experience... ...experience developing, deploying, and evaluating machine learning models in... ....AccommodationsIf you need a reasonable accommodation during the...Full timeImmediate startRemote work$262.5k - $299.6k
...Applied Researcher II (AI Foundations, LLM Core and Agentic AI) Overview... ..., our applications of AI & ML are bringing humanity and... ...from design through training, evaluation, validation, and implementation... ...extent required to provide needed reasonable accommodations. For technical...Full timePart timeLocal areaFlexible hours$262.5k - $299.6k
Applied Researcher II (AI Foundations) Overview: At Capital One, we... ...time, our applications of AI & ML are bringing humanity and... ...from design through training, evaluation, validation, and implementation... ...extent required to provide needed reasonable accommodations. For...Full timePart timeLocal areaFlexible hours- Carnaby Fox is seeking a Member of Technical Staff (AI Research) in San Francisco to help shape the research direction for frontier AI models. You will collaborate with world-class researchers to design experiments, evaluate LLMs, and improve data quality for high-stakes AI...
- ...customers almost immediately. No speculative research track here. If you want your work to hit... ..., and getting LLMs to understand and reason over audio directly. You'll take ideas... ...across ASR, TTS, neural codecs, and the frontier of LLM-audio understanding & speech-to-...Permanent employmentFull timeImmediate start
$200k - $280k
.... Our mandate is to push the frontier of efficient inference and RL... ...high-performance computing for ML. Are comfortable working from... ...training stack. Have a solid research foundation in your area(s) of... ...scale rollout collection and evaluation cheaper. Use these pipelines...Full time$100k - $150k
...your work firsthand. Push the frontier of neuroscience. You'll work... ...We are a team of passionate researchers and engineers developing technology... .... About the Role As an ML Researcher, you will help... ...preprocessing, augmentation, training, evaluation, and deployment Develop...Work at office- OpenAI is seeking a Researcher for Frontier Cybersecurity Risks to design and implement an end-to-end mitigation stack that reduces severe cyber... ...security, policy, product, and engineering teams. You will evaluate trade-offs in coverage, latency, model utility, and user...
$200k - $300k
Unsiloed AI — Founding ML Researcher Type: Full-time | On-site | San Francisco, CA Compensation... ...— research experimentation training evaluation production deployment — with full autonomy... ...shown on role page Top research labs / frontier AI — Google DeepMind, OpenAI, DeepSeek,...Full timeH1bWork at officeVisa sponsorshipFlexible hoursWeekend work$250k
...Transluce is a fast-moving nonprofit research lab building the public tech stack for AI evaluation and oversight. We are pioneering... .... You don't need to be a pure ML engineer, but you should be... ...directly with leading AI researchers, frontier AI labs, and prominent child...Visa sponsorship$84.13 - $91.34 per hour
AI Researcher - Efficient AI (Contractor) Step into the innovative world... ..., efficient inference, reasoning optimization, and next-generation... ...workflows. • Propose and evaluate novel compression methods (PTQ... ...or engineering experience in ML, efficient AI, model optimization...Full timeContract workTemporary workFor contractorsLocal areaImmediate start- ...profound global impact. About the Role Frontier AI is moving toward scientific reasoning and design: molecules, materials,... ...next era. We are looking for a Research Scientist who can help define... ...at the intersection of frontier AI/ML, quantum algorithms, scientific machine...Casual workVisa sponsorship
$218.4k - $273k
...accelerating the abundance of frontier data to pave the road... ...upon our prior model evaluation work with enterprise... ...team, part of Scale’s Research organization, brings... ...research in top ML venues (e.g., ACL, EMNLP... ...Familiarity with agentic reasoning methods such as STaR...Full time- ...think deeply about how machines reason about the physical world;... ...a track record of exceptional research or engineering achievement; move... ...pipeline from data generation to evaluation to product integration.... ...Are fluent in Python and modern ML stack such as PyTorch or JAX,...Relocation package
- ...Google DeepMind, xAI, Microsoft Research, etc.), where we built large-... ...capable of robust multi‑step reasoning, tool use, and long‑horizon... ...eval benchmarks: Build evaluation frameworks that capture real‑... ...Hugging Face or competitive ML achievements (Kaggle medals,...Full timeWork at office
$150k - $250k
...rearchitect critical operations for the frontier of AI. Our customers include the... ..., and global social organizations.We research and deploy technologies that power AI... ...many paradigms—retrieval pipelines, reasoning agents, evaluation harnesses, multimodal integrations, or...Work at office3 days per week- ...questions in real time, our applications of AI & ML bring humanity and simplicity to banking.... .... Our work touches every aspect of the research life cycle, from partnering with academia... ..., from design through training, evaluation, validation, and implementation. Engage in...Flexible hours
- ...accelerating the abundance of frontier data to pave the road... ...upon our prior model evaluation work with enterprise... ...of cutting‑edge AI research and practical... ...published research in top ML venues (e.g., ACL, EMNLP... ...Familiarity with agentic reasoning methods such as STaR...
- ...LLC is seeking accomplished biology and biophysics researchers to contribute to AI systems for scientific reasoning. This position is placed at a leading AI lab as... ...workforce in San Francisco. You will review and evaluate papers for rigor, author and review challenging problems...Part time
$234.3k - $349k
...AI. About the roleAI research at WRITER isn't just about... ...models, agentic reasoning, and the system-level... ...through model training, evaluation, and production deploymentDesign... ...WRITER at the frontier of the field and contributing... ...7+ years of hands-on ML research experience,...Full timeWork at officeLocal area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to ML Researcher — Frontier Reasoning & Evaluation. Be the first to apply!
- senior researcher San Francisco, CA
- machine learning researcher San Francisco, CA
- researcher San Francisco, CA
- senior design researcher San Francisco, CA
- design researcher San Francisco, CA
- qualitative researcher San Francisco, CA
- data collection researcher San Francisco, CA
- product researcher San Francisco, CA
- survey researcher San Francisco, CA
- legal researcher San Francisco, CA

