Research Scientist, STEM
$150k - $300kTuring
About Turing
Turing’s mission is to accelerate superintelligence to drive real economic progress. Headquartered in San Francisco, Turing works with frontier AI labs to generate high-quality datasets, reinforcement learning environments, and frontier research benchmarks that improve model capabilities in software engineering, enterprise knowledge work, and advanced STEM reasoning. In software engineering, Turing is the largest and longest-running data provider in the category. Turing also works with Fortune 500 enterprises across financial services, life sciences, healthcare, retail, automotive, and CPG to build and deploy end-to-end agentic AI systems inside mission-critical workflows. By operating on both sides, Turing closes the loop between frontier research and enterprise deployment, turning real-world deployment signals into better data, evaluations, and more capable models. Learn more at
The Role
Turing is seeking exceptional AI Research Scientists to join our STEM research organization and develop new ways to evaluate, train, and improve frontier AI systems.
This is a research-first role focused on problems where the right benchmark, dataset, or methodology often does not yet exist. You will identify important gaps in the literature, propose ambitious new research directions, and take projects from initial hypothesis through experimentation, benchmark construction, and publication.
Our research is deliberately focused on frontier STEM evaluation, synthetic data, hallucination and reliability, and agentic science . We are looking for scientists who can recognize important problems early, formulate them precisely, and design rigorous research programs to answer them.
What You'll Do
Frontier benchmarks and evaluation
- Identify high-impact gaps in existing benchmark and evaluation literature.
- Design novel benchmarks in and across STEM fields and on general model functionality.
- Develop evaluations for emerging model capabilities that are poorly captured by traditional static benchmarks.
- Design rigorous task-generation, grading, contamination-control, difficulty-calibration, and validation methodologies.
- Build benchmarks that can become both valuable research contributions and meaningful standards for evaluating frontier models.
Synthetic data and post-training
- Develop methods for generating high-quality synthetic STEM training data.
- Study how task selection, difficulty, diversity, verification, filtering, and data quality affect downstream performance.
- Explore methods for generating useful training signal in domains where expert human data is scarce or expensive.
- Design experiments that determine when synthetic data genuinely improves capabilities rather than simply increasing training volume.
Hallucination, reliability, and verification
- Study hallucination, uncertainty, calibration, and epistemic failure in technical domains.
- Develop evaluations and methods for improving factual reliability, self-correction, verification, citation, and appropriate abstention.
- Investigate when models should reason internally, invoke tools, seek external evidence, or recognize that they do not know.
Agentic science
- Research AI systems capable of performing extended scientific and technical work.
- Develop workflows involving literature search, coding, simulation, tool use, experimentation, verification, and iterative reasoning.
- Evaluate long-horizon scientific agents and identify the bottlenecks preventing them from reliably performing real research.
- Explore new approaches to human-AI and multi-agent scientific collaboration.
New research directions
The areas above are our core focus, not an exhaustive list. Researchers will also have significant latitude to propose new programs in areas such as reasoning, model evaluation, AI-for-science, data generation, and emerging capabilities.
What We’re Looking For
- PhD or equivalent research experience in machine learning, computer science, mathematics, physics, chemistry, biology, engineering, statistics, or another highly technical field .
- Demonstrated ability to formulate and execute original research.
- Strong understanding of modern LLMs and the frontier AI research landscape.
- Excellent experimental design, quantitative reasoning, and scientific judgment.
- Ability to rapidly understand unfamiliar technical literature and develop expertise in new areas.
- Strong Python skills and the ability to independently build research prototypes and evaluation pipelines.
- Excellent technical writing and communication.
- Comfort working in a fast-moving environment where the research agenda evolves with the frontier.
A strong publication record is valuable, but we care most about whether you can identify important questions, design rigorous ways to answer them, and execute quickly enough for the results to matter .
What Success Looks Like
You might:
- Identify a major capability that existing benchmarks fail to measure and create the benchmark that becomes the standard for evaluating it.
- Discover a failure mode in current synthetic-data pipelines and develop a method that materially improves post-training.
- Build a new evaluation that changes how frontier labs understand hallucination, reasoning, or scientific capability.
- Develop an agentic workflow that substantially advances performance on complex scientific research tasks.
- Launch an entirely new research direction that grows into a major program within Turing.
Why Turing
Frontier models are improving faster than the benchmarks, datasets, and research methodologies used to understand them. The STEM research team at Turing works on that gap directly. Our goal is not simply to apply existing evaluation methods, but to invent the benchmarks, data-generation methods, and research frameworks needed for the next generation of AI systems . If you want to define how frontier AI is evaluated and improved across science and technical reasoning, we’d like to hear from you.
This role is required to be in office five days a week, based in any of Turing's offices in San Francisco, Palo Alto, or Seattle.
Compensation: $150,000 to $300,000 OTE + Equity
Values
- We are client first: We put our clients at the center of everything we do, because their success is the ultimate measure of our value.
- We work at Start-Up Speed: We move fast, stay agile and favor action because momentum is the foundation of perfection
- We are AI forward: We help our clients build the future of Al and implement it in our own roles and workflow to amplify productivity.
Advantages of joining Turing
- Work at the frontier of AI , helping the world’s leading AI labs improve their most advanced models by building expert datasets, RL environments, and first-of-a-kind benchmarks.
- Contribute to leading-edge AI research and showcase your work at top conferences such as ICLR, ICML, and NeurIPS.
- Bring frontier AI innovation to the enterprise , applying lessons learned from leading AI labs to solve real-world business challenges.
- Collaborate with and learn from exceptional colleagues with deep AI experience from Google, Meta, Amazon, and other leading technology companies.
- Move at the pace of AI innovation , with the speed, ownership, and impact of a startup.
Turing is proud to be an equal opportunity employer. We do not discriminate on the basis of race, religion, color, national origin, gender, gender identity, sexual orientation, age, marital status, disability, protected veteran status, or any other legally protected characteristics. At Turing we are dedicated to building a diverse, inclusive and authentic workplace and celebrate authenticity, so if you’re excited about this role but your past experience doesn’t align perfectly with every qualification in the job description, we encourage you to apply anyways. You may be just the right candidate for this or other roles.
For applicants from the European Union, please review Turing's GDPR notice here.
$174k - $252k
Develop and improve AI agents that generate formally verified code, algorithms, and mathematical proofs using the Lean proof assistant.Formalize the semantics of programming languages (e.g., C/C++) in Lean and build verified static analyses on top of these formalizations...Suggested$147k - $210k
..., or core generative model development in an industry AI lab, research institute, or frontier AI organization.Preferred qualifications... ...freedom to emphasize specific types of work. As a Research Scientist, you'll setup large-scale tests and deploy promising ideas quickly...Suggested- ...Job Title: Research Scientist We are seeking a dedicated Research Scientist to join our team. In this role, you will lead cutting-edge research, design benchmarks, and and agent architectures. This role is for someone eager to push the boundaries of what's possible...SuggestedFull time
$126k - $248k
...semantic search, and more. It is backed by a strong team of AI researchers from Stanford, MIT, Berkeley, Princeton, and CMU, who have... ...end solutions.Position OverviewWe are seeking a Senior Research Scientist to join our team and contribute to the development of next-generation...SuggestedLocal areaWorldwideFlexible hours- ...Research Scientist Scaled Cognition is the world's only model lab dedicated exclusively to customer experience and pioneering agentic models purpose-built for reliable action-taking enterprise applications. Backed by Khosla Ventures, the company's flagship Agentic...Suggested
$120k - $140k
...algorithms and AI models in Python Excellent communication skills and ability to collaborate within an interdisciplinary team of scientists, engineers, and physicians Preferred Qualifications ~ First-author publications at peer-reviewed journals and...Work at office- ...Research Scientist As a Research Scientist at Simular, you will: Shape the future of agentic AI by pioneering new research directions in planning, reinforcement learning, multimodal reasoning, grounding, human-agent interaction, and alignment (e.g. reward modeling...
- Conduct critical research in RL and reasoning by proposing new RL algorithms, scaling RL or inference time compute for improving Gemini... ...or journals.In this role, you will join a team of research scientists working on humanity's critical path to AGI. You will focus on...
$207k - $300k
...language models, NLP, or Generative AI.Experience in an applied research setting.Experience working with machine learning (ML) models,... ...freedom to emphasize specific types of work. As a Research Scientist, you'll setup large-scale tests and deploy promising ideas quickly...$185k - $400k
...infrastructure built around real-time, multimodal generation and intelligent agentic platforms. We are seeking accomplished Research Scientists in Foundation Models with expertise in pre-training and mid-training large-scale multimodal foundation models to advance our...Remote work$147k - $210k
...efficiencies.Build generalizable horizontal agents to accelerate researcher velocity and automate end-to-end business value generation... ...the freedom to emphasize specific types of work. As a Research Scientist, you'll setup large-scale tests and deploy promising ideas quickly...- Responsibilities Participate in cutting edge research in machine intelligence and machine learning applications. Develop solutions for real world, large scale problems.Qualifications Minimum qualifications:PhD in Computer Science, related technical field or equivalent...Full timeWork experience placement
$100k - $157k
...-talks with Leaders from across XAbout the team: At X, we are researching volumetric light-matter interactions and laser induced material... ...these transitions.About the role:As a Resident Research Scientist, you will develop new glass materials. You will be responsible...Flexible hours$160.36k - $240.54k
...You have a Ph.D. (preferable) or M.Sc. with 2-3 years of experience working with generative models in the lab, in industry, or both.Research experiences in generative models, particularly diffusion models, flow matching and energy-based models, for robotics, including...Immediate startFlexible hours$185k - $400k
...infrastructure built around real-time, multimodal generation and intelligent agentic platforms. We are looking for a staff or lead-level Research Engineer, Data to architect and scale data engineering systems supporting model training for our advanced multimodal foundation...Remote work$160.36k - $240.54k
...scalable machine learning based planning and prediction systems to generate safe and feasible trajectories for autonomous driving.Research generative sequence modeling and sequential decision making. Research backgrounds we are looking for but not limited to are:Nice to...Immediate startFlexible hours$117.2k - $313.7k
...you are the future of Salesforce.The ExperienceSalesforce AI Research is a global leader in Enterprise AI, driving innovations that... ...Salesforce AI Research.We are looking for entrepreneurial Research Scientists who want to build, ship, and scale the next generation of...Full timeWorldwide$147k - $210k
Author research papers to share and generate impact of research results across the team and in the research community. Help in growing... ...freedom to emphasize specific types of work. As a Research Scientist, you'll setup large-scale tests and deploy promising ideas quickly...$207k - $300k
...data pipelines to detect model misbehavior and misuse end-to-end.Research and develop cross-context monitoring systems to detect... ...answers.Collaborate closely with infrastructure teams and data scientists to scale your work and regularly share results with the wider...$204k - $259k
...our work, we also initiate and foster collaborations with other research teams in Alphabet. AI Foundations areas that we are currently... .... In this hybrid role, you will report to a Principal Scientist. You will: Participate in Waymo’s Foundation World...Full timeTemporary workRemote work$88.6k - $140k
Join the HJF Team! This position is contingent upon contract award. HJF is seeking a Research Scientist to develop and supervise research projects. Investigates the feasibility of applying a wide variety of scientific principles and concepts to potential inventions, products...Contract workFor contractorsWork at officeLocal area$183.83k - $275.98k
...the Role In this role, you will be a member of the Perception & Behavior team, leveraging the cutting edge of machine learning research to solve challenging real-world robotics problems. This role is focused on bringing advancements in the field of ML and large-scale...$117.2k - $313.7k
...About the Role Salesforce AI Research is seeking outstanding AI Research Scientists / Research Engineers to build and deploy high‑impact AI solutions at scale. The team focuses on AI‑driven products that serve Salesforce’s enterprise customers, and works in Palo Alto...Full time- .... We want each team member to feel welcomed and included in every step of our exciting journey. The mission of the Waymo Research team is to develop machine learning solutions addressing open problems in autonomous driving, towards the goal of safely operating...InternshipSummer internshipLocal area
- ...Stanford University, and CEOs/Presidents of Google, AMD, Broadcom, Marvell, etc. We are a team of previous Stanford professors, SAIL researchers, Olympiad medalists (IPhO, IOI, etc.), CTOs of Synopsys & GlobalFoundries, Head of Sales & CRO of Cadence, former US Secretary of...
- ...Machine Learning Research Scientist At Autoscience Institute, we create AI systems that autonomously conduct AI research. Recently, we announced the first AI agent to autonomously create peer-reviewed literature (ICLR 2025 Workshops). We are passionate about pushing...Full timeFlexible hours
- Role Mission As Research Scientist in Speech Technologies, you will lead the research and engineering that makes Hippocratic AI's conversational platform not just intelligent, but genuinely conversational—accurate, fast, and trustworthy in the highest-stakes environments...Work at office
- Google is seeking a high-velocity applied research engineer to lead exploratory ML-audio work focused on real-time perception for humans and AI. You will shape foundational algorithms and bridge theory with validated prototypes in an open-ended research setting. The role...
$193.93k - $291.15k
...enable effective closed-loop training in simulation.If you are passionate about solving challenging new problems, leading impactful research, and seeing your work deployed onto real robots, we encourage you to apply!About the WorkDesign and build scalable, machine...Immediate startFlexible hours$190k - $250k
...long-tail scenario generation, and is distilled into the onboard models that drive our autonomous trucks. We are looking for a research scientist to lead the design and development of world models capable of generating multi-sensor, multi-view, temporally coherent...Temporary workWork at officeVisa sponsorship
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist, STEM. Be the first to apply!
- manufacturing scientist Palo Alto, CA
- analytical scientist Palo Alto, CA
- senior research scientist Palo Alto, CA
- support scientist Palo Alto, CA
- application scientist Palo Alto, CA
- drug safety scientist Palo Alto, CA
- machine learning scientist Palo Alto, CA
- lab scientist Palo Alto, CA
- safety scientist Palo Alto, CA
- validation scientist Palo Alto, CA


