Materials Science Researcher for AI Model Evaluation
$70 - $110 per hourSaidGig
Help advance frontier AI models by applying deep materials science expertise to the tasks, standards, and evaluations used by an AI research team. This role turns real-world materials engineering judgment into rigorous training and assessment work.
Key Responsibilities
- Review materials science knowledge tasks and model outputs for missing reasoning, unsupported structure-to-property claims, and technically plausible answers that do not withstand expert scrutiny.
- Write detailed instruction specifications and golden solutions for materials problems, defining clear standards for correct work.
- Create materials science tasks, challenging benchmarks, and evaluation sets that reflect real practice.
- Partner with researchers and adjacent-domain specialists to develop materials-specific skills and tools and translate expert judgment into consistent, teachable criteria.
Qualifications
- PhD in materials science, materials engineering, or a closely related field such as chemistry, chemical engineering, applied physics, or metallurgy. A master''s degree with exceptional industrial depth may be considered.
- At least 4 years of substantive materials research or industrial R&D experience at a research university, national laboratory, or industrial research organization. Graduate coursework alone does not qualify.
- Specialization in at least one area, such as energy storage and battery materials, semiconductors and electronic materials, polymers and soft matter, structural alloys and metallurgy, characterization and microscopy, or computational materials and simulation.
- Senior-level progression, such as Senior Scientist, Staff Scientist, Research Lead, Principal Investigator, or a senior industrial R&D role, with ownership of research direction.
- Peer-reviewed publications, granted patents, or delivered materials programs are strongly preferred.
- Hands-on professional use of large language models and the ability to distinguish sound technical reasoning from plausible but incorrect responses.
- Excellent written communication and the ability to provide precise, structured feedback.
Work Terms
- Full-time W-2 employment, working 40 hours per week for an initial 6-month engagement.
- Hybrid role based in the Bay Area, California, with on-site work alongside the client team multiple days per week when required.
- This is not a remote role. Candidates must live in the Bay Area or relocate there at their own expense before the engagement begins. Relocation assistance is not available.
- Client-issued accounts and equipment will be provided for work within the client''s tools and workflows.
- Employment, onboarding, payroll, benefits, and compliance are administered by the employer of record.
Compensation
$70 to $110 per hour.
Equal Opportunity and Accommodations
Qualified applicants are considered without regard to legally protected characteristics. Reasonable accommodations are available for qualified individuals with disabilities and disabled veterans throughout the application process.
$152k - $218.5k
...Postdoctoral Researcher At Toyota Research Institute (TRI... ...the state of the art in AI, robotics, driving, and material sciences. Our team uses design... ...we design, develop and evaluate novel interfaces to improve... ...image/video generative models. Strong programming...SuggestedWork at officeLocal areaShift work$136.44k - $265.11k
...rebuilding biotech for the AI era.When a breakthrough... ...AI are built into how science gets done.Benchling is... ...and run AI agents and models directly in their... ...ll build the datasets, evaluations, and systems that help... ...scientists, and external research partners.Desire to work...SuggestedWork at officeLocal areaMonday to FridayShift work$65 - $105 per hour
...Help advance frontier AI models by bringing rigorous life sciences research judgment into task design, evaluation, and model improvement. You will work closely with an AI lab''s research and program management teams to ensure models can reason credibly about real scientific...SuggestedHourly payFull timeFreelanceLive inRelocationRelocation package$143.33k
Explainable AI - Postdoctoral Researcher Mid-Senior Level | Full-time... ...changing impact advancing science and technology to... ...what modern AI models learn and how that knowledge... ...interfaces and evaluation methodology that let... ..., such as, material science, climate science...SuggestedFull timeWork from homeFlexible hours1 day per week$30 - $35 per hour
Model Evaluation Operator (Swing Shift) Milpitas, CA Why RoboForce RoboForce is an AI robotics company developing Physical AI-powered Robo-Labor... ...with engineers and researchers to improve model evaluation... ...Robotics, Engineering, Computer Science, or related technical...SuggestedHourly payMonday to FridayShift workAfternoon shift$100 - $150 per hour
...Overview Apply senior legal judgment to help a leading AI research organization improve how advanced AI models reason about real-world legal work. You will work... ...define high-quality legal tasks, standards, and evaluations. This role is designed for a practicing legal...Hourly payFull timeFreelanceInternshipLive inRelocationRelocation package$60 per hour
Prolific is seeking Biology Experts and Life Science Professionals to evaluate AI-generated science and ensure compliance with scientific standards. Responsibilities include reviewing biological inquiries, validating technical claims from public databases, and critiquing...Remote jobHourly payWork from homeFlexible hours- A leading AI company is seeking a legal professional for a contractor role focused on evaluating AI model outputs in legal contexts. Candidates must hold a Juris Doctor (J.D.) and have more than 3 years of experience in law. The role involves reviewing complex legal hypotheticals...For contractors10 hours per week
$65 - $105 per hour
...engineering judgment to help frontier AI models reason more accurately about... ...will work closely with an AI research and program management team... ...-quality engineering work, evaluate model performance, and turn... ...degree or higher in computer science, software engineering, or a...Hourly payFull timeFreelanceInternshipLive inRelocationRelocation package$196k - $230k
...AreNotion is the collaborative AI workspace where teams and... ...an experienced UX Researcher to define and scale how we evaluate Notion’s AI-powered experiences... ...looks like not only for model output quality, but for... ..., engineering, and data science can apply consistently....Local areaShift work$60 - $90 per hour
...Role Overview Help shape how an advanced performance-transfer model evaluates character animation, preserving an actor''s timing, emotion,... ...evaluation methods, and help build a reliable human-review process for AI-generated performance results. Key Responsibilities...Hourly payPart time$65 - $105 per hour
...Help improve how frontier AI models reason about real-world life sciences research. In this senior, hands-on role, you will apply deep scientific judgment to evaluate research tasks and model outputs, define rigorous standards for correct answers, and build benchmarks...Hourly payFull timeLive inRelocationRelocation package$70 per hour
...you’re a software engineer or researcher who’s curious and passionate... ...) Develop temporal graph modeling techniques to capture time-varying... ...PhD or Master’s) in Computer Science, Artificial Intelligence,... ...published papers in top-tier AI/ML, data mining, or computer...Hourly payFull timeInternshipSummer internshipRelocation package$400 per month
About the Role Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical assessments. The work focuses on realistic infrastructure engineering...$190k - $250k
...artificial intelligence (AI) powered technology... ...scale generative world models that learn to predict realistic... .... We are looking for a research scientist to lead the... ...radar outputsDesign evaluation frameworks that measure... ...bring:PhD in Computer Science, AI, Robotics, or a related...Temporary workWork at officeVisa sponsorship$75 - $115 per hour
...Role Overview Help advance frontier AI models by bringing rigorous pharmaceutical research and development judgment to the evaluation, design, and improvement of domain-specific... ...pharmacology, molecular biology, pharmaceutical sciences, or toxicology. An MD with substantial...Hourly payFull timeContract workLive inRelocationRelocation package$100 - $150 per hour
...Role Overview Provide senior legal subject-matter expertise to a GenAI research team, creating authoritative instruction specifications, golden solutions, and evaluation benchmarks so frontier AI models reason correctly about real legal work. This role centers on hands-on...Hourly payFull timeFreelanceInternshipLive inLocal areaRelocationRelocation package$193.93k - $352.29k
...driver, combining cutting-edge AI with automotive-grade... ...About The Team Frontier models are fungible. Any team can rent... ...system's output can be trusted — evaluation, verification, and the discipline... ...output of every engineer and researcher at Nuro by 100x. Not a better...Full time- ...Founded by a team of Stanford researchers and entrepreneurs with deep... ...world's first real-time speech AI platform capable of accent... ...team combines deep expertise in model innovation and systems... ...'s model families, build the evaluation infrastructure to measure it...
$75k - $155k
Group Research and Development (GRD) performs strategic... ...shaping the future of AI assurance. The AI Assurance... ..., risk, and assurance science. The program is... ...prototypes to test and evaluate methods relevant to embodied... ...experience in machine learning model development and digital...Temporary workWork at office$296.3k - $374.8k
Meet the Team The AI Software & Platform Group builds... ...AI platforms and models, bringing together AI with... ...portfolio. Our research team is working on a fundamental... ...for training and evaluating models and agents before... ...Qualifications Masters in Computer Science, Machine Learning,...Full timeTemporary workLocal areaFlexible hours$150k - $250k
About Distyl AI Distyl is an applied AI technology company partnering... ...social organizations.We research and deploy technologies that power... ...own construction processes, evaluate them, and evolve. They draw... ...builders).Experience Building with Models, Not Just Building Models: We...Work at office3 days per week$200k - $287.5k
...Full-time /HybridAt Toyota Research Institute (TRI), we’re... ...the state of the art in AI, robotics, driving, and material sciences.The TeamThe Automated... ...Policy and Large Behavior Models (LBM).The OpportunityWe... ...scenarios.Perform closed-loop evaluations in sensor simulations...Full timeLocal areaShift work- ...REF8691KJob Code: SES.4 Science & Engineering MTS 4 /... ...from fundamental research in machine learning, development... ...of large scale AI models, to applied AI and analysis... ...density physics, material science, predictive medicine... ..., implement, and evaluate new machine learning and...Minimum wagePermanent employmentFull timeFor contractorsLocal areaRelocationFlexible hours
- ...specified location(s).As an AI Researcher within Schwab’s AI Strategy... ...considerations matter as much as model performance.In this role,... ...through modeling, evaluation, deployment, and iteration,... ...degree or higher in Computer Science, Machine Learning, Mathematics...Full timeWork at office
$163.2k - $264k
...and Inclusion. We weave AI into the fabric of... ...when it’s needed. This model supports real-time problem... ...CareerAs a Sr Principal AI Researcher, you will advance the... ...research, rigorous evaluation, and applied LLM post-training... .../MS degree in Computer Science, Machine Learning,...Full timeWork at office$216.3k - $280.8k
Meet the TeamAt Foundation AI, we are leading frontier AI research across Cisco. Our mission is to advance the... ...AI.Our research spans foundation models, agentic AI, multimodal learning,... ...systems, scalable training algorithms, evaluation science, inference optimization, and AI...Full timeTemporary workLocal areaFlexible hours$262.5k - $299.6k
Applied Researcher II (AI Foundations, LLM Core and Agentic AI) Overview:... ...building world-class applied science and engineering teams and continue... .... Build AI foundation models through all phases of... ...from design through training, evaluation, validation, and implementation...Full timePart timeLocal areaFlexible hours$262.5k - $299.6k
Applied Researcher II (AI Foundations) Overview: At Capital One, we are creating... ...world-class applied science and engineering teams and continue... .... Build AI foundation models through all phases of... ...from design through training, evaluation, validation, and implementation...Full timePart timeLocal areaFlexible hours- ...committed to fostering its own research and development... ...smart electric vehicle models under the NIO brand, two... ...seeking exceptional AI Robotics Researchers to... ...control. Develop and evaluate large-scale multimodal... ...in Robotics, Computer Science, Artificial Intelligence...Worldwide
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Materials Science Researcher for AI Model Evaluation. Be the first to apply!
- data collection researcher California
- building materials California
- building materials sales California
- construction materials testing California
- preferred materials California
- construction materials project manager California
- material control California
- raw materials California
- construction materials manager California
- materials supervisor California



