Research Scientist: Model Evaluation & Metrics
Sanas.AI Inc.
Sanas is a leading force in real-time speech AI, advancing evaluation-driven research across accent translation, noise cancellation, and language translation. We seek a Research Scientist to define meaningful progress metrics and build rigorous evaluation infrastructure that ties research advances to real-world impact. You will shape evaluation strategies, run listener studies and MOS/MUSHRA protocols, and collaborate with product and infra teams to ensure scalable, production-ready quality #J-18808-Ljbffr Sanas.AI Inc.
- Sanas in Palo Alto is seeking a Research Scientist focused on rigorous evaluation of speech AI models. You will define meaningful metrics, build scalable evaluation pipelines, and align research progress with product impact across Accent Translation, Noise Cancellation,...Suggested
- ...by a team of Stanford researchers and entrepreneurs with... ...combines deep expertise in model innovation and systems... ...that automated metrics struggle to capture —... ...looking for a Research Scientist who can define what "better... ...families, build the evaluation infrastructure to...Suggested
$190k - $250k
...large-scale generative world models that learn to predict... ...trucks. We are looking for a research scientist to lead the design and development... ..., and radar outputsDesign evaluation frameworks that measure world... ...quality beyond pixel-level metrics, including scenario fidelity...SuggestedTemporary workWork at officeVisa sponsorship- ...industry-leading generative video models into world models: interactive,... ...into a generated world, and own the metrics that define success. It fits a researcher with deep generative-modeling or... ...tasks (planning, control, evaluation). Enthusiasm for open-sourcing frontier...Suggested
$195.2k - $262.2k
...enterprises from data and model training through to... ...Token Factory needs scientists who can turn frontier... ...inference bottlenecks into research problems, publish... ...production impact. Invent, evaluate, and productionize... ...ablations, baselines, metrics, statistical reasoning...SuggestedTemporary workImmediate startRemote work$272k - $431.25k
...generating it! Our world model team is pushing the... ...are looking for a Senior Research Manager to lead world-model evaluation and benchmarking across... ...Lead a team of Research Scientists focused on world-model evaluation... ...closed-loop benchmarks, metrics, failure taxonomy, model...Full time$185k - $400k
...infrastructure built around real-time, multimodal generation and intelligent agentic platforms. We are seeking accomplished Research Scientists in Foundation Models with expertise in pre-training and mid-training large-scale multimodal foundation models to advance our mission of...Remote work$193.93k - $352.29k
...About The Team Frontier models are fungible. Any team can rent... ...system's output can be trusted — evaluation, verification, and the... ...output of every engineer and researcher at Nuro by 100x. Not a better... ...it optimizes hard against a metric that means nothing. Post-train...Full time- About the Institute of Foundation Models We are a dedicated research lab for building, understanding, using... ...world‑class researchers, data scientists, and engineers, tackling the most fundamental... ...‑training and post‑training, and evaluation benchmarks. The role combines...
- Tencent’s Technology Engineering Group (TEG) seeks a research-focused engineer to advance large-scale video world models, including data set design, model pre-training, SFT, RL, and downstream applications. You will analyze R&D challenges, optimize training and inference...
- The Role We are looking for a Research Scientist to join the Multi-Embodiment Generalist Agent... ...founding member. MEGA is building foundation models for general-purpose robots beyond not... ...and training pipelines, designing evaluations and deploying policies on physical...Full timeWork from home
- Google is seeking a Research Data Scientist for Ads Metrics to work onsite in Mountain View, CA. You will manage a portfolio of businesses, gaining deep insight into their marketing objectives and success metrics, and translate that understanding into data-driven advertising...
$260k - $350k
...strategy, and owning the corporate operating model. They will help determine where the... ...analysis Partner with R&D leadership to evaluate roadmap trade-offs, investment priorities... ...COGS as a strategic operating metric, delivering measurable savings while improving...Full timeImmediate startFlexible hours- A leading autonomous vehicle technology company in Mountain View is seeking a World Model Research Scientist. The successful candidate will design and train generative models, requiring expertise in AI and robotics. Responsibilities include developing techniques for realistic...
$224k - $356.5k
...computing. As a Senior / Principal Deep Learning Engineer — Model Evaluation & AI Systems, you will play a meaningful role in crafting the... ...unclear technical challenges and communicate effectively across research, engineering, and product teams.Ways to stand out from the...Full time$158.3k - $297k
...innovation.What the Role Entails1. Engage in the research and development of large-scale video world models, including the design and construction of training... ...related to pre-training, SFT, and RL, model capability evaluation, and exploration of downstream application...Full timeRelocation package$192.2k - $260k
...at Amazon's Delivery Foundation Model team, where you'll work alongside world-class scientists and engineers to pioneer the... ...technical direction for specific research initiatives, ensuring robust performance... ...and our extensive training and evaluation infrastructure- Guide and...Local areaWorldwideFlexible hours$174k - $252k
Conduct AI research and development in the field of health... ...health-related metrics and outcomes.Design architectures... ...for multimodal AI models and set up tests to... ...and maintaining evaluation systems.Communicate research... ...work. As a Research Scientist, you'll setup large-scale...$147k - $210k
Author research papers to share and generate impact of research results across the team... ...structure, framework, design, and evaluation metrics for research solution development and... ...specific types of work. As a Research Scientist, you'll setup large-scale tests and deploy...$180k
.... About the Role As a multimodal engineer on the Imagine Model Team, you will develop cutting‑edge AI experiences beyond... ...studies, particularly for visual and audio data. Design evaluation frameworks, metrics, benchmarks, evals, and reward models tailored to image/video...Temporary work- NVIDIA Gruppe is seeking a Senior Research Manager to lead world-model evaluation and benchmarking efforts in Santa Clara, California. The successful candidate... ...models for Physical AI and build a team of research scientists focused on innovative evaluation techniques. The...
$175k - $350k
...future with human-centered AI models that unite emotional... ...iterate on the fun parts. Balance research curiosity with product pragmatism... ..., hyper-parameter search, evaluation, and rollout—using PyTorch, Torchtune... ...easy to trace. Define the metrics that matter; run A/B tests...$165k - $185k
Company DescriptionThe Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania... ..., our AI research in Silicon Valley focuses on Foundation Models, Big Data Visual Analytics, Explainable AI (XAI), Natural Language...Work experience placementWorldwide$174k - $252k
...experience. 2 years of experience leading a research agenda. 1 year of experience in a... ...science field. Experience with AI model training, testing, evaluation, and tuning processes, as well as... ..., framework, design, and evaluation metrics for research solution development and...- ServiceNow in Santa Clara, CA is seeking a leader for the Build Agent evaluation framework. You will own eval strategy, roadmap, telemetry, and... ...deliver scalable, data‑driven improvements. You will benchmark models, decide on model support, and communicate results to...
- Lightspeed Studios is seeking candidates for a role focused on the research and development of large-scale video world models. This position involves designing datasets, foundational model algorithms, and evaluating model capabilities. Ideal candidates should have a Bachelor'...
- ...based in Sunnyvale, CA, is seeking a Research Scientist to join the MEGA team as a founding member... ...role focuses on building foundation models for general-purpose robots beyond self... ...architectures, data pipelines, and evaluations, and deploy policies on physical robots...
$150k
...THE ROLE: You will join the Grok Voice Model team to help build the world's best voice... ...enable high-quality model training and evaluation. Work on pre-training and post-training... ...evaluation framework covering objective metrics (accuracy, quality, latency,...Temporary work$147k - $211k
...year of experience owning and initiating research agendas. About the job As an... ...specific types of work. As a Research Scientist, you'll setup large-scale tests and deploy... ...data structure, framework, design, and evaluation metrics for research solution development and...Full time$174k - $252k
Scope and drive research efforts to improve complex frontier Gemini... ....Curate and generate data to evaluate and improve Gemini capabilities... .... Experience in core ML model development.Experience with LLM... ...types of work. As a Research Scientist, you'll setup large-scale tests...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Research Scientist: Model Evaluation & Metrics. Be the first to apply!
- molecular biology scientist Palo Alto, CA
- water quality scientist Palo Alto, CA
- machine learning scientist Palo Alto, CA
- scientist antibody discovery Palo Alto, CA
- image scientist Palo Alto, CA
- machine learning research scientist Palo Alto, CA
- materials scientist Palo Alto, CA
- health scientist Palo Alto, CA
- lead scientist Palo Alto, CA
- scientist Palo Alto, CA


