Machine Learning Scientist
Arena
About Arena Intelligence Arena Intelligence is the open platform for evaluating how AI models perform in the real world. Created by researchers from UC Berkeley’s SkyLab, our mission is to measure and advance the frontier of AI for real-world use. Millions of people use Arena Intelligence each month to explore how frontier systems perform — and we use our community’s feedback to build transparent, rigorous, and human-centered model evaluations. Leading enterprises and AI labs rely on our evaluations to understand real-world reliability, alignment, and impact. Our leaderboards are the gold standard for AI performance — trusted by leaders across the AI community and shaping the global conversation on model reliability and progress. We’re a team of researchers, engineers, academics, and builders from places like UC Berkeley, Google, Stanford, DeepMind, and Discord. We seek truth, move fast, and value craftsmanship, curiosity, and impact over hierarchy. We’re building a company where thoughtful, curious people from all backgrounds can do their best work. Everyone on our team is a deep expert in their field — our office radiates excellence, energy, and focus. About the Role Arena Intelligence is seeking a variety of Machine Learning Scientist to help advance how we evaluate and understand AI models. You’ll help design and analyse experiments that uncover what makes models useful, trustworthy and capable through human preference signals. Your work will contribute to the scientific foundations of understanding AI at scale. This role is deeply interdisciplinary. You’ll work closely with engineers, product teams, marketing and the broader research community to develop new methods for comparing models, analyzing preference data, and disentangling performance factors like style, reasoning, and robustness. Your work will inform both the public leaderboard and the tools we provide to model developers. If you’re excited by open-ended questions, rigorous evaluation, and research that’s grounded in real-world impact, you’ll find a meaningful home here. We’re looking for: Hands-on experience training large-scale models, including reward models, preference models, and fine-tuning LLMs with methods like RLHF, DPO, and contrastive learning. Strong foundation in ML and statistics, with a track record of designing novel training objectives, evaluation schemes, or statistical frameworks to improve model reliability and alignment. Fluent in the full experimental stack, from dataset design and large-batch training to rigorous evaluation and ablation, with an eye for what scales to production. Deeply collaborative mindset, working closely with engineers to productionize research insights and iterating with product teams to align modeling goals with user needs. You’ll Design and conduct experiments to evaluate AI model behavior across reasoning, style, robustness, and user preference dimensions Develop new metrics, methodologies, and evaluation protocols that go beyond traditional benchmarks Analyze large-scale human voting and interaction data to uncover insights into model performance and user preferences Collaborate with engineers to implement and scale research findings into production systems Prototype and test research ideas rapidly, balancing rigor with iteration speed Author internal reports and external publications that contribute to the broader ML research community Partner with model providers to shape evaluation questions and support responsible model testing Contribute to the scientific integrity and transparency of the Arena Intelligence leaderboard and tools You’ll have PhD or equivalent research experience in Machine Learning, Natural Language Processing, Statistics, or a related field Strong understanding of LLMs and modern deep learning architectures (e.g., Transformers, diffusion models, reinforcement learning with human feedback) Proficiency in Python and ML research libraries such as PyTorch, JAX, or TensorFlow Demonstrated ability to design and analyze experiments with statistical rigor Experience publishing research or working on open-source projects in ML, NLP, or AI evaluation Comfortable working with real-world usage data and designing metrics beyond standard benchmarks Ability to translate research questions into practical systems and collaborate across engineering and product teams Passion for open science, reproducibility, and community-driven research. What we offer We offer competitive compensation and equity aligned to the markets where our team members are based. The base salary range will depend on the candidate’s permanent work location. Comprehensive health and wellness benefits, including medical, dental, vision, and additional support programs. The opportunity to work on cutting-edge AI with a small, mission-driven team A culture that values transparency, trust, and community impact Come help build the space where anyone can explore and help shape the future of AI. Arena Intelligence provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion, sex, national origin, age, disability, genetics, sexual orientation, gender identity, or gender expression. We are committed to a diverse and inclusive workforce and welcome people from all backgrounds, experiences, perspectives, and abilities. #J-18808-Ljbffr Arena
$160k - $280k
...deploy our state of the art ML models trained with an H100/scientist ratio of 100x. Check out our Suno version of the job here! What... ...of data engineering, designing, training and evaluating machine learning models Track record showing independent ownership of entire...SuggestedWork at officeFlexible hours$200k - $275k
...Job Description: Senior Machine Learning Scientist sf, ca Senior Machine Learning Scientist, you will play a leading role in designing the next generation of foundation models of gene regulatory networks powered by Tahoe's large scale single-cell datasets...SuggestedWork at officeVisa sponsorshipFree visa3 days per week$160k - $280k
...deploy our state of the art ML models trained with an H100/scientist ratio of 100x. What You'll Need ~5+ years experience... ...stack of data engineering, designing, training and evaluating machine learning models ~ Track record showing independent ownership of...SuggestedFull timeWork at officeLocal area- ...their field — our office radiates excellence, energy, and focus. About the Role Arena Intelligence is seeking a variety of Machine Learning Scientist to help advance how we evaluate and understand AI models. You’ll help design and analyse experiments that uncover what...SuggestedPermanent employmentWork at office
- ...health systems. We are a growing team of practicing MDs, AI scientists, PhDs, creatives, technologists, and engineers working together... ...to delivering key takeaways, our trailblazing work in machine learning research makes the Abridge experience possible. We're currently...SuggestedCurrently hiringWork at officeRelocation packageFlexible hours
$150k - $300k
...time, causality, and context. As a Research Scientist, you will tackle fundamental problems in... ...micro-events into durable knowledge, and learn patterns that predict events before it... ...~5+ years building novel systems in machine learning, NLP, knowledge graphs, or related...Relocation packageFlexible hours$180k - $270k
...too much just yet, our team is tackling cutting-edge engineering challenges to bring revolutionary products to life. As a Machine Learning Scientist, you will develop cutting-edge AI models to integrate and decode complex, multimodal data streams from our custom sensing...Full timeVisa sponsorship$142.8k - $193.2k
...Applied Scientist, Prime Video - Title Lifecycle Presentation Job ID: 10372570 | Amazon.com... ..., and real‑time signals Reinforcement learning frameworks that create continuous improvement... ...techniques with strong fundamentals in machine learning and statistical methods...Flexible hours- ...energetic, curious, and driven Postdoctoral Scientist interested in systems neuroscience to... ...during vocal behaviors. Memory and learning during offline periods (sleep). Long‑range... ...neuronal communication and stroke. Brain‑machine interfaces to enhance learning and performance...
$225k - $300k
...Research Scientist About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access,... ...record contains extraordinary depth. ML at Latent Health The Machine Learning team is responsible for building systems that run in real clinical...Work at officeImmediate start$229k - $269k
...commitment to patients with cancers harboring mutations in the RAS signaling pathway. The Opportunity We are seeking a Senior Machine Learning Scientist to help accelerate drug discovery through advanced analytics and artificial intelligence. This role will develop...Full timeLocal area$160k - $250k
...Overview Research Scientist - Mountain View, CA at Granica. This range is provided by Granica... ...experience — talk with your recruiter to learn more. Base pay range $160,000.00/yr -... ...learning. What you’ll bring PhD in Machine Learning, Statistics, Applied Mathematics...Flexible hours- ...by training RNA foundation models that learn the patterns that shape disease progression... ...unique. We’re a technical team of AI scientists and engineers from companies including Recursion... ...haves PhD (or equivalent experience) in Machine Learning, Computational Biology, or...
$176k - $304k
...Cambridge, MA USA; San Francisco, CA USA As a Machine Learning Research Scientist I/II in LLM Inference you will lead research on how we train and serve large language models for scientific applications. What You’ll Be Building Develop and optimize LLM post-training strategies...- About the Role This role operates at the forefront of AI research and real-world implementation, with a strong focus on reasoning within large language models (LLMs). The ideal candidate will study the data types critical for advancing LLM-based agents, including browser...
- ...team is made up of mathematicians, physicists, and computer scientists who are deeply passionate about their craft. If you thrive on... ..., this is the place for you. The Role We’re looking for a Machine Learning Scientist to push the limits of small, high-performance language...Work at office
$117.2k - $313.7k
...is looking for outstanding AI Research Scientists and Research Engineers to discover new research... ...: LLM‑powered agents, reinforcement learning (RL), reasoning and planning, autonomous... .... Core Modeling and Post‑Training: Machine learning methodology, pre‑training and post...$114.2k - $306.6k
...customers by harnessing the latest deep learning techniques. As part of our team you’ll learn... ...is looking for outstanding AI Research Scientists / Research Engineers.**Our team... ...latency deployment.* **Core Modeling:** Machine learning methodology, pre-training/post-...Full time$153k - $235k
...highly motivated individuals to be foundational members of our Machine Learning and Data Platform team. You will partner across the company... ...impact goes a long way here. As our next Machine Learning Scientist you should have 5+ years of experience, plus: Bachelor’s degree...Work experience placementRemote workWork from homeHome officeFlexible hours$229k - $269k
...commitment to patients with cancers harboring mutations in the RAS signaling pathway. The Opportunity We are seeking a Senior Machine Learning Scientist to help accelerate drug discovery through advanced analytics and artificial intelligence. This role will develop...Full timeLocal area- Use machine-learning, applied computer science, and techniques from high performance computing to develop and refine compilers and frameworks for reducing engineering complexity and time to market of mobile applications. An ideal candidate will have practical experience...Remote work
$176k - $304k
ML Scientist I / II, Foundation Models for Life Sciences San Francisco, CA USA Lila is building a platform where AI and automation... ...to foundation model research at the intersection of machine learning and life science data. You will work on generative models spanning...Full timeWork at officeLocal areaFlexible hours$175k - $215k
...of billions in simulation across 15+ U.S. states. The Predictive Planning team (PrePlan) develops and deploys state-of-the-art machine learning solutions that predict the future state of the world and plan the Waymo Driver's behavior. Our mission is to transform Waymo's...Full timeInternshipRemote work$214.5k - $244.8k
...For years, Capital One has been leading the industry in using machine learning to create real‑time, intelligent, automated customer... .... In this role Partner with a cross‑functional team of data scientists, software engineers, machine learning engineers and product...Full timePart timeLocal areaFlexible hours$144k - $187k
...Overview MSCI is establishing a Machine Learning Center of Excellence within the Research & Development team to develop machine learning models that power investment tools for institutional clients. We are seeking exceptional early‑career researchers with recent PhDs...Flexible hours$150k - $300k
...AI research lab and infrastructure provider working on LLM interpretability and context optimization. The team builds custom machine learning models that analyze and compress token contexts before they reach the underlying model, cutting inference costs by roughly 50%...Full timeH1bVisa sponsorship$140k - $250k
...thoughts from the brain, entirely non-invasively. We apply deep learning research to large scale EEG datasets collected on affordable... ...society as a whole. About the Role We're seeking a talented Machine Learning Researcher to join our core R&D team. This role...Work from homeVisa sponsorship- ...environments. Our robotic arms run on tight compute and power budgets, so learning systems have to be fast, reliable, and deeply integrated with... ...that let robots see, reason, and act. This work runs on real machines, not benchmarks. About the role As an AI Researcher at...
$250k - $350k
...training systems, and modalities to create novel products for our customers. Your work will span from exploring new architectures and learning methods to optimizing latency and efficiency, with the goal of delivering better models to customers. Your north star is...Work at office- ...pharma\'s drug discovery workflows, where scientists use them to solve the highest value drug... ..., feature extraction, representation learning, model training, evaluation, inference,... ...or independently, that shows exceptional machine learning ability. You are deeply technical...
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Machine Learning Scientist. Be the first to apply!
- scientist 1 San Francisco, CA
- application scientist San Francisco, CA
- pharmaceutical scientist San Francisco, CA
- scientist biology San Francisco, CA
- research scientist machine learning deep learning San Francisco, CA
- scientist immunology San Francisco, CA
- senior analytical scientist San Francisco, CA
- materials scientist San Francisco, CA
- image scientist San Francisco, CA
- senior principal scientist San Francisco, CA

