Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Machine Learning Scientist

Doist

About Arena Intelligence Arena is the platform for evaluating how AI models perform in the real world. Founded by researchers from UC Berkeley's SkyLab, we are on a mission to measure and advance the frontier of AI for real‑world use, and to build the foundation for everyone to understand, shape, and benefit from it. Tens of millions of people use Arena each month to evaluate how frontier systems handle the work they actually do. The preferences they share power the most transparent, rigorous, and human‑centered evaluations in AI. Leading AI labs, enterprises, and independent researchers rely on our work and open datasets to understand how models behave in real workflows: agentic coding, creative generation, professional productivity, and beyond. We go beyond leaderboards and decompose what human experience reveals about AI, so models advance toward the work people actually do. We’re a team of researchers, academics, builders, and creatives from UC Berkeley, Google, Stanford, and DeepMind. We seek truth, move fast, and value craftsmanship, curiosity, and impact over hierarchy. We’re building a company where thoughtful, curious people from all backgrounds can do their best work together, in an office culture that radiates excellence, energy, and focus. About the Role Arena Intelligence is seeking a variety of Machine Learning Scientists to help advance how we evaluate and understand AI models. You’ll help design and analyse experiments that uncover what makes models useful, trustworthy, and capable through human preference signals. Your work will contribute to the scientific foundations of understanding AI at scale. This role is deeply interdisciplinary. You’ll work closely with engineers, product teams, marketing, and the broader research community to develop new methods for comparing models, analysing preference data, and disentangling performance factors like style, reasoning, and robustness. Your work will inform both the public leaderboard and the tools we provide to model developers. If you’re excited by open‑ended questions, rigorous evaluation, and research that’s grounded in real‑world impact, you’ll find a meaningful home here. We’re looking for: Hands‑on experience training large‑scale models, including reward models, preference models, and fine‑tuning LLMs with methods like RLHF, DPO, and contrastive learning. Strong foundation in ML and statistics, with a track record of designing novel training objectives, evaluation schemes, or statistical frameworks to improve model reliability and alignment. Fluent in the full experimental stack, from dataset design and large‑batch training to rigorous evaluation and ablation, with an eye for what scales to production. Deeply collaborative mindset, working closely with engineers to productionise research insights and iterating with product teams to align modelling goals with user needs. You’ll Design and conduct experiments to evaluate AI model behaviour across reasoning, style, robustness, and user preference dimensions. Develop new metrics, methodologies, and evaluation protocols that go beyond traditional benchmarks. Analyze large‑scale human voting and interaction data to uncover insights into model performance and user preferences. Collaborate with engineers to implement and scale research findings into production systems. Prototype and test research ideas rapidly, balancing rigour with iteration speed. Author internal reports and external publications that contribute to the broader ML research community. Partner with model providers to shape evaluation questions and support responsible model testing. Contribute to the scientific integrity and transparency of the Arena Intelligence leaderboard and tools. You’ll have PhD or equivalent research experience in Machine Learning, Natural Language Processing, Statistics, or a related field. Strong understanding of LLMs and modern deep learning architectures (e.g., Transformers, diffusion models, reinforcement learning with human feedback). Proficiency in Python and ML research libraries such as PyTorch, JAX, or TensorFlow. Demonstrated ability to design and analyse experiments with statistical rigour. Experience publishing research or working on open‑source projects in ML, NLP, or AI evaluation. Comfortable working with real‑world usage data and designing metrics beyond standard benchmarks. Ability to translate research questions into practical systems and collaborate across engineering and product teams. Passion for open science, reproducibility, and community‑driven research. What we offer Competitive compensation and equity aligned to the markets where our team members are based. The base salary range will depend on the candidate’s permanent work location. Comprehensive health and wellness benefits, including medical, dental, vision, and additional support programs. Opportunity to work on cutting‑edge AI with a small, mission‑driven team. A culture that values transparency, trust, and community impact. Arena Intelligence provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion, sex, national origin, age, disability, genetics, sexual orientation, gender identity, or gender expression. We are committed to a diverse and inclusive workforce and welcome people from all backgrounds, experiences, perspectives, and abilities. #J-18808-Ljbffr Doist

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Machine Learning Scientist in San Francisco, CA vacancy
  • $200k - $275k

     ...Job Description: Senior Machine Learning Scientist sf, ca Senior Machine Learning Scientist, you will play a leading role in designing the next generation of foundation models of gene regulatory networks powered by Tahoe's large scale single-cell datasets... 
    Suggested
    Work at office
    Visa sponsorship
    Free visa
    3 days per week

    ESR Healthcare

    San Francisco, CA
    1 day ago
  • $160k - $280k

     ...deploy our state of the art ML models trained with an H100/scientist ratio of 100x. Check out our Suno version of the job here! What...  ...of data engineering, designing, training and evaluating machine learning models Track record showing independent ownership of entire... 
    Suggested
    Work at office
    Flexible hours

    Menlo Ventures

    San Francisco, CA
    5 days ago
  • $160k - $280k

     ...deploy our state of the art ML models trained with an H100/scientist ratio of 100x. What You'll Need ~5+ years experience...  ...stack of data engineering, designing, training and evaluating machine learning models ~ Track record showing independent ownership of... 
    Suggested
    Full time
    Work at office
    Local area

    SUNO

    San Francisco, CA
    2 days ago
  • $142.8k - $193.2k

     ...Applied Scientist, Prime Video - Title Lifecycle Presentation Job ID: 10372570 | Amazon.com...  ..., and real‑time signals Reinforcement learning frameworks that create continuous improvement...  ...techniques with strong fundamentals in machine learning and statistical methods... 
    Suggested
    Flexible hours

    Amazon

    San Francisco, CA
    1 day ago
  •  ...health systems. We are a growing team of practicing MDs, AI scientists, PhDs, creatives, technologists, and engineers working together...  ...to delivering key takeaways, our trailblazing work in machine learning research makes the Abridge experience possible. We're currently... 
    Suggested
    Currently hiring
    Work at office
    Relocation package
    Flexible hours

    Gravity Engineering Services Pvt Ltd.

    San Francisco, CA
    3 days ago
  •  ...their field — our office radiates excellence, energy, and focus. About the Role Arena Intelligence is seeking a variety of Machine Learning Scientist to help advance how we evaluate and understand AI models. You’ll help design and analyse experiments that uncover what... 
    Permanent employment
    Work at office

    Arena

    San Francisco, CA
    3 days ago
  • $150k - $300k

     ...time, causality, and context. As a Research Scientist, you will tackle fundamental problems in...  ...micro-events into durable knowledge, and learn patterns that predict events before it...  ...~5+ years building novel systems in machine learning, NLP, knowledge graphs, or related... 
    Relocation package
    Flexible hours

    Dynamis Labs

    San Francisco, CA
    2 days ago
  •  ...partnering with a AI‑native therapeutics company applying frontier machine learning to understand human biology and improve clinical outcomes in...  .... Position Summary We are seeking a Machine Learning Scientist to conduct original, high‑impact machine learning research and... 

    Harnham

    San Francisco, CA
    3 days ago
  •  ...energetic, curious, and driven Postdoctoral Scientist interested in systems neuroscience to...  ...during vocal behaviors. Memory and learning during offline periods (sleep). Long‑range...  ...neuronal communication and stroke. Brain‑machine interfaces to enhance learning and performance... 

    Howard Hughes Medical Institute

    San Francisco, CA
    5 days ago
  • $225k - $300k

     ...Research Scientist About Latent Health Healthcare today is only truly personalized for two groups: those with wealth and access,...  ...record contains extraordinary depth. ML at Latent Health The Machine Learning team is responsible for building systems that run in real clinical... 
    Work at office
    Immediate start

    Latent

    San Francisco, CA
    5 days ago
  • $160k - $250k

     ...Overview Research Scientist - Mountain View, CA at Granica. This range is provided by Granica...  ...experience — talk with your recruiter to learn more. Base pay range $160,000.00/yr -...  ...learning. What you’ll bring PhD in Machine Learning, Statistics, Applied Mathematics... 
    Flexible hours

    Granica

    San Francisco, CA
    5 days ago
  •  ...team is made up of mathematicians, physicists, and computer scientists who are deeply passionate about their craft. If you thrive on...  ..., this is the place for you. The Role We’re looking for a Machine Learning Scientist to push the limits of small, high-performance language... 
    Work at office

    Relace Inc

    San Francisco, CA
    3 days ago
  • $117.2k - $313.7k

     ...is looking for outstanding AI Research Scientists and Research Engineers to discover new research...  ...: LLM‑powered agents, reinforcement learning (RL), reasoning and planning, autonomous...  .... Core Modeling and Post‑Training: Machine learning methodology, pre‑training and post... 

    Salesforce.Com Inc

    San Francisco, CA
    3 days ago
  • $229k - $269k

     ...commitment to patients with cancers harboring mutations in the RAS signaling pathway. The Opportunity We are seeking a Senior Machine Learning Scientist to help accelerate drug discovery through advanced analytics and artificial intelligence. This role will develop... 
    Full time
    Local area

    REVOLUTION Medicines

    San Francisco, CA
    3 days ago
  • $153k - $235k

     ...highly motivated individuals to be foundational members of our Machine Learning and Data Platform team. You will partner across the company...  ...impact goes a long way here. As our next Machine Learning Scientist you should have 5+ years of experience, plus: Bachelor’s degree... 
    Work experience placement
    Remote work
    Work from home
    Home office
    Flexible hours

    Whatnot

    San Francisco, CA
    3 days ago
  • $114.2k - $306.6k

     ...customers by harnessing the latest deep learning techniques. As part of our team you’ll learn...  ...is looking for outstanding AI Research Scientists / Research Engineers.**Our team...  ...latency deployment.* **Core Modeling:** Machine learning methodology, pre-training/post-... 
    Full time

    Niebles

    San Francisco, CA
    1 day ago
  • About the Role This role operates at the forefront of AI research and real-world implementation, with a strong focus on reasoning within large language models (LLMs). The ideal candidate will study the data types critical for advancing LLM-based agents, including browser...

    Gravity Engineering Services Pvt Ltd.

    San Francisco, CA
    3 days ago
  •  ...by training RNA foundation models that learn the patterns that shape disease progression...  ...unique. We’re a technical team of AI scientists and engineers from companies including Recursion...  ...haves PhD (or equivalent experience) in Machine Learning, Computational Biology, or... 

    blank

    San Francisco, CA
    1 day ago
  • $176k - $304k

     ...Cambridge, MA USA; San Francisco, CA USA As a Machine Learning Research Scientist I/II in LLM Inference you will lead research on how we train and serve large language models for scientific applications. What You’ll Be Building Develop and optimize LLM post-training strategies... 

    Lila Sciences

    San Francisco, CA
    1 day ago
  • $229k - $269k

     ...commitment to patients with cancers harboring mutations in the RAS signaling pathway. The Opportunity We are seeking a Senior Machine Learning Scientist to help accelerate drug discovery through advanced analytics and artificial intelligence. This role will develop... 
    Full time
    Local area

    REVOLUTION Medicines

    San Francisco, CA
    4 days ago
  • $176k - $304k

    ML Scientist I / II, Foundation Models for Life Sciences San Francisco, CA USA Lila is building a platform where AI and automation...  ...to foundation model research at the intersection of machine learning and life science data. You will work on generative models spanning... 
    Full time
    Work at office
    Local area
    Flexible hours

    Lila Sciences

    San Francisco, CA
    4 days ago
  • Use machine-learning, applied computer science, and techniques from high performance computing to develop and refine compilers and frameworks for reducing engineering complexity and time to market of mobile applications. An ideal candidate will have practical experience... 
    Remote work

    Peoples Grocers LLC.

    San Francisco, CA
    5 days ago
  • $175k - $215k

     ...of billions in simulation across 15+ U.S. states. The Predictive Planning team (PrePlan) develops and deploys state-of-the-art machine learning solutions that predict the future state of the world and plan the Waymo Driver's behavior. Our mission is to transform Waymo's... 
    Full time
    Internship
    Remote work

    Waymo

    San Francisco, CA
    2 days ago
  • $214.5k - $244.8k

     ...For years, Capital One has been leading the industry in using machine learning to create real‑time, intelligent, automated customer...  .... In this role Partner with a cross‑functional team of data scientists, software engineers, machine learning engineers and product... 
    Full time
    Part time
    Local area
    Flexible hours

    Capital One

    San Francisco, CA
    5 days ago
  • $140k - $250k

     ...thoughts from the brain, entirely non-invasively. We apply deep learning research to large scale EEG datasets collected on affordable...  ...society as a whole. About the Role We're seeking a talented Machine Learning Researcher to join our core R&D team. This role... 
    Work from home
    Visa sponsorship

    Alljoined

    San Francisco, CA
    2 days ago
  • $250k - $350k

     ...training systems, and modalities to create novel products for our customers. Your work will span from exploring new architectures and learning methods to optimizing latency and efficiency, with the goal of delivering better models to customers. Your north star is... 
    Work at office

    Inference

    San Francisco, CA
    3 days ago
  • $144k - $187k

     ...Overview MSCI is establishing a Machine Learning Center of Excellence within the Research & Development team to develop machine learning models that power investment tools for institutional clients. We are seeking exceptional early‑career researchers with recent PhDs... 
    Flexible hours

    MSCI

    San Francisco, CA
    1 day ago
  • $150k - $300k

     ...AI research lab and infrastructure provider working on LLM interpretability and context optimization. The team builds custom machine learning models that analyze and compress token contexts before they reach the underlying model, cutting inference costs by roughly 50%... 
    Full time
    H1b
    Visa sponsorship

    David Joseph & Company

    San Francisco, CA
    3 days ago
  •  ...pharma\'s drug discovery workflows, where scientists use them to solve the highest value drug...  ..., feature extraction, representation learning, model training, evaluation, inference,...  ...or independently, that shows exceptional machine learning ability. You are deeply technical... 

    Axiom company

    San Francisco, CA
    4 days ago
  •  ...environments. Our robotic arms run on tight compute and power budgets, so learning systems have to be fast, reliable, and deeply integrated with...  ...that let robots see, reason, and act. This work runs on real machines, not benchmarks. About the role As an AI Researcher at... 

    Droyd Robotics

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Machine Learning Scientist. Be the first to apply!