User Researcher, AI Evaluations (San Francisco)
$196k - $230kNotion Labs
Who We AreNotion is the collaborative AI workspace where teams and agents think together. We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion.Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work.About the Role:We’re seeking an experienced UX Researcher to define and scale how we evaluate Notion’s AI-powered experiences—focusing on what “good” looks like not only for model output quality, but for the end-to-end product experience where people discover, set goals, delegate work, review results, and build trust over time with AI.This role sits at the intersection of research craft and evaluation operations: you’ll run studies that uncover user mental models, expectations, and failure/recovery behaviors, then translate those insights into reusable rubrics, workflows, and measurement approaches that product, design, engineering, and data science can apply consistently.This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You'll Achieve:Define what “good” looks like (frameworks & rubrics): Establish clear, reusable evaluation criteria that reflect real user expectations—helpfulness, trust, tone, control, and transparency. You’ll translate qualitative insight into scoring guidance that can be applied consistently across teams and over time.Run recurring evals (longitudinal & feature-specific): Run recurring longitudinal and feature-specific surveys and studies to measure experience quality over time against defined rubrics. Lead qualitative studies, side-by-side comparisons, and human-in-the-loop evaluation efforts to deepen understanding of where experiences break down and how they can improve. You’ll help teams spot regressions, benchmark improvements, and understand when expectations shift.Anchor evaluation in real workflows (context > isolated feedback): Ensure evals reflect jobs-to-be-done, user intent, and the full interaction journey (goal setting, delegation, review, iteration), not just decontextualized thumbs up/down. You’ll help teams understand who is evaluating, what they’re trying to do, and why outputs succeed or fail.Identify failure modes & recovery behavior (guardrails): Uncover breakdowns, regressions, and edge cases across the system—from model behavior to UI and integrations—and study how people notice issues, correct them, and continue their work. You’ll turn these insights into actionable guidance for guardrails, fixes, and prioritization.Operationalize evaluation with partners (process & tooling): Collaborate closely with Product, Design, Engineering, and Data Science to align on target use cases and build scalable evaluation loops (human-in-the-loop review, longitudinal studies, and calibration of automated/LLM-judge approaches against human judgment).Skills You'll Need to Bring: Ability to operationalize insight into measurement: You’re comfortable turning “soft” user expectations (trust, tone, usefulness, clarity) into concrete rubrics, scoring guidelines, and observable metrics.AI fluency and systems thinking: You’re curious and hands-on with AI products, and can reason about how model behavior, uncertainty, and system constraints shape user experience. You also have experience evaluating AI-enabled products (LLMs, agents, generative UI/workflow automation) and working with Data Science/ML partners on measurement strategy and evaluation tooling.Clear communication and impact orientation: You can align diverse partners around shared definitions of quality and create artifacts that enable teams to act consistently. You tailor storytelling to different audiences, connect research to business outcomes, and drive follow-through so insights translate into product change.Strong UX research craft (quant + qual): You can choose the right methods for the question— interviews, benchmarking, surveys, experiments—and synthesize into actionable guidance. You also can prioritize ruthlessly, work through ambiguity, and balance scrappy iteration with deep dives when needed.Pragmatism in fast-moving environments: You can prioritize ruthlessly, work through ambiguity, and balance scrappy iteration with deep dives when needed.Experience: 5+ years doing UX research in industryNice to Haves:Familiarity with LLM-as-judge methods, prompt design for evaluators, or “golden dataset” creationExperience using AI research tooling for rapid synthesis and communication (e.g., Dovetail, Listen Labs, Maze, Outset, etc.), as well as AI observability tooling like BraintrustExperience using data querying languages (e.g., SQL), scripting languages (e.g., Python), or statistical/mathematical software (e.g., R, SAS, Matlab, etc.)Master’s or PhD in HCI, Psychology, Behavioral Science, Anthropology, Sociology, or a related fieldYou’re familiar with the work of computing heroes like Douglas Engelbart, Alan Kay, Bret Victor, etc. — and understand why we're big fans.Notion is committed to providing highly competitive cash compensation, equity, and benefits. The compensation offered for this role will be based on multiple factors such as location, the role’s scope and complexity, and the candidate’s experience and expertise, and may vary from the range provided below. For roles based in San Francisco or New York City, the estimated base salary range for this role is $196,000-$230,000 per year.By clicking “Submit Application”, I understand and agree that Notion and its affiliates and subsidiaries will collect and process my information in accordance with Notion’s Global Recruiting Privacy Policy and NYLL 144.#LI-OnsiteA Note on AIYou don’t need deep AI expertise for every role, but we do expect every Notino to be intellectually curious, drawn to tinkering and discovery, and excited to use AI as a real collaborator in their work. For some roles, AI fluency is a core requirement — when that’s the case, we'll say so explicitly in the qualifications. People who thrive here don’t treat AI as a novelty. They use it to think better, and make their work easier for others to build on.Equal Opportunity & AccommodationsWe hire talented people from a wide range of backgrounds. If you’re excited about this role but don’t meet every bullet, we still encourage you to apply. Notion is an equal opportunity employer and does not discriminate on the basis of any legally protected characteristic. Consistent with applicable law, we will consider for employment qualified applicants with arrest and conviction records. Notion provides reasonable accommodations during the application process; if you need one, please let your recruiter know.Notion is proud to be an equal opportunity employer. We do not discriminate in hiring or any employment decision based on race, color, religion, national origin, age, sex (including pregnancy, childbirth, or related medical conditions), marital status, ancestry, physical or mental disability, genetic information, veteran status, gender identity or expression, sexual orientation, or other applicable legally protected characteristic. Notion considers qualified applicants with criminal histories, consistent with applicable federal, state and local law. Notion is also committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or an accommodation due to a disability, please let your recruiter know.LocationSan Francisco, California; New York, New YorkEmployment TypeFull timeLocation TypeHybridDepartmentUser Research and Product Operations
- ...looking for a Staff UX Researcher to join our Design... ...Research team and drive the user research agenda that... ...to work from our San Francisco or Helsinki offices at... ...research across generative, evaluative, and strategic studies... ...spaces — particularly AI- and LLM-powered...SuggestedPart timeWork at officeLocal areaRemote work
$172.5k - $260.1k
...SalesforceSalesforce is the #1 AI CRM, where humans with... ...delivering scalable, user-centered solutions... ...the broader portfolio.Research & Insight: Partner with... ...recruiters assess and evaluate candidates’ resumes and... ...following link: to the San Francisco Fair Chance Ordinance and...SuggestedFull timePart time$139.7k - $265.8k
...The OpportunityOn the Design Research & Strategy team, our mission is to lay a user-centered foundation for Adobe product... ...research that shapes Adobe’s future AI investments across our creative... ...and civil liability.SummaryLocation: San Francisco; San JoseType: Full time...SuggestedFull timeTemporary workPart timeFor contractorsLocal areaWorldwide$170k - $200k
...build better, more human customer experiences with AI. We are primarily an in-person company based in San Francisco, with growing offices in Atlanta, New York,... ...precisely match the job description. We strive to evaluate all applicants consistently without regard to race...SuggestedFull timePart timeFlexible hours- ...based on-site in our San Francisco office. Joining Schwab... ...financial planning.Schwab’s AI Strategy &... ...our clients. As an AI Researcher on AI.x, you will play... ...research programs and evaluation methodologies that examine... ...platforms that impact users at scale.Experience designing...SuggestedFull timePart timeWork at office
$197.3k - $313.7k
...SalesforceSalesforce is the #1 AI CRM, where humans with... ...collaborative, diverse team of researchers at Agentforce Operations. The... ...developing, deploying, and evaluating machine learning models in production... ...the following link: to the San Francisco Fair Chance Ordinance and...Full timePart timeImmediate startRemote work- ...site in the specified location(s).As an AI Researcher within Schwab’s AI Strategy &... ...problem formulation through modeling, evaluation, deployment, and iteration, contributing... ...capabilities.This is a hybrid role based in San Francisco, CA, with regular in‑office...Full timePart timeWork at office
$210k - $250k
...easily accessible for all users.We're looking for a... ...design and AI engineering: someone who... ...with AI, training it, evaluating it, and questioning every... ...day on-site role in our San Francisco office. What You'll DoStart... ...to PMs, engineers, ML researchers, and executives.A...Full timePart timeWork at office$137.8k - $186.4k
...will have the opportunity to do hands-on research and collaborate closely with engineers.... ...experiencesHave experience incorporating AI tooling as part of your design workflows... ...Employee DiscountPursuant to the San Francisco Fair Chance Ordinance, we will consider...Part timeFlexible hours$140.4k - $190k
...team of designers and researchers, working closely with... ...member experience.Bring AI fluency to every stage... ..., and usability evaluations, to deeply understand... ...product experiences for end users that feel natural, trustworthy... ...the US and Hybrid for San Francisco (SF). Candidates must...Full timePart timeCurrently hiringWork at officeLocal areaRemote work3 days per week$147k - $187k
..., and control spend effortlessly. Brex’s AI-native automation and world-class service... ...strategy, or harnessing AI to empower our users, we obsess over quality and clarity. This... ...’ll workThis role will be based in our San Francisco, Seattle, or New York office. We are a hybrid...Part timeFreelanceWork at officeRemote workWork from home3 days per week$120k - $180k
...Cut Pro, Cinema 4D, and Rive, along with experience utilizing AI platforms to enhance creativity and efficiency.Business Alignment... ...more than 2,550 employers. The company is headquartered in San Francisco with additional offices in Montreal and Bangalore. Learn more at...Part timeLocal areaWorldwide- ...building the real-world AI interface for... ...loved by over 2,000,000 users worldwide since 2023.... ...Delaware-incorporated, San Francisco-based company pushing... ...you will do Execute research programs end-to-end: design... ...contextual inquiry, ergonomic evaluation, form factor...WorldwideShift work
$164k - $190k
...Notion is the collaborative AI workspace where teams and agents... ...’re seeking an experienced User Researcher to deliver insights that... ...role can be based in either San Francisco or New York City. We work from... ...testing, and post‑ship evaluation — with clear recommendations...Local area$120k - $135k
...CliftonStrengths through research that anchors what we... ...and why. As a senior user researcher at Gallup,... .... Lead generative and evaluative research using a range... ...Thoughtfully incorporate AI tools into research... ...working on‑site at Gallup’s San Francisco office at least three...Work at office3 days per week$212k - $292k
...Advanced Technology Group (ATG) is the research division of the company. ATG’s mission is... ...science and electrical engineering, such as AI/ML, algorithms, digital signal... ...about our innovative research: #LI-PO1The San Francisco/Bay Area base salary range for this full...Full timePart timeLocal areaWorldwideFlexible hours$180k - $210k
...ship looks and feels like it belongs to one of the most exciting AI products in the world — clear, beautiful, and fast.You're a great... ...Range: $180K - $210KLocationNew York City; Remote (United States); San FranciscoEmployment TypeFull timeDepartmentDesignCompensation$180...Part timeLocal areaRemote work$146.3k - $274.3k
...Experience Orchestration (CXO) AI Foundation, the team building... ...Management, Engineering, Research, and Marketing partners. The... ...exceptional experiences for Adobe users.About AdobeAdobe empowers... ...civil liability.SummaryLocation: San Francisco; San JoseType: Full time...Full timeTemporary workPart timeLocal areaImmediate startWorldwide$240k - $300k
..., and control spend effortlessly. Brex’s AI-native automation and world-class service... ...strategy, or harnessing AI to empower our users, we obsess over quality and clarity. This... ...’ll workThis role will be based in our San Francisco or New York office. We are a hybrid environment...Temporary workPart timeWork at officeRemote workWork from home- ...office. We’re looking for an experienced, San Francisco based Senior Content Designer who’s... ..., physiology and behavior science, and AI. An ideal candidate can write and prompt... ...clinical, legal, regulatory, engineering, user research, and localization teams to create...Part timeWork at officeLocal areaRemote workHome officeFlexible hoursNight shiftEarly shift
$175k - $190k
...San Francisco, New York CityMarketing /Full Time /HybridAbout FinchWe are on a mission to revolutionize employment by building the infrastructure... ...components.Build templates, libraries, and standards—and the AI-enabled workflows around them—that empower the marketing team...Full timePart timeLive inWork at office2 days per week$212.3k - $275.8k
...applications are received.Meet the TeamAt Foundation AI, we are leading frontier AI research across Cisco. Our mission is to advance the state of artificial... ..., reasoning systems, scalable training algorithms, evaluation science, inference optimization, and AI systems...Full timeTemporary workPart timeLocal areaFlexible hours$148.5k - $223.9k
...SalesforceSalesforce is the #1 AI CRM, where humans with agents... ...Systems role sits within our User Interface & User Experience... ...product suite. Based in the San Francisco Bay Area with hybrid flexibility... ...our recruiters assess and evaluate candidates’ resumes and qualifications...Full timePart time$169k - $303k
...translating designs into code, or iterating with AI. From idea to product, Figma empowers... ...and become the bridge connecting our AI research innovations with world-class design... ...models' design capabilities and oversee evaluation quality.This is an exciting opportunity to...Minimum wageFull timePart timeFor contractorsLocal areaRemote workFlexible hours$190k - $290k
...customer experiences with AI. We are primarily an in... ...company based in San Francisco, with growing offices in... ...iterate to ensure we solve user problems effectively.... ...to various stakeholders.Research Skills: Experience... ...description. We strive to evaluate all applicants consistently...Full timePart timeFlexible hours$118k - $309.8k
...Educational Research Scientist Department of Surgery and the Center for... ...: University of California, San FranciscoThe University of California, San Francisco (UCSF) Department of Surgery and... ...advance education scholarship and evaluate the effectiveness of educational...Full timePart timeWork at office$30 per hour
...AI Trainer – Visual & Graphic Design Expert (Remote) About Prolific Prolific... ...in the world. Over 35,000 AI developers, researchers, and organizations use Prolific to gather... ...act as Domain Experts for a high‑level AI evaluation project. AI models are evolving beyond...Remote workWork from homeFlexible hours$170k - $250k
...Who We AreHP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global... ..., world-class team—engineers, designers, researchers, and product minds—focused on creating an... ...simple, delightful, and transformative user experiences. We’re passionate about...Full timeTemporary workPart timeLocal areaFlexible hours$15 - $30 per hour
...scenes and assets to join our global network for a high level gaming project. This is an exciting opportunity to contribute to improving AI systems for video generation and to impact how AI models understand the logic of the physical world. Rather than traditional VFX...Hourly payRemote work$102.6k - $309.8k
...Translational ResearcherThe Department of Psychiatry and Behavioral Sciences at the University of California, San Francisco is dedicated to building strong, research-focused clinical programs across the department. In support of this goal, the department is committed to...Part time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to User Researcher, AI Evaluations (San Francisco). Be the first to apply!


















