User Researcher, AI Evaluations
$196k - $230kNotion Labs
Who We AreNotion is the collaborative AI workspace where teams and agents think together. We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion.Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work.About the Role:We’re seeking an experienced UX Researcher to define and scale how we evaluate Notion’s AI-powered experiences—focusing on what “good” looks like not only for model output quality, but for the end-to-end product experience where people discover, set goals, delegate work, review results, and build trust over time with AI.This role sits at the intersection of research craft and evaluation operations: you’ll run studies that uncover user mental models, expectations, and failure/recovery behaviors, then translate those insights into reusable rubrics, workflows, and measurement approaches that product, design, engineering, and data science can apply consistently.This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You'll Achieve:Define what “good” looks like (frameworks & rubrics): Establish clear, reusable evaluation criteria that reflect real user expectations—helpfulness, trust, tone, control, and transparency. You’ll translate qualitative insight into scoring guidance that can be applied consistently across teams and over time.Run recurring evals (longitudinal & feature-specific): Run recurring longitudinal and feature-specific surveys and studies to measure experience quality over time against defined rubrics. Lead qualitative studies, side-by-side comparisons, and human-in-the-loop evaluation efforts to deepen understanding of where experiences break down and how they can improve. You’ll help teams spot regressions, benchmark improvements, and understand when expectations shift.Anchor evaluation in real workflows (context > isolated feedback): Ensure evals reflect jobs-to-be-done, user intent, and the full interaction journey (goal setting, delegation, review, iteration), not just decontextualized thumbs up/down. You’ll help teams understand who is evaluating, what they’re trying to do, and why outputs succeed or fail.Identify failure modes & recovery behavior (guardrails): Uncover breakdowns, regressions, and edge cases across the system—from model behavior to UI and integrations—and study how people notice issues, correct them, and continue their work. You’ll turn these insights into actionable guidance for guardrails, fixes, and prioritization.Operationalize evaluation with partners (process & tooling): Collaborate closely with Product, Design, Engineering, and Data Science to align on target use cases and build scalable evaluation loops (human-in-the-loop review, longitudinal studies, and calibration of automated/LLM-judge approaches against human judgment).Skills You'll Need to Bring: Ability to operationalize insight into measurement: You’re comfortable turning “soft” user expectations (trust, tone, usefulness, clarity) into concrete rubrics, scoring guidelines, and observable metrics.AI fluency and systems thinking: You’re curious and hands-on with AI products, and can reason about how model behavior, uncertainty, and system constraints shape user experience. You also have experience evaluating AI-enabled products (LLMs, agents, generative UI/workflow automation) and working with Data Science/ML partners on measurement strategy and evaluation tooling.Clear communication and impact orientation: You can align diverse partners around shared definitions of quality and create artifacts that enable teams to act consistently. You tailor storytelling to different audiences, connect research to business outcomes, and drive follow-through so insights translate into product change.Strong UX research craft (quant + qual): You can choose the right methods for the question— interviews, benchmarking, surveys, experiments—and synthesize into actionable guidance. You also can prioritize ruthlessly, work through ambiguity, and balance scrappy iteration with deep dives when needed.Pragmatism in fast-moving environments: You can prioritize ruthlessly, work through ambiguity, and balance scrappy iteration with deep dives when needed.Experience: 5+ years doing UX research in industryNice to Haves:Familiarity with LLM-as-judge methods, prompt design for evaluators, or “golden dataset” creationExperience using AI research tooling for rapid synthesis and communication (e.g., Dovetail, Listen Labs, Maze, Outset, etc.), as well as AI observability tooling like BraintrustExperience using data querying languages (e.g., SQL), scripting languages (e.g., Python), or statistical/mathematical software (e.g., R, SAS, Matlab, etc.)Master’s or PhD in HCI, Psychology, Behavioral Science, Anthropology, Sociology, or a related fieldYou’re familiar with the work of computing heroes like Douglas Engelbart, Alan Kay, Bret Victor, etc. — and understand why we're big fans.Notion is committed to providing highly competitive cash compensation, equity, and benefits. The compensation offered for this role will be based on multiple factors such as location, the role’s scope and complexity, and the candidate’s experience and expertise, and may vary from the range provided below. For roles based in San Francisco or New York City, the estimated base salary range for this role is $196,000-$230,000 per year.By clicking “Submit Application”, I understand and agree that Notion and its affiliates and subsidiaries will collect and process my information in accordance with Notion’s Global Recruiting Privacy Policy and NYLL 144.#LI-OnsiteA Note on AIYou don’t need deep AI expertise for every role, but we do expect every Notino to be intellectually curious, drawn to tinkering and discovery, and excited to use AI as a real collaborator in their work. For some roles, AI fluency is a core requirement — when that’s the case, we'll say so explicitly in the qualifications. People who thrive here don’t treat AI as a novelty. They use it to think better, and make their work easier for others to build on.Equal Opportunity & AccommodationsWe hire talented people from a wide range of backgrounds. If you’re excited about this role but don’t meet every bullet, we still encourage you to apply. Notion is an equal opportunity employer and does not discriminate on the basis of any legally protected characteristic. Consistent with applicable law, we will consider for employment qualified applicants with arrest and conviction records. Notion provides reasonable accommodations during the application process; if you need one, please let your recruiter know.Notion is proud to be an equal opportunity employer. We do not discriminate in hiring or any employment decision based on race, color, religion, national origin, age, sex (including pregnancy, childbirth, or related medical conditions), marital status, ancestry, physical or mental disability, genetic information, veteran status, gender identity or expression, sexual orientation, or other applicable legally protected characteristic. Notion considers qualified applicants with criminal histories, consistent with applicable federal, state and local law. Notion is also committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or an accommodation due to a disability, please let your recruiter know.LocationSan Francisco, California; New York, New YorkEmployment TypeFull timeLocation TypeHybridDepartmentUser Research and Product Operations
$164k - $190k
...Who We Are Notion is the collaborative AI workspace where teams and agents think together... ...the Role: We’re seeking an experienced User Researcher to deliver insights that improve Notion’... .../usability testing, and post‑ship evaluation — with clear recommendations that drive...SuggestedLocal area$120k - $150k
...Role Summary Senior User Experience Researcher – Montefiore’s Experience Design and End‑User Research... ...Responsibilities Research Lead generative and evaluative research across web, mobile, voice,... ...for emerging technologies such as AI, wearables, or voice interfaces. Experience...SuggestedMonday to FridayShift work- ...Role Overview As a User Researcher at Glance, you'll investigate how users interact with lock screen... ...lock screen content Collaborate with AI/ML teams to improve content... ...Create and maintain research frameworks for evaluating user engagement with short-form content...SuggestedLocal area
$159k - $230k
Drive research strategy by independently identifying, prioritizing, and... ...opportunities based on user needs, product telemetry, and... ...Conduct tactical and strategic evaluations to streamline user journeys, identifying... ...operates, from accelerating AI/ML deployments to reducing...Suggested$55 - $65 per hour
...ours. We are searching for a Part-Time UX Researcher for our faith-based tech client. In this... ...If you’re passionate about understanding user needs and bringing purpose-driven work to... ...insights directly into each article, started with the help of AI. #J-18808-Ljbffr...SuggestedContract workPart timeRemote work$45 - $50 per hour
...Senior User Researcher, Insights (Contractor) Senior User Researcher, Insights (Contractor) This range is provided by Swell Partners. Your... ...community knowledge in a new way. Experts add insights directly into each article, started with the help of AI. #J-18808-Ljbffr...Full timeContract workFor contractorsFreelanceRemote work- ...Robinhood is seeking a Staff User Researcher to lead cross‑product research for Brokerage, partnering with product and business leaders to... ...quantitative methods, and integrate customer insights into product decisions while embracing AI-enabled workflows #J-18808-Ljbffr...
$30 per hour
...Prolific is not just another player in the AI space – we are building the biggest pool... ...world. Over 35,000 AI developers, researchers, and organizations use Prolific to gather... ...as Domain Experts for a high-level AI evaluation project. AI models are evolving beyond simple...Remote workWork from homeFlexible hours$214k - $255k
...experiences across traditional UI and AI‑powered (agentic) workflows, shaping how users interact with data, automation,... ...range. Experiment with and evaluate emerging AI design tools, advocating... ...generative design tools, LLM‑assisted research synthesis, AI‑powered prototyping...Live inWork at office- ...opportunities for improvement. • Proficiency in user research methods such as interviews, surveys,... ...• Experience conducting accessibility evaluations and ensuring service designs meet or... ...(e.g., WCAG, ADA compliance), utilizing AI-assisted tools to automate accessibility...
$180.3k - $244k
...interaction vocabulary alongside researchers, scientists, and engineers.... ...iterative and demo-led: build fast, evaluate honestly, and let each round... ...should change - Direct user research that exposes our blind... ...assistants, dialogue systems, AI companions) - Experience with...Flexible hours$20 per hour
A technology-focused AI company is seeking a UI/UX Product Designer to enhance and evaluate AI systems' design capabilities. This remote role allows you to choose projects and work on your own schedule. Ideal candidates should have a strong background in UI/UX design,...Remote job- ...operating model to deliver on it. The organization is becoming AI-native and agentic-supported, putting enterprise AI platforms in... ...healthcare or pharmaceutical sectors.Experience with AI experience evaluation approaches (conversation quality testing, scenario-based risk...
$80k - $90k
...assisting with scripting, storyboarding, and basic editing using AI-driven and traditional tools Support the administration of... ...accessibility and alignment with L&D strategy Participate in the evaluation of learning effectiveness by gathering feedback and analyzing learner...Full timeTemporary workWork at officeLocal area$20 per hour
A technology-driven design firm in the United States is seeking an Experience Designer to enhance AI models that generate and evaluate design work. The role involves critiquing AI outputs in UI/UX and working flexibly on various projects. Candidates should have a background...Remote job- ...role will be pivotal in defining user experiences that deliver... ...expertise in UX design, user research, and system thinking, with the... ...surveys, interviews, and heuristic evaluations. Define and review user... ...our teams harness cutting-edge AI and breakthrough technologies...WorldwideFlexible hours
- ...Associate Consultant - Senior User Experience Designer... ...workstreams in designing and delivering AI-powered digital products, predictive... ...strategy, translate research into scalable GenAI design patterns... ...Conduct or support generative and evaluative research, synthesizing...Work at officeLocal areaWork from homeWorldwideFlexible hours
$160k - $200k
...Senior UX Designer - AI SaaS Platform $160K-$200K Base + Bonus + Equity... ...complex data and workflows into intuitive user experiences. You'll collaborate... ...Hands-on experience conducting user research, usability testing, and evaluations ~ Excellent communication skills,...Remote workFlexible hours2 days per week3 days per week- ...Position SummaryShape the long-term user experience vision for web, mobile, and AI experiences that help millions of... ...focused solutions.Partner with UX research for discovery work, research... ...accessibility, and long-term experience value.Evaluate emerging technologies and industry...Second jobWork at officeRemote work
$25 - $26 per hour
...landing pages: create visually appealing, user-friendly, and effective website layouts... ...(like WCAG) in mind Staying current: research and evaluate new design trends, techniques, software,... ...document design. General knowledge of AI and modern technologies to drive innovation...Hourly payFull timeWork at officeImmediate startMonday to Friday$123.6k - $151.85k
...feedback, program data, and evolving business prioritiesLeverage AI-based tools and emerging learning technologies to accelerate... ...sized, engaging learning in the flow of workParticipate in program evaluation efforts, analyzing learner feedback and completion data to recommend...Local area$20 per hour
...DataAnnotation is committed to creating quality AI. Join our team to help train AI chatbots... ...responses to demonstrate excellence, and evaluate different model outputs based on accuracy... ...performance of different AI models Research and fact‑check AI responses...Hourly payFull timeContract workPart timeFor contractorsSelf employmentFreelanceRemote work$20 per hour
...Experience Designer to help train and improve cutting‑edge AI models. You’ll assess and enhance how AI systems understand, generate, and evaluate design work — including interfaces, layouts, visuals, and user experiences. In this role, you’ll use your design expertise...Full timeContract workPart timeRemote work$50 - $80 per hour
...Android experiences that support users in diverse, real-world working... .... Explore and design AI-enabled experiences, including... .... Advocate for users through research, feedback analysis, and data-informed... ...process your personal data to evaluate your candidacy and share...Hourly payContract workRemote work$25 - $40 per hour
DataAnnotation is seeking an experienced Web Developer/Designer to help train AI models by reviewing and critiquing AI-generated UI/UX designs. This role involves evaluating design quality to improve the user experience and aesthetic appeal of AI outputs. Applicants must be...Remote jobFor contractorsFlexible hours$110k - $150k
...comfortable exploring emerging tools such as Cavalry, Jitter, and AI-powered motion tools.You have a robust portfolio of work that... ...team’s efficiency and improve the candidate experience — not to evaluate or decide on your candidacy. Participation in AI-supported interviews...Permanent employmentFreelanceLive inFlexible hours$197.3k - $313.7k
...DetailsAbout SalesforceSalesforce is the #1 AI CRM, where humans with agents drive... ...Salesforce.Join a collaborative, diverse team of researchers at Agentforce Operations. The... ...possess experience developing, deploying, and evaluating machine learning models in production or...Full timeImmediate startRemote work- Our Client is a rapidly growing AI-powered B2B SaaS start-up that is disrupting the life insurance market. Their AI-powered intelligent backend is used in the life insurance space to improve workflows, automate manual processes, and disrupt outdated practices. They are...Full timeWork at officeLocal area
$200k - $300k
Hudson River Trading (HRT) is hiring an AI Researcher to join the HAIL team. HAIL (HRT AI Labs) is the team at HRT responsible for developing... ...instructed or agreed upon. We employ various methods to evaluate the authenticity of candidate responses. If we determine that...Work experience placementWork at officeLocal areaImmediate start$200k - $300k
Hudson River Trading (HRT) is seeking an LLM-focused AI Researcher to join the HAIL team. HAIL (HRT AI Labs) is the team at HRT responsible... ..., dataset curation and mixing, pretraining, post-training, evaluation design, inference, and live trading. HAIL researchers have...Work at officeLocal areaImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to User Researcher, AI Evaluations. Be the first to apply!


