Director, Research - AI Evals
$258k - $348kFigma
Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us!The Figma Research team is hiring an Director, Research - AI Evals to own how we measure the quality of Figma's AI-powered experiences. As Figma ships more AI capabilities across our products, the question "is this really good?" has never mattered more — and answering it rigorously is what this role exists to do. You'll define what "good" means for our AI features, build the frameworks and quality bars to measure it, and turn that into trusted signal that product teams rely on to decide what to ship.The ideal candidate brings deep, hands-on experience evaluating AI/LLM-powered products — blending human evaluation with automated, model-based approaches — along with the product instinct and communication skills to make evaluation genuinely useful. Partnering with Product, Design, Engineering, and Data Science, you'll sit upstream of nearly every AI shipping decision at Figma and directly shape the quality of features used by millions of people.This is a full time role that can be held from one of our US hubs or remotely in the United States.What you'll do at Figma:Own AI evaluation methods and operations for Figma's AI-powered experiences — define quality dimensions, design how we measure them, and turn results into decision-ready signalBuild and maintain evaluation frameworks, rubrics, golden datasets, and quality bars, combining human evaluation with automated/model-based approaches (e.g., LLM-as-judge) where appropriatePartner with engineering to stand up repeatable, reproducible evaluation pipelines and regression testing, so evaluation is a routine part of how AI features are built and shippedProduce clear readouts and dashboards that let stakeholders confidently make go/no-go and prioritization decisionsSocialize a shared definition of quality so evaluation standards are adopted across teams rather than re-invented — and advocate for evaluation as a strategic partner in the product processManage a small team to execute our AI evals in partnership with contractors, internal staff, and/or LLMsWe'd love to hear from you if you have:10+ years of experience in product, research, applied research, or a closely related field, including 2+ years of management experienceDirect, hands-on experience owning the evaluation of AI/LLM-powered productsExpertise designing and running AI evaluation — human evaluation programs, rubric and benchmark/golden-dataset construction, inter-rater reliability — and sound judgment about when and how to apply automated/model-based approaches (e.g., LLM-as-judge), including their limitationsStrength across both qualitative and quantitative methods, comfort with data and metrics, and the ability to reason about model behaviorDemonstrated success in identifying the riskiest assumptions behind an ambiguous quality question, prioritizing them, and designing right-sized evaluation to build confidenceA proven track record of gaining buy-in from executive and cross-disciplinary stakeholders — transcending methodology to articulate a larger user story and the "so what" to inspire actionWhile it's not required, it's an added plus if you also have:Experience building or co-building automated evaluation pipelines and regression testing in partnership with engineering, or familiarity with eval tooling (e.g., Braintrust, LangSmith, DeepEval, or equivalents)Experience standing up a new function, practice, or discipline from scratch2+ years in product design, user-centric product management, data science, product development, and/or front-end engineeringA familiarity and depth of experience using Figma's productsAt Figma, one of our values is Grow as you go. We believe in hiring smart, curious people who are excited to learn and develop their skills. If you’re excited about this role but your past experience doesn’t align perfectly with the points outlined in the job description, we encourage you to apply anyways. You may be just the right candidate for this or other roles.Pay Transparency DisclosureJob level and actual compensation will be decided based on factors including, but not limited to, individual qualifications objectively assessed during the interview process (including skills and prior relevant experience, potential impact, and scope of role), market demands, and specific work location. Figma offers equity to employees, as well as a competitive package of additional benefits, including health, dental, and vision coverage; retirement benefits with company contributions; parental leave and reproductive or family planning support; mental health and wellness benefits; and paid time off. Figma provides paid sick leave, holidays, and other leave benefits in compliance with applicable federal, state, and local laws, including the requirements of the Washington Minimum Wage Act and related regulations. Exempt employees are eligible for employer‑provided paid flexible PTO in addition to flexible paid sick leave. PTO is subject to manager approval. Additional benefits may include company recharge days, cell phone and home internet reimbursements, and a number of lifestyle spending accounts. Figma also offers sales incentive compensation for most sales roles and an annual bonus plan for eligible non-sales roles. All compensation and benefits are subject to applicable plan terms and may be modified by Figma at any time, consistent with applicable law.Annual Base Salary Range:$258,000—$348,000 USDAt Figma we celebrate and support our differences. We know employing a team rich in diverse thoughts, experiences, and opinions allows our employees, our product and our community to flourish. Figma is an equal opportunity workplace - we are dedicated to equal employment opportunities regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity/expression, veteran status, or any other characteristic protected by law. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements.We will work to ensure individuals with disabilities are provided reasonable accommodation to apply for a role, participate in the interview process, perform essential job functions, and receive other benefits and privileges of employment. If you require accommodation, please reach out to View email address on click.appcast.io. These modifications enable an individual with a disability to have an equal opportunity not only to get a job, but successfully perform their job tasks to the same extent as people without disabilities. Examples of accommodations include but are not limited to: Holding interviews in an accessible locationEnabling closed captioning on video conferencingEnsuring all written communication be compatible with screen readersChanging the mode or format of interviews To ensure the integrity of our hiring process and facilitate a more personal connection, we require all candidates keep their cameras on during video interviews. Additionally, if hired you will be required to attend in person onboarding.By applying for this job, the candidate acknowledges and agrees that any personal data contained in their application or supporting materials will be processed in accordance with Figma's Candidate Privacy Notice.
- ...EPAM's new Frontier AI business unit partners directly with leading AI labs and advanced AI organizations, translating their research and post-training objectives into technically rigorous... ..., and training data. The Senior Director, Frontier AI Research is EPAM's...Suggested
$365k
...create reliable, interpretable, and steerable AI systems. We want AI to be safe and... ...is a quickly growing group of committed researchers, engineers, policy experts, and business... ...move across research areas like compute, evals, RL environments, and emerging research initiatives...SuggestedWork at officeVisa sponsorshipFlexible hoursShift work- ...create reliable, interpretable, and steerable AI systems. We want AI to be safe and... ...is a quickly growing group of committed researchers, engineers, policy experts, and business... ...adapted Epoch's Capabilities Index to our evals); all the data in When AI Builds Itself comes...SuggestedFull time
$250k
Our client, an AI-driven Healthcare company, are hiring a Head of AI Research to join their team in San Francisco. The successful candidate will work on building high-impact, complex AI systems across clinical intelligence, outcome modelling and advanced learning frameworks...SuggestedFull time- A cutting-edge AI company in San Francisco is seeking a Head of Research to drive their research agenda in LLM efficiency. The ideal candidate will lead an applied research team, define the strategy for model routing and training, and work closely with engineering to translate...Suggested
$340k - $425k
A leading AI research organization is seeking a Manager for its Interpretability team in San Francisco. The ideal candidate will have a strong background in managing technical teams and a passion for AI safety research. This role involves overseeing project execution, supporting...Work at officeFlexible hours- ...chain management, and scalable manufacturing. As robotics, Physical AI, AI infrastructure, and advanced manufacturing continue to... ...frontier technology initiatives. Track market and technology trends, research startups and technology companies, prepare management briefings,...
- ...driven transformation in medicine. Driven by the push for streamlined drug development, the market for advanced analytics and AI in clinical research is expanding exponentially. The PReDiCTR-TB Consortium is not just following industry standards—we are creating and...TraineeshipWorldwide
$130k - $190k
...poster from Two Point Consulting Vice President of Recruitment at Two Point Consulting Responsibilities Legal/professional services AI tools Creating AI automated workflows Working with IT and knowledge management teams Vetting new AI tools Providing trainings and...Full time- ...AI Research Engineer (Robot Learning) San Francisco AI & Software In office Full-time mimic is an early stage deep tech robotics... ..., to data collection, model training and real world evals Write clean, maintainable python code using deep learning frameworks...Full timeWork at officeImmediate start
- ...Washington D.C., London and Amsterdam. We are the Data Foundation & AI team within Plaid’s Data organization. Our mission is to build... ...cases and build scalable, repeatable pipelines that translate research into production impact. You will also partner closely with teams...Full timeWork experience placementLocal areaImmediate start
- ...Team The Fleet team builds core components to enable productive research from small to state of the art scale across OpenAI, with the... ...Kubernetes Experienced in CI/CD About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that...Full timeWork at officeRelocation package
- ...data usability, quality, and impact on key benchmarks and guide data generation, augmentation, and curation. Collaborate with AI researchers, applied AI teams, and data producers to align evaluations with training objectives. Own benchmarking, evaluation, and...Full timeWork at officeRelocation package
$155k - $269k
...Job Description Job Description Waabi, founded by AI visionary Raquel Urtasun, is the leader in Physical AI. With a world-class... ...realistic, scalable, controllable, and efficient simulation. As a Research Engineer in the World Models team, you will develop algorithms...Full timeWork at officeWork from homeFlexible hours- ...integrate successful approaches into production systems. Track research in LLMs, information retrieval, and developer tooling.... ...to $25,000 in relocation assistance. Opportunity to work on AI-powered code review used by thousands of developers and more than...Full timeWork at officeRelocation package
$140k - $160k
...at Consumer Reports. We're seeking a highly capable Senior Research Engineer to join our team and drive forward exploratory research... ...engineering with particular focus on innovative applications of AI, machine learning, and emerging technologies. In this role, you'...Full timeLocal areaRemote workRelocation package$134k - $235k
...Job Description Job Description Waabi, founded by AI visionary Raquel Urtasun, is the leader in Physical AI. With a world-class... ...impact the world in a positive way. To learn more visit: As a Research Engineer in Neural Rendering, you will create the next...Full timeWork at officeWork from homeFlexible hours- ...company based in San Francisco, California. The Role: As a Research Engineer - Brain Computer Interface Models , you will be a core... ...low as possible We all enjoy what we do and love discussing AI Benefits and Perks: Comprehensive medical, dental, vision...Work at officeRelocation package
- ...’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and... ...society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working...Full timeVisa sponsorshipShift work
- ...Job Description Job Description About the Role This is a Research Engineer role focused on building privacy and anonymization systems that make sensitive, real-world data safe and useful for AI training. You will own the full pipeline for protecting privacy without...Visa sponsorship
- ...Job Description Job Description Research Engineer — AI Alignment & Evaluation AI Safety / Research Engineering | San Francisco, CA | Hybrid / In-Person About the Company We are representing a high-growth AI research organization working at the intersection of...Full timeWork at officeRelocationVisa sponsorship
- ...Job Description Job Description We are Genmo, a research lab developing the world’s most sophisticated video world models to understand... ...research team in advancing the frontiers of visual generative AI. As a Research Engineer, you'll work alongside experienced researchers...Work at office
$110.7k - $379.2k
Position Summary Research Engineer — Post-Training & Small Language Models (SLMs), Healthcare AI Three hundred fifty million Americans rely on a healthcare system whose decision-making has become slow, costly, and adversarial — care delayed by prior authorization...Local areaVisa sponsorship- ...company based in San Francisco, California. The Role: As a Research Engineer - Agency and Reasoning , you will be a core contributor... ...low as possible We all enjoy what we do and love discussing AI Benefits and Perks: Comprehensive medical, dental, vision...Work at officeRelocation package
$140k - $250k
...Location: Remote (United States) Work Model: Remote Industry: AI training data infrastructure Compensation: $140K-$250K base,... ...This is the company's top hiring priority and a genuinely hard research problem. Because data flows through a decentralized marketplace,...Full timeWork at officeRemote work- ...company based in San Francisco, California. The Role: As a Research Engineer - Audio & Speech Models , you will be a core contributor... ...low as possible We all enjoy what we do and love discussing AI Benefits and Perks: Comprehensive medical, dental,...Work at officeRelocation package
$150k - $180k
...measure, improve, and scale training data for frontier AI agents. You will sit at the intersection of research and engineering, leading a team that defines what... ...environments. ~ Deep, research-oriented understanding of AI evals and post-training, beyond surface-level agent...Visa sponsorship- ...physics for an exciting ongoing project. Successful candidates will contribute to high-quality data creation that informs the future of AI innovation. Experts are expected to complete 4-6 tasks per week and must have exceptional communication skills and rigorous physics...
$318.1k - $363.1k
...Overview Senior Director, Applied Research Overview: At Capital One, we are creating trustworthy and reliable AI systems, changing banking for good. For years, Capital One has been leading the industry in using machine learning to create real-time, intelligent...Full timePart timeLocal areaFlexible hours- ...create reliable, interpretable, and steerable AI systems. We want AI to be safe and... ...is a quickly growing group of committed researchers, engineers, policy experts, and business... ...adapted Epoch's Capabilities Index to our evals); all the data in When AI Builds Itself comes...Full time
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Director, Research - AI Evals. Be the first to apply!
- research manager San Francisco, CA
- senior research manager San Francisco, CA
- director institutional research San Francisco, CA
- clinical research manager remote San Francisco, CA
- director of research San Francisco, CA
- research data coordinator San Francisco, CA
- research supervisor San Francisco, CA
- account manager market research San Francisco, CA
- research coordinator San Francisco, CA
- associate director clinical research San Francisco, CA


