Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Director, Research - AI Evals

$258k - $348k

Figma

Figma is growing our team of passionate creatives and builders on a mission to make design accessible to all. Figma’s platform helps teams bring ideas to life—whether you're brainstorming, creating a prototype, translating designs into code, or iterating with AI. From idea to product, Figma empowers teams to streamline workflows, move faster, and work together in real time from anywhere in the world. If you're excited to shape the future of design and collaboration, join us!The Figma Research team is hiring an Director, Research - AI Evals to own how we measure the quality of Figma's AI-powered experiences. As Figma ships more AI capabilities across our products, the question "is this really good?" has never mattered more — and answering it rigorously is what this role exists to do. You'll define what "good" means for our AI features, build the frameworks and quality bars to measure it, and turn that into trusted signal that product teams rely on to decide what to ship.The ideal candidate brings deep, hands-on experience evaluating AI/LLM-powered products — blending human evaluation with automated, model-based approaches — along with the product instinct and communication skills to make evaluation genuinely useful. Partnering with Product, Design, Engineering, and Data Science, you'll sit upstream of nearly every AI shipping decision at Figma and directly shape the quality of features used by millions of people.This is a full time role that can be held from one of our US hubs or remotely in the United States.What you'll do at Figma:Own AI evaluation methods and operations for Figma's AI-powered experiences — define quality dimensions, design how we measure them, and turn results into decision-ready signalBuild and maintain evaluation frameworks, rubrics, golden datasets, and quality bars, combining human evaluation with automated/model-based approaches (e.g., LLM-as-judge) where appropriatePartner with engineering to stand up repeatable, reproducible evaluation pipelines and regression testing, so evaluation is a routine part of how AI features are built and shippedProduce clear readouts and dashboards that let stakeholders confidently make go/no-go and prioritization decisionsSocialize a shared definition of quality so evaluation standards are adopted across teams rather than re-invented — and advocate for evaluation as a strategic partner in the product processManage a small team to execute our AI evals in partnership with contractors, internal staff, and/or LLMsWe'd love to hear from you if you have:10+ years of experience in product, research, applied research, or a closely related field, including 2+ years of management experienceDirect, hands-on experience owning the evaluation of AI/LLM-powered productsExpertise designing and running AI evaluation — human evaluation programs, rubric and benchmark/golden-dataset construction, inter-rater reliability — and sound judgment about when and how to apply automated/model-based approaches (e.g., LLM-as-judge), including their limitationsStrength across both qualitative and quantitative methods, comfort with data and metrics, and the ability to reason about model behaviorDemonstrated success in identifying the riskiest assumptions behind an ambiguous quality question, prioritizing them, and designing right-sized evaluation to build confidenceA proven track record of gaining buy-in from executive and cross-disciplinary stakeholders — transcending methodology to articulate a larger user story and the "so what" to inspire actionWhile it's not required, it's an added plus if you also have:Experience building or co-building automated evaluation pipelines and regression testing in partnership with engineering, or familiarity with eval tooling (e.g., Braintrust, LangSmith, DeepEval, or equivalents)Experience standing up a new function, practice, or discipline from scratch2+ years in product design, user-centric product management, data science, product development, and/or front-end engineeringA familiarity and depth of experience using Figma's productsAt Figma, one of our values is Grow as you go. We believe in hiring smart, curious people who are excited to learn and develop their skills. If you’re excited about this role but your past experience doesn’t align perfectly with the points outlined in the job description, we encourage you to apply anyways. You may be just the right candidate for this or other roles.Pay Transparency DisclosureJob level and actual compensation will be decided based on factors including, but not limited to, individual qualifications objectively assessed during the interview process (including skills and prior relevant experience, potential impact, and scope of role), market demands, and specific work location. Figma offers equity to employees, as well as a competitive package of additional benefits, including health, dental, and vision coverage; retirement benefits with company contributions; parental leave and reproductive or family planning support; mental health and wellness benefits; and paid time off. Figma provides paid sick leave, holidays, and other leave benefits in compliance with applicable federal, state, and local laws, including the requirements of the Washington Minimum Wage Act and related regulations. Exempt employees are eligible for employer‑provided paid flexible PTO in addition to flexible paid sick leave. PTO is subject to manager approval. Additional benefits may include company recharge days, cell phone and home internet reimbursements, and a number of lifestyle spending accounts. Figma also offers sales incentive compensation for most sales roles and an annual bonus plan for eligible non-sales roles. All compensation and benefits are subject to applicable plan terms and may be modified by Figma at any time, consistent with applicable law.Annual Base Salary Range:$258,000—$348,000 USDAt Figma we celebrate and support our differences. We know employing a team rich in diverse thoughts, experiences, and opinions allows our employees, our product and our community to flourish. Figma is an equal opportunity workplace - we are dedicated to equal employment opportunities regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity/expression, veteran status, or any other characteristic protected by law. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements.We will work to ensure individuals with disabilities are provided reasonable accommodation to apply for a role, participate in the interview process, perform essential job functions, and receive other benefits and privileges of employment. If you require accommodation, please reach out to View email address on click.appcast.io. These modifications enable an individual with a disability to have an equal opportunity not only to get a job, but successfully perform their job tasks to the same extent as people without disabilities. Examples of accommodations include but are not limited to: Holding interviews in an accessible locationEnabling closed captioning on video conferencingEnsuring all written communication be compatible with screen readersChanging the mode or format of interviews To ensure the integrity of our hiring process and facilitate a more personal connection, we require all candidates keep their cameras on during video interviews. Additionally, if hired you will be required to attend in person onboarding.By applying for this job, the candidate acknowledges and agrees that any personal data contained in their application or supporting materials will be processed in accordance with Figma's Candidate Privacy Notice.

Vacancy posted 2 days ago
Similar jobs that could be interesting for youBased on the Director, Research - AI Evals in San Francisco, CA vacancy
  • $320k

     ...that powers Lila's work in the life sciences. We are looking for a Research Engineer to own and build the core codebase that our models plug...  .... We believe science is the most inspiring frontier for AI. Rather than hard‑coding expert knowledge into tools, LILA builds... 
    Suggested
    Full time
    Work at office
    Local area
    Flexible hours
    Shift work

    Lilasciences

    San Francisco, CA
    3 days ago
  •  ...EPAM's new Frontier AI business unit partners directly with leading AI labs and advanced AI organizations, translating their research and post-training objectives into technically rigorous...  ..., and training data. The Senior Director, Frontier AI Research is EPAM's... 
    Suggested

    EPAM Systems

    San Francisco, CA
    3 days ago
  • Google in San Francisco seeks a Director of Research for Search Verticals to lead the user research and content design teams, guiding insights that shape the Google Search experience and its AI-forward direction. You will partner with product management, design and engineering... 
    Suggested

    Google

    San Francisco, CA
    2 days ago
  •  ...tuned Lila model variants for priority platform and commercial use cases. You will define data needs, gaps, evals, and release criteria, and partner with AI Research to ship capabilities. You will coordinate with research and app teams to ensure validatedFine-tunes flow... 
    Suggested

    Lila Sciences

    San Francisco, CA
    4 days ago
  • $150k - $250k

     ...engineering team building infrastructure for AI evaluation and reinforcement learning...  ...partner problems quickly. Coordinate with research and go-to-market teams to keep deployments...  ...~ Experience working on benchmarks and evals — with solid judgment about task realism,... 
    Suggested
    Full time
    Relocation
    Visa sponsorship

    Clera

    San Francisco, CA
    3 days ago
  • $150k - $250k

    About Distyl AI Distyl is an applied AI technology company partnering with the world’s...  ...goods, and global social organizations.We research and deploy technologies that power AI-native...  ..., reward modeling, synthetic data, evals, or related post-training techniquesStrong... 
    Work at office
    3 days per week

    Distyl AI

    San Francisco, CA
    22 hours ago
  • $365k

     ...create reliable, interpretable, and steerable AI systems. We want AI to be safe and...  ...is a quickly growing group of committed researchers, engineers, policy experts, and business...  ...move across research areas like compute, evals, RL environments, and emerging research initiatives... 
    Work at office
    Visa sponsorship
    Flexible hours
    Shift work

    Anthropic

    San Francisco, CA
    5 days ago
  •  ...Console is building the AI agents that power autonomous enterprises. We use AI to autonomously...  ...us. About the role We're hiring a Research Engineer to push Console's core agent loop...  ...production traffic into better agents, evals, and specialist models. This role sits... 
    Work at office

    Console Systems, Inc

    San Francisco, CA
    4 days ago
  • About Sable Sable built Aidan, the first AI employee who can lead customer calls using realtime voice, vision, and browser use. Aidan...  ...orchestration, browser/computer use, and applied ML such as multimodal evals Comfortable in a small team where you own problems end to end... 

    Sable AI, Inc

    San Francisco, CA
    2 days ago
  •  ...individuals, agents, enterprises, and even nation states. Our team of AI researchers and company builders come from DeepMind, OpenAI, Google Brain,...  .... This is a foundational role. Reflection is building model evals and safety from the ground up, and this RPM will be at the... 
    Full time
    Relocation package

    aijoblist

    San Francisco, CA
    3 days ago
  • $250k - $350k

     ...San Francisco, California, Turing is the world’s leading research accelerator for frontier AI labs and a trusted partner for global enterprises...  ...end-to-end the creation of datasets, RL environments, and evals for frontier AI labs in the domain of coding agents and... 
    For contractors

    Turing

    San Francisco, CA
    5 days ago
  • $150k - $250k

     ...engineering role at the intersection of applied research and customer deployment, sitting within a...  ...infrastructure for RL environments and AI evaluation. You'll own end-to-end...  ...~ Experience working on benchmarks and evals, with sound judgment about what makes a task... 
    Work at office
    Remote work
    Visa sponsorship

    Clera

    San Francisco, CA
    4 days ago
  • $160k - $250k

    Join to apply for the Founding Research Engineer role at Adam Join to apply for the Founding...  ...’re tackling a frontier problem: training AI models to intelligently interpret and edit...  ...vector representations of CAD features Design evals to measure geometric accuracy in 3D space... 
    Full time

    Adam

    San Francisco, CA
    2 days ago
  • Fundamental AI Research Institute Come join one of the only research institutions globally with resources to compete with top AI companies...  ...and benchmarks 2+ years industry experience, focused on RL, Evals, Reasoning, Agents or AI infra Comfortable navigating ambiguity... 

    Storm3

    San Francisco, CA
    3 days ago
  • $193k - $272k

     ...team empowers organizations to build deeper relationships with customers through innovative strategies, advanced analytics, Generative AI, transformative technologies, and creative design. We can enhance customer experiences and drive sustained growth and customer value... 
    Local area
    Visa sponsorship

    Deloitte

    San Francisco, CA
    4 days ago
  • $270k - $340k

    Principal AI Research Scientist, Research Director - AI ScalingP-1227About Databricks AIAt Databricks, we are obsessed with enabling data teams to solve the world’s toughest problems, from security threat detection to cancer drug development, by building and running the... 
    Local area
    Worldwide

    DataBricks

    San Francisco, CA
    1 day ago
  • $250k

    Our client, an AI-driven Healthcare company, are hiring a Head of AI Research to join their team in San Francisco. The successful candidate will work on building high-impact, complex AI systems across clinical intelligence, outcome modelling and advanced learning frameworks... 
    Full time

    Alldus International Consulting Ltd

    San Francisco, CA
    more than 2 months ago
  • A cutting-edge AI company in San Francisco is seeking a Head of Research to drive their research agenda in LLM efficiency. The ideal candidate will lead an applied research team, define the strategy for model routing and training, and work closely with engineering to translate... 

    Datawizz

    San Francisco, CA
    1 day ago
  • $400k

    Base pay range $400,000.00/yr - $400,000.00/yr About the Role We are seeking an exceptional Head of Research to lead and scale a world‑class AI research organization focused on advancing the state of the art in machine learning, large language models, and applied AI systems... 
    Full time

    Stealth AI Startup

    San Francisco, CA
    1 day ago
  • $250k - $400k

     ...San Francisco, California, Turing is the world’s leading research accelerator for frontier AI labs and a trusted partner for global enterprises...  ...at AI labs, translate those needs into environment goals Evals & Post-training: Demonstrate proof of value for your environments... 

    Turing

    San Francisco, CA
    5 days ago
  • Mercor is looking for First-Line Supervisors of Police and Detectives to assist in AI research projects remotely from San Francisco. Candidates should have over 4 years of experience and excellent written communication skills. This role involves creating deliverables, reviewing... 
    Remote job
    Contract work
    Flexible hours

    Mercor

    San Francisco, CA
    1 day ago
  •  ..., harnessing, and deployment—and connect that research to the patients, clinicians, and real‑world outcomes...  ...learning (RL) / post‑training, or evals; and researchers with real depth in developing frontier biomedical AI capabilities. Prior experience in healthcare is... 
    Work at office
    Relocation package

    Triwill Group

    San Francisco, CA
    1 day ago
  • $135.05k - $202.59k

     ...implementation planning, and stakeholder engagement.Familiarity with AI tools and emerging technologies, including practical...  ...business plan development, including qualitative and quantitative research and analysis.Well-developed communication and interpersonal skills... 
    Full time
    Work at office
    Shift work

    Sutter Health

    San Francisco, CA
    5 days ago
  • $148.1k - $282.1k

    The OpportunityThe Research and AI team builds foundational generative AI models and applications for Adobe products, enabling customers to ideate, build, and scale content in new ways.The role of Principal PM involves defining and advancing foundational generative capabilities... 
    Full time
    Temporary work
    Local area
    Worldwide

    Adobe Systems

    San Francisco, CA
    5 days ago
  • $210k

     ...About AfterQuery AfterQuery is an applied research lab curating data solutions for foundation model development. We serve every frontier AI lab with the mission of delivering the best data to power the best models. In doing so, we can make expertise that once took... 
    Full time
    Local area

    AfterQuery

    San Francisco, CA
    3 days ago
  •  ...driven transformation in medicine. Driven by the push for streamlined drug development, the market for advanced analytics and AI in clinical research is expanding exponentially. The PReDiCTR-TB Consortium is not just following industry standards—we are creating and leading... 
    Traineeship
    Worldwide

    UCSF Health

    San Francisco, CA
    4 days ago
  • $340k - $425k

    A leading AI research organization is seeking a Manager for its Interpretability team in San Francisco. The ideal candidate will have a strong background in managing technical teams and a passion for AI safety research. This role involves overseeing project execution, supporting... 
    Work at office
    Flexible hours

    Jobleads-US

    San Francisco, CA
    2 days ago
  •  ...Involves gathering, analyzing, and interpreting a wide variety of research data. Designs and conducts research including selecting data...  ...streamlined drug development, the market for advanced analytics and AI in clinical research is expanding exponentially. What You'll... 
    Traineeship
    Worldwide

    University of California , San Francisco

    San Francisco, CA
    4 days ago
  • Anthropic is seeking a Product Manager for the Research team to own ideation and deployment of frontier AI models and products, collaborating with research to productize applied work and identify high-potential use cases grounded in customer needs. You will translate complex... 

    Anthropic

    San Francisco, CA
    3 days ago
  • $175k

    Thinking Machines Lab is seeking a Research Product Manager based in San Francisco, California, to drive complex technical products and programs. You will coordinate large-scale research products, translate technical ideas into actionable plans, and collaborate across... 

    Thinking Machines Lab

    San Francisco, CA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Director, Research - AI Evals. Be the first to apply!