Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Red Teamer (LLM Generalist)

$66.56k - $197.6k
Full-time

Handshake

AI Red Teamer (LLM Generalist)

Location: Seattle, WA (candidates must reside in the Seattle metro area or be willing to relocate prior to start)
Type: Contract, 40 hours per week

About the Role

As an AI Red Teamer, you will stress-test large language models by intentionally trying to break them. Rather than checking whether an answer is correct, you will design creative, adversarial prompts that expose vulnerabilities: unsafe content, bias, broken guardrails, hallucinations, prompt injection weaknesses, and unexpected behaviors. Your work directly supports AI safety and model robustness for leading research labs.

This is a generalist red teaming role. You will probe models across the full spectrum of risk categories, including content safety, CBRN (chemical, biological, radiological, nuclear), cybersecurity, persuasion and influence operations, child safety, self-harm, over-companionship, and regulatory compliance. Red teaming may span text, image, voice, and agentic model capabilities depending on project needs.

This role requires creativity, curiosity, and an ability to think like an adversary while operating with strong ethical judgment.

Day-to-Day Responsibilities

  • Craft creative prompts and multi-turn scenarios to stress-test AI guardrails across diverse risk categories

  • Discover ways around safety filters, restrictions, and defenses using jailbreak, evasion, and prompt injection techniques

  • Explore edge cases to provoke disallowed, harmful, or incorrect outputs

  • Evaluate and score model responses against structured harm taxonomies and severity rubrics

  • Document experiments clearly, including what you tried, why you tried it, and what it revealed

  • Review and refine adversarial prompts generated by other team members

  • Contribute to harm taxonomy development, calibration exercises, and inter-rater reliability work

  • Collaborate with engineers, data scientists, and researchers to share findings and strengthen defenses

  • Work with potentially disturbing content on a regular basis (see Content Warning below)

  • Stay current on jailbreaks, attack methods, and evolving model behaviors

Desired Capabilities

Core

  • Strong hands-on experience using multiple LLMs (ChatGPT, Claude, Gemini, open-source models, etc.)

  • Intuition for crafting adversarial prompts; familiarity with jailbreak or evasion techniques is a strong plus

  • Creative, adversarial problem-solving skills

  • Clear and thoughtful written communication

  • Strong ethical judgment and the ability to separate adversarial thinking from personal values

  • Self-directed, collaborative, and comfortable in feedback-heavy environments

  • Curiosity, persistence, and comfort with frequent failure in experimentation

Nice to Have

  • Familiarity with Python or other scripting languages

  • Experience working with LLM APIs or evaluation tooling

  • Comfort with structured data annotation and rubric-based scoring

  • Prior work in trust and safety, content moderation, QA, or security research

  • Subject matter expertise in any high-risk domain (cybersecurity, chemistry, biology, medicine, law, finance, etc.)

You Will Thrive Here If

  • You treat every model response as a hypothesis to challenge

  • You can switch between creative free-association and rigorous documentation in the same session

  • You go deep into unusual interests (fandoms, niche internet cultures, gaming exploits, Wikipedia rabbit holes, etc.)

  • You come from a creative background: writing, visual art, improv, puzzle design, or similar

  • You are energized by finding the thing nobody else thought to try

  • You are genuinely passionate about AI and follow the space closely

Content Warning

This role involves regular and deliberate exposure to harmful content. You will encounter and intentionally generate content involving violence, self-harm, hate speech, sexually explicit material, child safety scenarios, and other categories of harmful output as part of structured adversarial testing. Candidates must be able to engage with this material professionally and sustainably. Support resources are available.

About Handshake AI

Handshake AI partners with leading AI research labs to make models safer and more robust. Our red teaming operations help identify vulnerabilities before they reach users, contributing directly to the responsible development of frontier AI systems.

Vacancy posted 6 days ago
Similar jobs that could be interesting for youBased on the AI Red Teamer (LLM Generalist) in Seattle, WA vacancy
  •  ...job is to support the following businesses with the most advanced AI technology:- Combat any kinds of risks/violations issues in E-...  ...basic algorithms such as NLP, vision, multimodal, search, graph, LLM, etc. to provide support for governance business, explore cutting... 
    Suggested
    Overseas

    TikTok

    Seattle, WA
    5 days ago
  • $184k - $287.5k

    Join our team at NVIDIA and help bring AI solutions to our largest customers. We are seeking an expert Solutions Architect to assist...  ...understanding performance aspects related to tasks like large scale LLM training and inference.Conducting regular technical customer... 
    Suggested
    Full time

    Nvidia

    Seattle, WA
    5 days ago
  •  ...era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity...  ...ML/AI workloads inside Snowflake. Cortex Training is our LLM post-training platform: it turns scarce, expensive GPU capacity... 
    Suggested

    Snowflake

    Bellevue, WA
    1 day ago
  • $168.1k - $227.4k

     ...shopping, buying, and pre and post purchase experience.This role is at the forefront of Generative AI innovation within the Sub-Same Day organization. As part of the LLM Foundations team, you will design, build, and scale AI-powered tools that leverage Large Language Models... 
    Suggested
    Internship
    Flexible hours

    Amazon

    Bellevue, WA
    4 days ago
  •  ...knowledge integration- Design prompt engineering and reasoning workflows that connect structured features, risk indicators, and real-time LLM-based decisions.- Knowledge Distillation and BERT-style architectures- Build agentic workflows for complex cases, including modular... 
    Suggested
    Worldwide

    TikTok

    Seattle, WA
    5 days ago
  •  ...Experience with Prometheus/OpenTelemetry, graph databases (e.g., Neo4j), and developing alert and event platforms.- A Passion for AIOps/ML/LLM Practices:- A keen interest in the latest advancements in Large Models and Agent technologies, with thoughtful insights or hands-on... 

    TikTok

    Seattle, WA
    5 days ago
  • $136k - $184k

     ...the right methods (statistical, causal, ML, LLM, hybrid) for each problem and justify trade...  ...compliancy (Shepherd risk, App Security red-certification, Kale, Legal, Threat Models,...  ...experience- 1+ years of working with or evaluating AI systems experience- 1+ years of creating or... 
    Flexible hours

    Amazon

    Bellevue, WA
    4 days ago
  • $193.4k - $420.2k

     ...that wants you to grow and succeed. The context engine that makes AI enterprise ready. Anyone can build an AI agent. What makes SAP's...  ...production. Develop AI capabilities including generative AI and LLM-based solutions using enterprise business data, knowledge graphs,... 
    Permanent employment
    Full time
    Worldwide
    Flexible hours

    SAP

    Bellevue, WA
    2 days ago
  •  ...The Opportunity   We are building a dedicated AI Red Team to rigorously test and harden enterprise-scale AI products. We are looking...  .... This role focuses on identifying vulnerabilities in LLM-driven systems, breaking model guardrails, exploiting data pathways... 
    Full time

    C-serv

    Seattle, WA
    1 day ago
  • $236k - $330k

     ...era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity...  ...Snowflake AI Research team and advance the state of the art in LLM inference systems and optimization.Our mission is to build the... 
    Shift work

    Snowflake

    Bellevue, WA
    4 days ago
  •  ...-efficient tuning, RAG architectures, vector-store technologies, LLM evaluationExceptional time management to meet your responsibilities...  ...client needs, develop impactful advanced analytics and AI solutions, optimize code, and solve complex business challenges across... 
    Apprenticeship
    Work at office
    Local area
    Easy work

    McKinsey & Company

    Seattle, WA
    3 days ago
  • $293.5k

     ...Information Job Title Expert Senior Manager, AI Engineering Job ID 104335 Work...  ...design:Implement guardrails, fallbacks, red-teaming strategies, and human-in-the-loop (...  ..., and data labeling strategies for LLM appsDeep experience with:Advanced RAG architectures... 
    Permanent employment
    Full time
    Apprenticeship
    Work at office
    Local area
    Work from home
    Home office
    3 days per week

    Bain & Company

    Seattle, WA
    4 days ago
  • $110.7k - $379.2k

     ...& Small Language Models (SLMs), Healthcare AI Three hundred fifty million Americans rely...  ...including frameworks such as vLLM, TensorRT-LLM, or TGI. Small language models & open-weight...  ...-per-dollar. Evaluation, safety & red teaming • Design evaluation frameworks covering... 
    Local area
    Visa sponsorship

    Deloitte

    Seattle, WA
    2 days ago
  •  ...that wants you to grow and succeed. The context engine that makes AI enterprise ready.Anyone can build an AI agent. What makes SAP's...  ...scale problems.Develop AI capabilities — including generative AI and LLM-based solutions — using enterprise business data, knowledge... 
    Permanent employment
    Full time
    Local area
    Worldwide
    Flexible hours

    SAP

    Bellevue, WA
    4 days ago
  • $110k - $220k

     ...provide technical leadership and long‑term vision for next‑generation AI systems and platforms. This senior individual contributor will...  ...multiple autonomous AI agents (not just experimented with LLM APIs) - You’re Level 5+ on Steve Yegge’s Vibe Coding scale and can... 
    Full time
    Temporary work
    Part time
    Home office

    Walmart

    Bellevue, WA
    1 day ago
  • $203.5k

    General Information Job Title Lead, AI Engineering Job ID 102641 Work Areas...  ...design:Implement guardrails, fallbacks, red-teaming strategies, and human-in-the-loop (...  ...frameworks, and data labeling strategies for LLM applicationsExperience with:RAG architectures... 
    Permanent employment
    Full time
    Apprenticeship
    Work at office
    Local area
    Work from home
    Home office
    3 days per week

    Bain & Company

    Seattle, WA
    3 days ago
  • $184k - $287.5k

     ...with strategic customers to implement and enhance groundbreaking AI workloads. We partner with the world's most innovative AI companies...  ...and external customers’ AI and ML initiatives, including LLM performance evaluation and supporting new hardware in open-source... 
    Full time
    Remote work

    Nvidia

    Seattle, WA
    2 days ago
  • $166k - $258k

     ...team covering a large surface area, and AI is how we do it.This is a player-coach role...  ..., no story points. Our engineers are generalists who work across the whole portfolio, pair...  ...practiceFamiliarity with Conversational Analytics, LLM-powered data tools, or semantic layer... 
    Full time

    Nordstrom

    Seattle, WA
    1 day ago
  • $143.7k - $194.4k

     ...testing, and operational excellence- Knowledge of Machine Learning and LLM fundamentals, including transformer architecture, training/...  ...developing and deploying LLMs in production on GPUs, Neuron, TPU or other AI acceleration hardwareAmazon is an equal opportunity employer and... 
    Internship
    Flexible hours

    Amazon

    Seattle, WA
    4 days ago
  • $204.51k - $269.94k

     ...critical part of our product strategy and future growth.As Director of AI, you will lead the team that builds and ships iSpot's AI-powered...  ...workflows in production.Familiarity with modern MLOps practices, LLM observability, and cloud-native AI platforms.Experience with... 
    Full time
    Part time
    Work experience placement
    Work at office
    Local area
    Remote work
    Work from home
    Flexible hours
    Shift work
    3 days per week
    1 day per week

    iSpot.tv

    Bellevue, WA
    5 days ago
  • AI/ML Engineer**** Please note: This role is not eligible for 100% remote work. Employees must live within a commutable distance of a...  ...that solve business problems and improve how people work* Develop LLM-powered applications using RAG, agents, prompt engineering, tool... 
    Temporary work
    Work at office
    Local area
    3 days per week

    Slalom

    Seattle, WA
    1 day ago
  •  ...networking at massive scale.Responsibilities:- Design, implementation and deployment of high-speed network technologies to support AI/LLM applications.- Design and development of platforms/systems for monitoring, analysis and diagnosis of large scale AI/LLM network.- Research... 

    TikTok

    Seattle, WA
    5 days ago
  • $342.7k

     ...strategy for Splunk and Cisco, and we are building best-in-class AI capabilities into our platform, security and observability products...  ...record in Generative AI technology and Large Language Models (LLM).Demonstrated experience in leading impactful AI product development... 
    Full time
    Temporary work
    Work at office
    Local area
    Remote work
    Flexible hours

    CISCO Systems

    Seattle, WA
    1 day ago
  • $30 - $85 per hour

     ...million+ employers, and 1,600 educational institutions. Handshake AI works directly with frontier AI lab researchers to create...  ...human expertise. About the Role As an AI Model Policy Trainer, Generalist, you will turn complex customer policies into consistent, well-... 
    Hourly pay
    Full time
    Monday to Friday
    Flexible hours

    Handshake

    Seattle, WA
    9 days ago
  • $143.7k - $194.4k

    Do you want to build software systems powered by generative AI that serve hundreds of millions of customers? The OAS Offers Tech organization...  ...) of new and current systems- Knowledge of Machine Learning and LLM fundamentals, including transformer architecture, training/... 
    Temporary work
    Worldwide
    Flexible hours

    Amazon

    Seattle, WA
    4 days ago
  • $142.8k - $193.2k

     ...personalized framework for content and subscription optimization. As an AI/ML expert, you will partner directly with product owners to...  ...get to tackle in this role, such as optimizing/fine-tuning GenAI/LLM solutions for Prime personalization, building GenAI foundation... 
    Temporary work
    Flexible hours

    Amazon

    Seattle, WA
    1 day ago
  • $198.5k - $300k

     ...and Chewy revenue.This is a rare seat: you will define how modern AI—LLMs, generative AI, foundation models, and multimodal systems—...  ...balance.Define and execute Chewy Ads’ AI / GenAI strategy, including LLM-powered workstreams such as:Generative creative and offer... 
    Local area
    Flexible hours

    Chewy

    Bellevue, WA
    4 days ago
  • $189.72k - $332.01k

     ...best work. Creating a career you love? It’s Possible.At Pinterest, AI isn't just a feature, it's a powerful partner that augments our...  ...development, debugging, testing, and refactoringFamiliarity with LLM-powered productivity tools for documentation search, experiment analysis... 
    Local area
    Relocation package

    Pinterest

    Seattle, WA
    3 days ago
  • $163.42k - $285.98k

     ...best work. Creating a career you love? It’s Possible.At Pinterest, AI isn't just a feature, it's a powerful partner that augments our...  ...development, debugging, testing, and refactoringFamiliarity with LLM-powered productivity tools for documentation search, experiment analysis... 
    Local area
    Relocation package

    Pinterest

    Seattle, WA
    3 days ago
  •  ...positive impact in critical industries through AI transformation. We specialize in physics-...  ...real system.Define the agent seam: where LLM agents may assist (translation, hypothesis...  ...never assert around formal tooling.Strong generalist engineering skills: Python fluency,... 
    Full time
    Remote work
    Work visa

    AZX

    Seattle, WA
    5 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Red Teamer (LLM Generalist). Be the first to apply!