Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI Red Teamer (LLM Generalist)

Cacheflow

About the Role As an AI Red Teamer, you will stress-test large language models by intentionally trying to break them. Rather than checking whether an answer is correct, you will design creative, adversarial prompts that expose vulnerabilities: unsafe content, bias, broken guardrails, hallucinations, prompt injection weaknesses, and unexpected behaviors. Your work directly supports AI safety and model robustness for leading research labs. This is a generalist red teaming role. You will probe models across the full spectrum of risk categories, including content safety, CBRN (chemical, biological, radiological, nuclear), cybersecurity, persuasion and influence operations, child safety, self-harm, over-companionship, and regulatory compliance. Red teaming may span text, image, voice, and agentic model capabilities depending on project needs. This role requires creativity, curiosity, and an ability to think like an adversary while operating with strong ethical judgment. Responsibilities Craft creative prompts and multi-turn scenarios to stress-test AI guardrails across diverse risk categories Discover ways around safety filters, restrictions, and defenses using jailbreak, evasion, and prompt injection techniques Explore edge cases to provoke disallowed, harmful, or incorrect outputs Evaluate and score model responses against structured harm taxonomies and severity rubrics Document experiments clearly, including what you tried, why you tried it, and what it revealed Review and refine adversarial prompts generated by other team members Contribute to harm taxonomy development, calibration exercises, and inter-rater reliability work Collaborate with engineers, data scientists, and researchers to share findings and strengthen defenses Work with potentially disturbing content on a regular basis (see Content Warning below) Stay current on jailbreaks, attack methods, and evolving model behaviors Desired Capabilities Strong hands‑on experience using multiple LLMs (ChatGPT, Claude, Gemini, open‑source models, etc.) Intuition for crafting adversarial prompts; familiarity with jailbreak or evasion techniques is a strong plus Creative, adversarial problem‑solving skills Clear and thoughtful written communication Strong ethical judgment and the ability to separate adversarial thinking from personal values Self‑directed, collaborative, and comfortable in feedback‑heavy environments Curiosity, persistence, and comfort with frequent failure in experimentation Extra Credit Familiarity with Python or other scripting languages Experience working with LLM APIs or evaluation tooling Comfort with structured data annotation and rubric‑based scoring Prior work in trust and safety, content moderation, QA, or security research Subject matter expertise in any high‑risk domain (cybersecurity, chemistry, biology, medicine, law, finance, etc.) You Will Thrive Here If You treat every model response as a hypothesis to challenge You can switch between creative free‑association and rigorous documentation in the same session You go deep into unusual interests (fandoms, niche internet cultures, gaming exploits, Wikipedia rabbit holes, etc.) You come from a creative background: writing, visual art, improv, puzzle design, or similar You are energized by finding the thing nobody else thought to try You are genuinely passionate about AI and follow the space closely Content Warning This role involves regular and deliberate exposure to harmful content. You will encounter and intentionally generate content involving violence, self‑harm, hate speech, sexually explicit material, child safety scenarios, and other categories of harmful output as part of structured adversarial testing. Candidates must be able to engage with this material professionally and sustainably. Support resources are available. #J-18808-Ljbffr Cacheflow

Vacancy posted 4 days ago
Similar jobs that could be interesting for youBased on the AI Red Teamer (LLM Generalist) in Seattle, WA vacancy
  • Handshake is seeking an AI Red Teamer in Seattle, WA, to stress-test large language models by designing adversarial prompts. The role involves evaluating model outputs, documenting findings, and collaborating with a multidisciplinary team to strengthen AI safety. Ideal... 
    Suggested

    Handshake

    Seattle, WA
    3 days ago
  •  ...job is to support the following businesses with the most advanced AI technology:- Combat any kinds of risks/violations issues in E-...  ...basic algorithms such as NLP, vision, multimodal, search, graph, LLM, etc. to provide support for governance business, explore cutting... 
    Suggested
    Overseas

    TikTok

    Seattle, WA
    3 days ago
  • $184k - $287.5k

    Join our team at NVIDIA and help bring AI solutions to our largest customers. We are seeking an expert Solutions Architect to assist...  ...understanding performance aspects related to tasks like large scale LLM training and inference.Conducting regular technical customer... 
    Suggested
    Full time

    Nvidia

    Seattle, WA
    3 days ago
  • Agoda is seeking a Technical Product Manager for ML & LLM Platforms to drive AI transformation. You will define the product vision and support teams in productionizing ML models. The ideal candidate has over 5 years of experience in ML engineering, platform engineering,... 
    Suggested
    Worldwide
    Relocation

    Agoda

    Seattle, WA
    1 day ago
  •  ...era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity...  ...ML/AI workloads inside Snowflake. Cortex Training is our LLM post-training platform: it turns scarce, expensive GPU capacity... 
    Suggested

    Snowflake

    Bellevue, WA
    4 days ago
  •  ...LLM/Prompt-Context Engineer – Fullstack Python (AI Agents, LangGraph, Context Engineering) Location – 1st Atlanta, 2nd Dallas, 3rd Seattle (Onsite no remote). Onsite interview required We are looking for a highly skilled LLM/Prompt-Context Engineer with a strong... 
    Remote work

    Diversity Nexus

    Seattle, WA
    4 days ago
  •  ...knowledge integration- Design prompt engineering and reasoning workflows that connect structured features, risk indicators, and real-time LLM-based decisions.- Knowledge Distillation and BERT-style architectures- Build agentic workflows for complex cases, including modular... 
    Worldwide

    TikTok

    Seattle, WA
    3 days ago
  •  ...at the heart of our culture, fueling our curiosity and innovation. Technical Product Manager - ML & LLM Platforms The Opportunity Agoda is rapidly accelerating its AI transformation. Our ambition is to make ML and LLM application development fast, scalable, reliable,... 
    Worldwide
    Relocation
    Relocation package

    Agoda

    Seattle, WA
    1 day ago
  •  ...Experience with Prometheus/OpenTelemetry, graph databases (e.g., Neo4j), and developing alert and event platforms.- A Passion for AIOps/ML/LLM Practices:- A keen interest in the latest advancements in Large Models and Agent technologies, with thoughtful insights or hands-on... 

    TikTok

    Seattle, WA
    3 days ago
  • $230k - $280k

     ...CTEM). The HackerOne Platform unites agentic AI solutions with the ingenuity of the world’s...  ...disclosure, agentic pentesting, AI red teaming, and code security, HackerOne delivers...  ...memory systems, RAG, long-horizon tasks, and LLM-based models—into the HackerOne platform, applying... 
    Full time
    Apprenticeship
    Work at office
    Local area
    Remote work
    Flexible hours
    Shift work
    1 day per week

    Hackerone

    Seattle, WA
    20 hours ago
  •  ...building the future of trustworthy AI. Grounded in behavioral science...  ...work in areas like RL gyms, red teaming, and benchmarking, we...  ...experts, annotators, reviewers, red teamers, contractors, and quality...  ...rsum. Experience in AI safety, LLM evaluation, or trust & safety operations... 
    Contract work
    For contractors

    mpathic

    Seattle, WA
    4 days ago
  •  ...The Opportunity   We are building a dedicated AI Red Team to rigorously test and harden enterprise-scale AI products. We are looking...  .... This role focuses on identifying vulnerabilities in LLM-driven systems, breaking model guardrails, exploiting data pathways... 
    Full time

    C-serv

    Seattle, WA
    20 hours ago
  •  ...Senior Data Scientist The AI & Data Analytics division is seeking to hire a Senior Data Scientist. This...  ...hyperspectral imagery. Hands on experience with LLM/LVM/Foundation Model and Frontier AI evaluation, red teaming, uncertainty analysis, or safety control implementation... 
    Work experience placement
    Local area
    Remote work
    Flexible hours

    PNNL

    Seattle, WA
    1 day ago
  • $293.5k

     ...Information Job Title Expert Senior Manager, AI Engineering Job ID 104335 Work...  ...design:Implement guardrails, fallbacks, red-teaming strategies, and human-in-the-loop (...  ..., and data labeling strategies for LLM appsDeep experience with:Advanced RAG architectures... 
    Permanent employment
    Full time
    Apprenticeship
    Work at office
    Local area
    Work from home
    Home office
    3 days per week

    Bain & Company

    Seattle, WA
    2 days ago
  •  ...-efficient tuning, RAG architectures, vector-store technologies, LLM evaluationExceptional time management to meet your responsibilities...  ...client needs, develop impactful advanced analytics and AI solutions, optimize code, and solve complex business challenges across... 
    Apprenticeship
    Work at office
    Local area
    Easy work

    McKinsey & Company

    Seattle, WA
    1 day ago
  • $203.5k

    General Information Job Title Lead, AI Engineering Job ID 102641 Work Areas...  ...design:Implement guardrails, fallbacks, red-teaming strategies, and human-in-the-loop (...  ...frameworks, and data labeling strategies for LLM applicationsExperience with:RAG architectures... 
    Permanent employment
    Full time
    Apprenticeship
    Work at office
    Local area
    Work from home
    Home office
    3 days per week

    Bain & Company

    Seattle, WA
    1 day ago
  • $110.7k - $379.2k

     ...& Small Language Models (SLMs), Healthcare AI Three hundred fifty million Americans rely...  ...including frameworks such as vLLM, TensorRT-LLM, or TGI. Small language models & open-weight...  ...-per-dollar. Evaluation, safety & red teaming • Design evaluation frameworks covering... 
    Local area
    Visa sponsorship

    Deloitte

    Seattle, WA
    20 hours ago
  • $166k - $258k

     ...team covering a large surface area, and AI is how we do it.This is a player-coach role...  ..., no story points. Our engineers are generalists who work across the whole portfolio, pair...  ...practiceFamiliarity with Conversational Analytics, LLM-powered data tools, or semantic layer... 
    Full time

    Nordstrom

    Seattle, WA
    4 days ago
  • $169k - $338k

     ...provide technical leadership and long‑term vision for next‑generation AI systems and platforms. This senior individual contributor will...  ...multiple autonomous AI agents (not just experimented with LLM APIs) - You’re Level 5+ on Steve Yegge’s Vibe Coding scale and can... 
    Full time
    Temporary work
    Part time
    Home office

    Walmart

    Bellevue, WA
    4 days ago
  •  ...prioritize a diverse F5 community where each individual can thrive.AI Engineer — Customer Success & Services (F5) Location: Hybrid (San...  ...architectures, own model lifecycle and MLOps, implement safe RAG/LLM systems and observability, and partner closely with Product,... 
    Full time
    Local area

    F5 Networks

    Seattle, WA
    1 day ago
  • $109.5k - $190k

     ...Scientist II to join our growing team. In this role, you will help build AI-powered and machine learning solutions that improve customer and...  ....Practical knowledge of deep learning, neural networks, GenAI/LLM concepts, or AI system development.Experience working with SQL... 
    Local area
    Flexible hours

    Chewy

    Bellevue, WA
    3 days ago
  •  ...networking at massive scale.Responsibilities:- Design, implementation and deployment of high-speed network technologies to support AI/LLM applications.- Design and development of platforms/systems for monitoring, analysis and diagnosis of large scale AI/LLM network.- Research... 

    TikTok

    Seattle, WA
    3 days ago
  • $202.16k - $368.22k

    Senior Research Engineer / Scientist - Storage for LLM Location: Seattle Team: Infrastructure Employment Type: Regular Job Code: A1526...  ...research and engineering group focused on building next‑generation AI‑native data infrastructure. Positioned at the intersection of... 
    Temporary work
    Local area

    ByteDance

    Seattle, WA
    20 hours ago
  • $128k - $252.5k

     ..., and help us create the next generation of tools, products, and AI services. You will work closely with clients to understand their...  ...execution of complex AI data science solutions1+ year of experience with LLM/GenAI use cases and developing RAG solutions, tools, and services... 
    Local area

    Deloitte

    Seattle, WA
    20 hours ago
  • $143.7k - $194.4k

     ...testing, and operational excellence- Knowledge of Machine Learning and LLM fundamentals, including transformer architecture, training/...  ...developing and deploying LLMs in production on GPUs, Neuron, TPU or other AI acceleration hardwareAmazon is an equal opportunity employer and... 
    Internship
    Flexible hours

    Amazon

    Seattle, WA
    2 days ago
  • $140k - $180k

     ...AI Builder Artera is hiring an AI Builder to build the agentic systems that deliver exceptional patient experiences for our customers...  .... Experience writing evals and quality benchmarks for LLM-driven systems. $140,000 - $180,000 a year This is an exempt... 
    Temporary work
    Summer work
    Summer holiday
    Work at office
    Local area
    Relocation
    Flexible hours

    Artera Corporation

    Seattle, WA
    4 days ago
  • $148.5k - $260.1k

     ...CategorySoftware EngineeringJob DetailsAbout SalesforceSalesforce is the #1 AI CRM, where humans with agents drive customer success together....  ...cosine similarity) to decouple intent handling from expensive LLM calls.Developer Automation: Experience deploying and integrating... 
    Full time
    Contract work

    Salesforce

    Seattle, WA
    1 day ago
  • $342.7k

     ...strategy for Splunk and Cisco, and we are building best-in-class AI capabilities into our platform, security and observability products...  ...record in Generative AI technology and Large Language Models (LLM).Demonstrated experience in leading impactful AI product development... 
    Full time
    Temporary work
    Work at office
    Local area
    Remote work
    Flexible hours

    CISCO Systems

    Seattle, WA
    3 days ago
  • $166k - $203k

     ...CTEM). The HackerOne Platform unites agentic AI solutions with the ingenuity of the world’s...  ...disclosure, agentic pentesting, AI red teaming, and code security, HackerOne delivers...  ...building and integrating AI capabilities such as LLM-powered workflows, RAG pipelines, or... 
    Full time
    Apprenticeship
    Work at office
    Local area
    Remote work
    Flexible hours
    Shift work
    1 day per week

    Hackerone

    Seattle, WA
    20 hours ago
  • $123.5k - $185.3k

     ...thrive. Role Overview F5 is expanding its AI Center of Excellence and is hiring a Specialist...  ...deep expertise in AI, Data Science, and LLM behavior to support our AI Runtime Security...  ...outcomes from proofs of concept (POCs), red-teaming exercises, and runtime guardrail evaluations... 
    Local area

    F5

    Seattle, WA
    3 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI Red Teamer (LLM Generalist). Be the first to apply!