AI Red Teamer (LLM Generalist)
Cacheflow
About the Role As an AI Red Teamer, you will stress-test large language models by intentionally trying to break them. Rather than checking whether an answer is correct, you will design creative, adversarial prompts that expose vulnerabilities: unsafe content, bias, broken guardrails, hallucinations, prompt injection weaknesses, and unexpected behaviors. Your work directly supports AI safety and model robustness for leading research labs. This is a generalist red teaming role. You will probe models across the full spectrum of risk categories, including content safety, CBRN (chemical, biological, radiological, nuclear), cybersecurity, persuasion and influence operations, child safety, self-harm, over-companionship, and regulatory compliance. Red teaming may span text, image, voice, and agentic model capabilities depending on project needs. This role requires creativity, curiosity, and an ability to think like an adversary while operating with strong ethical judgment. Responsibilities Craft creative prompts and multi-turn scenarios to stress-test AI guardrails across diverse risk categories Discover ways around safety filters, restrictions, and defenses using jailbreak, evasion, and prompt injection techniques Explore edge cases to provoke disallowed, harmful, or incorrect outputs Evaluate and score model responses against structured harm taxonomies and severity rubrics Document experiments clearly, including what you tried, why you tried it, and what it revealed Review and refine adversarial prompts generated by other team members Contribute to harm taxonomy development, calibration exercises, and inter-rater reliability work Collaborate with engineers, data scientists, and researchers to share findings and strengthen defenses Work with potentially disturbing content on a regular basis (see Content Warning below) Stay current on jailbreaks, attack methods, and evolving model behaviors Desired Capabilities Strong hands‑on experience using multiple LLMs (ChatGPT, Claude, Gemini, open‑source models, etc.) Intuition for crafting adversarial prompts; familiarity with jailbreak or evasion techniques is a strong plus Creative, adversarial problem‑solving skills Clear and thoughtful written communication Strong ethical judgment and the ability to separate adversarial thinking from personal values Self‑directed, collaborative, and comfortable in feedback‑heavy environments Curiosity, persistence, and comfort with frequent failure in experimentation Extra Credit Familiarity with Python or other scripting languages Experience working with LLM APIs or evaluation tooling Comfort with structured data annotation and rubric‑based scoring Prior work in trust and safety, content moderation, QA, or security research Subject matter expertise in any high‑risk domain (cybersecurity, chemistry, biology, medicine, law, finance, etc.) You Will Thrive Here If You treat every model response as a hypothesis to challenge You can switch between creative free‑association and rigorous documentation in the same session You go deep into unusual interests (fandoms, niche internet cultures, gaming exploits, Wikipedia rabbit holes, etc.) You come from a creative background: writing, visual art, improv, puzzle design, or similar You are energized by finding the thing nobody else thought to try You are genuinely passionate about AI and follow the space closely Content Warning This role involves regular and deliberate exposure to harmful content. You will encounter and intentionally generate content involving violence, self‑harm, hate speech, sexually explicit material, child safety scenarios, and other categories of harmful output as part of structured adversarial testing. Candidates must be able to engage with this material professionally and sustainably. Support resources are available. #J-18808-Ljbffr Cacheflow
- Handshake is seeking an AI Red Teamer in Seattle, WA, to stress-test large language models by designing adversarial prompts. The role involves evaluating model outputs, documenting findings, and collaborating with a multidisciplinary team to strengthen AI safety. Ideal...Suggested
- ...job is to support the following businesses with the most advanced AI technology:- Combat any kinds of risks/violations issues in E-... ...basic algorithms such as NLP, vision, multimodal, search, graph, LLM, etc. to provide support for governance business, explore cutting...SuggestedOverseas
$184k - $287.5k
Join our team at NVIDIA and help bring AI solutions to our largest customers. We are seeking an expert Solutions Architect to assist... ...understanding performance aspects related to tasks like large scale LLM training and inference.Conducting regular technical customer...SuggestedFull time- Agoda is seeking a Technical Product Manager for ML & LLM Platforms to drive AI transformation. You will define the product vision and support teams in productionizing ML models. The ideal candidate has over 5 years of experience in ML engineering, platform engineering,...SuggestedWorldwideRelocation
- ...era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity... ...ML/AI workloads inside Snowflake. Cortex Training is our LLM post-training platform: it turns scarce, expensive GPU capacity...Suggested
- ...LLM/Prompt-Context Engineer – Fullstack Python (AI Agents, LangGraph, Context Engineering) Location – 1st Atlanta, 2nd Dallas, 3rd Seattle (Onsite no remote). Onsite interview required We are looking for a highly skilled LLM/Prompt-Context Engineer with a strong...Remote work
- ...knowledge integration- Design prompt engineering and reasoning workflows that connect structured features, risk indicators, and real-time LLM-based decisions.- Knowledge Distillation and BERT-style architectures- Build agentic workflows for complex cases, including modular...Worldwide
- ...at the heart of our culture, fueling our curiosity and innovation. Technical Product Manager - ML & LLM Platforms The Opportunity Agoda is rapidly accelerating its AI transformation. Our ambition is to make ML and LLM application development fast, scalable, reliable,...WorldwideRelocationRelocation package
- ...Experience with Prometheus/OpenTelemetry, graph databases (e.g., Neo4j), and developing alert and event platforms.- A Passion for AIOps/ML/LLM Practices:- A keen interest in the latest advancements in Large Models and Agent technologies, with thoughtful insights or hands-on...
$230k - $280k
...CTEM). The HackerOne Platform unites agentic AI solutions with the ingenuity of the world’s... ...disclosure, agentic pentesting, AI red teaming, and code security, HackerOne delivers... ...memory systems, RAG, long-horizon tasks, and LLM-based models—into the HackerOne platform, applying...Full timeApprenticeshipWork at officeLocal areaRemote workFlexible hoursShift work1 day per week- ...building the future of trustworthy AI. Grounded in behavioral science... ...work in areas like RL gyms, red teaming, and benchmarking, we... ...experts, annotators, reviewers, red teamers, contractors, and quality... ...rsum. Experience in AI safety, LLM evaluation, or trust & safety operations...Contract workFor contractors
- ...The Opportunity We are building a dedicated AI Red Team to rigorously test and harden enterprise-scale AI products. We are looking... .... This role focuses on identifying vulnerabilities in LLM-driven systems, breaking model guardrails, exploiting data pathways...Full time
- ...Senior Data Scientist The AI & Data Analytics division is seeking to hire a Senior Data Scientist. This... ...hyperspectral imagery. Hands on experience with LLM/LVM/Foundation Model and Frontier AI evaluation, red teaming, uncertainty analysis, or safety control implementation...Work experience placementLocal areaRemote workFlexible hours
$293.5k
...Information Job Title Expert Senior Manager, AI Engineering Job ID 104335 Work... ...design:Implement guardrails, fallbacks, red-teaming strategies, and human-in-the-loop (... ..., and data labeling strategies for LLM appsDeep experience with:Advanced RAG architectures...Permanent employmentFull timeApprenticeshipWork at officeLocal areaWork from homeHome office3 days per week- ...-efficient tuning, RAG architectures, vector-store technologies, LLM evaluationExceptional time management to meet your responsibilities... ...client needs, develop impactful advanced analytics and AI solutions, optimize code, and solve complex business challenges across...ApprenticeshipWork at officeLocal areaEasy work
$203.5k
General Information Job Title Lead, AI Engineering Job ID 102641 Work Areas... ...design:Implement guardrails, fallbacks, red-teaming strategies, and human-in-the-loop (... ...frameworks, and data labeling strategies for LLM applicationsExperience with:RAG architectures...Permanent employmentFull timeApprenticeshipWork at officeLocal areaWork from homeHome office3 days per week$110.7k - $379.2k
...& Small Language Models (SLMs), Healthcare AI Three hundred fifty million Americans rely... ...including frameworks such as vLLM, TensorRT-LLM, or TGI. Small language models & open-weight... ...-per-dollar. Evaluation, safety & red teaming • Design evaluation frameworks covering...Local areaVisa sponsorship$166k - $258k
...team covering a large surface area, and AI is how we do it.This is a player-coach role... ..., no story points. Our engineers are generalists who work across the whole portfolio, pair... ...practiceFamiliarity with Conversational Analytics, LLM-powered data tools, or semantic layer...Full time$169k - $338k
...provide technical leadership and long‑term vision for next‑generation AI systems and platforms. This senior individual contributor will... ...multiple autonomous AI agents (not just experimented with LLM APIs) - You’re Level 5+ on Steve Yegge’s Vibe Coding scale and can...Full timeTemporary workPart timeHome office- ...prioritize a diverse F5 community where each individual can thrive.AI Engineer — Customer Success & Services (F5) Location: Hybrid (San... ...architectures, own model lifecycle and MLOps, implement safe RAG/LLM systems and observability, and partner closely with Product,...Full timeLocal area
$109.5k - $190k
...Scientist II to join our growing team. In this role, you will help build AI-powered and machine learning solutions that improve customer and... ....Practical knowledge of deep learning, neural networks, GenAI/LLM concepts, or AI system development.Experience working with SQL...Local areaFlexible hours- ...networking at massive scale.Responsibilities:- Design, implementation and deployment of high-speed network technologies to support AI/LLM applications.- Design and development of platforms/systems for monitoring, analysis and diagnosis of large scale AI/LLM network.- Research...
$202.16k - $368.22k
Senior Research Engineer / Scientist - Storage for LLM Location: Seattle Team: Infrastructure Employment Type: Regular Job Code: A1526... ...research and engineering group focused on building next‑generation AI‑native data infrastructure. Positioned at the intersection of...Temporary workLocal area$128k - $252.5k
..., and help us create the next generation of tools, products, and AI services. You will work closely with clients to understand their... ...execution of complex AI data science solutions1+ year of experience with LLM/GenAI use cases and developing RAG solutions, tools, and services...Local area$143.7k - $194.4k
...testing, and operational excellence- Knowledge of Machine Learning and LLM fundamentals, including transformer architecture, training/... ...developing and deploying LLMs in production on GPUs, Neuron, TPU or other AI acceleration hardwareAmazon is an equal opportunity employer and...InternshipFlexible hours$140k - $180k
...AI Builder Artera is hiring an AI Builder to build the agentic systems that deliver exceptional patient experiences for our customers... .... Experience writing evals and quality benchmarks for LLM-driven systems. $140,000 - $180,000 a year This is an exempt...Temporary workSummer workSummer holidayWork at officeLocal areaRelocationFlexible hours$148.5k - $260.1k
...CategorySoftware EngineeringJob DetailsAbout SalesforceSalesforce is the #1 AI CRM, where humans with agents drive customer success together.... ...cosine similarity) to decouple intent handling from expensive LLM calls.Developer Automation: Experience deploying and integrating...Full timeContract work$342.7k
...strategy for Splunk and Cisco, and we are building best-in-class AI capabilities into our platform, security and observability products... ...record in Generative AI technology and Large Language Models (LLM).Demonstrated experience in leading impactful AI product development...Full timeTemporary workWork at officeLocal areaRemote workFlexible hours$166k - $203k
...CTEM). The HackerOne Platform unites agentic AI solutions with the ingenuity of the world’s... ...disclosure, agentic pentesting, AI red teaming, and code security, HackerOne delivers... ...building and integrating AI capabilities such as LLM-powered workflows, RAG pipelines, or...Full timeApprenticeshipWork at officeLocal areaRemote workFlexible hoursShift work1 day per week$123.5k - $185.3k
...thrive. Role Overview F5 is expanding its AI Center of Excellence and is hiring a Specialist... ...deep expertise in AI, Data Science, and LLM behavior to support our AI Runtime Security... ...outcomes from proofs of concept (POCs), red-teaming exercises, and runtime guardrail evaluations...Local area
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Red Teamer (LLM Generalist). Be the first to apply!


