Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Adversarial AI Safety Engineer (Remote)

Obsidian

Mercor is building a remote AI red team to stress test conversational models and surface actionable vulnerabilities. The role emphasizes human data generation, structured evaluation, and clear reporting of risk. You will work across projects with clear guidelines and wellness resources in a safety-focused environment. You will review outputs on sensitive topics and contribute reproducible artifacts to improve safety across customer deployments. #J-18808-Ljbffr Obsidian

Vacancy posted 3 days ago
Similar jobs that could be interesting for youBased on the Adversarial AI Safety Engineer (Remote) in San Francisco, CA vacancy
  • Mercor is assembling a remote red team to test and strengthen AI systems through adversarial inputs. You will annotate failures, classify vulnerabilities, and surface systemic risks using rigorous taxonomies and playbooks. The role rewards clear communication of risks to... 
    Remote job

    Neon

    New York, NY
    1 day ago
  • A biosecurity innovation firm seeks a Software Engineer to enhance evaluations of frontier AI systems, focusing on security and misuse risk. In this remote role, you'll manage evaluations for new AI models, ensuring thorough analysis and collaboration with research scientists... 
    Remote work

    SecureBio, LLC

    United States
    2 days ago
  • $320k

     ...NVIDIA is seeking a Distinguished Engineer to serve as the founding technical leader for our AI Safety & Security Engineering team. Rooted in the firm belief that open...  ....SummaryLocation: US, CA, Santa Clara; US, NC, Remote; US, NY, Remote; US, TN, Remote; US, FL, Remote;... 
    Remote work
    Full time

    Nvidia

    North Carolina
    2 days ago
  • $65 per hour

    A leading AI consulting firm is seeking an AI Tutor specialized in Coding. This part-time freelance role involves evaluating AI models...  ...security. With flexible hours, this position allows you to work remotely on challenging AI projects that enhance your expertise.... 
    Remote job
    Part time
    Freelance
    Flexible hours

    Mindrift

    San Antonio, TX
    1 day ago
  • $152k - $241.5k

     ...weight models are foundational to American AI leadership and cybersecurity, and that...  ...and broad scientific scrutiny. Our AI Safety & Security Engineering team builds and evaluates AI-powered...  ...: US, CA, Santa Clara; US, NC, Remote; US, NY, Remote; US, TN, Remote; US, FL... 
    Remote work
    Full time

    Nvidia

    Santa Clara, CA
    1 day ago
  • Cohere, a leading security-first enterprise AI company, is seeking a Member of Technical Staff in the Safety for Agents team to advance safer LLMs through data generation...  ...autonomy and impact across global offices, while remote-friendly policies support flexible work... 
    Remote work
    Flexible hours

    Cohere

    New York, NY
    1 day ago
  • $145.9k - $234.2k

     ...execute, and evolve advanced adversarial simulation campaigns to test...  ...environments with a focus on AI and the attack surface it exposes...  ...Intelligence, and Detection Engineering to convert findings into...  ...SummaryLocation: Cambridge, Massachusetts; Remote - US; DigitalType: Full time
    Remote work
    Permanent employment
    Full time
    Work at office
    Work from home

    Moderna Therapeutics

    Cambridge, MA
    1 day ago
  •  ...We are searching for an AI Safety Specialist who will play a crucial role in enhancing...  ...deployment of AI systems by conducting adversarial testing, implementing protective measures...  ...Qualifications Background in cybersecurity, prompt engineering, or adversarial ML. Experience with... 

    Hyphen Connect Limited

    San Francisco, CA
    2 days ago
  •  ...are currently seeking a Gen AI Engineer to join our team in a hybrid...  ...Engineer robust guardrails for safety, compliance, and least-privilege...  ...: Build validator models, adversarial prompts, and policy checks into...  .... While many positions offer remote or hybrid work options, these... 
    Remote work
    Work at office
    Flexible hours
    3 days per week

    NTT DATA

    New York, NY
    3 days ago
  • $86.8k - $198k

    Operational Technology AI EngineerThe Opportunity: As a cyber warfare engineer, you know how critical...  ...vulnerabilities before adversaries can. At Booz Allen, you...  ...designed around the unique safety, availability, latency,...  ...on during meetings.Remote: If this position is listed... 
    Remote work
    Full time
    Contract work
    Part time
    Work at office
    Local area

    Booz Allen Hamilton

    Chantilly, Loudoun County, VA
    6 hours ago
  • $163.2k - $220.8k

     ...Wilson Sonsini is looking for a Senior AI Security Engineer to join the Security Operations team....  ...red teaming capabilities — developing adversarial test suites for prompt injection, jailbreaking...  ...York; Boulder; Los Angeles; Virtual / Remote; Seattle; Delaware; Washington, D.C.;... 
    Remote work
    Full time
    Work experience placement
    Worldwide
    Shift work

    Wilson Sonsini Goodrich & Rosati

    San Francisco, CA
    3 days ago
  • $147k - $184.8k

     ...location.Position Summary:The AI Security Engineer will be a key member of the...  ..., model integrity, and adversarial attack mitigation.Secure data...  ...a cloud environment.Travel:Remote employees will need to be...  ...personal information and online safety are a top priority for us.... 
    Remote work
    Contract work
    Work at office

    Structure Therapeutics

    South San Francisco, CA
    4 days ago
  • Role Description We are seeking an AI/LLM Safety Engineer to join our AI team and take ownership of how safely our models and agents behave...  ...Teaming ~Design and maintain a safety evaluation framework—adversarial prompt sets, scenario-based test suites, and regression... 
    Full time

    Propio

    Remote
    a month ago
  •  ...AI Red Team Engineer We're looking for an AI Red Team Engineer to break LLM-powered systems responsibly...  ...conversations. You'll own hands-on adversarial testing end to end: find the failure,...  ...write it up. White Circle is an AI Safety company building the safety,... 
    Remote work
    Local area

    Pumpkin Intelligence, Inc.

    United States
    3 days ago
  • $100k - $120k

     ...AI Cybersecurity Engineer – Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI...  ...security architecture review processes. Familiarity with adversarial ML, prompt injection, and model abuse research.... 
    Remote work
    Full time
    H1b
    Local area
    Immediate start
    Visa sponsorship

    Bright Vision Technologies

    United States
    2 days ago
  • Anthropic’s Safeguards team is seeking a Red Team Engineer to help ensure the safety of our deployed AI systems and products. You will take an adversarial approach to uncover vulnerabilities across our product ecosystem before they can be exploited by malicious actors.... 

    Anthropic

    San Francisco, CA
    4 days ago
  • $135.7k - $251.9k

     ...development code, integrating autonomy, AI or machine learning algorithms...  ...a motivated and experienced engineer with a strong technical...  ...who enjoy thinking like the adversary, breaking complex systems to make...  ...a Red Team leaderFlexible remote work, collaborative team culture... 
    Remote work
    Full time
    Temporary work
    Work experience placement
    Casual work
    Flexible hours

    Lockheed Martin

    Shelton, CT
    3 days ago
  • $102k - $170k

     ...Secret What You Will Do: Implement AI strategy to increase adoption of AI...  ...Collaborate with data scientists, platform engineers, and product teams to iterate on use cases...  ...of AI security concepts, including adversarial robustness, model access controls, and data... 
    Remote work
    Temporary work
    Flexible hours

    Guidehouse

    United States
    3 days ago
  • $190k - $250k

     ...comfort. About this role As a Senior AI Engineer at Atria, you will own the design,...  ...every model and prompt change. Accuracy, safety, cost, and latency are first-class...  ...considerations for LLM applications: guardrails, adversarial testing, and the failure modes that... 
    Remote work
    Summer work
    Flexible hours

    Atria Health and Research Institute

    United States
    4 days ago
  • UL Research Institutes, within the Digital Safety Research Institute (DSRI), is seeking an AI Research Engineer for a remote role. You will study incident reports related to LLMs and NLP, and contribute to safety assessments and publications in peer-reviewed venues. You... 
    Remote job

    Lucas James Talent Partners

    Evanston, IL
    3 days ago
  • Reflection is seeking a highly skilled safety researcher to own red-teaming and adversarial evaluation for our open-weight models. You will probe for failure modes...  ...as a gatekeeper for model releases, balancing risk with ambitious AI capabilities. #J-18808-Ljbffr Visa Hunt

    Visa Hunt

    San Francisco, CA
    3 days ago
  • B Capital in San Francisco is looking for individuals passionate about AI safety to own the red-teaming and adversarial evaluation pipeline for their models. The role involves collaborating with the Alignment team and ensuring all releases meet safety thresholds. Candidates... 

    B Capital

    San Francisco, CA
    2 days ago
  • $145.5k - $249.5k

     ...leadership role within the Optum Tech AI Transformation team,...  ...leadership for a squad of AI/ML engineers, driving strategic...  ...enjoy the flexibility to work remotely * from anywhere within the U....  ...explainability like SHAP/LIME, and adversarial red-teaming)Knowledge of Graph... 
    Remote work
    Minimum wage
    Full time
    Work experience placement
    Local area

    UnitedHealth Group

    Eden Prairie, MN
    11 hours ago
  •  ...AI Evaluation Engineer Location: USA Remote Employment Type: Contract Position Summary We are...  ...meet production quality and safety standards. Key Responsibilities...  ...suites. AI red-teaming or adversarial testing experience. Experience... 
    Remote work
    Contract work

    Acunor Inc

    New Jersey
    3 days ago
  •  ...Senior Translational Data And AI Engineer We are seeking a contract Senior Translational Data and AI Engineer to help modernize how...  ...assisted workflows that make future work faster. Participate in adversarial design and code reviews, identifying edge cases and pushing... 
    Remote work
    Contract work
    For contractors
    Local area

    Kaztronix

    United States
    1 day ago
  •  ...Member of Technical Staff to join our Applied AI team in San Francisco. You will deploy...  ...partners, focusing on reliability, safety, and performance. You will design scalable...  ...applications, collaborate with researchers and engineers, and build evaluation pipelines and... 
    Remote job

    AI Breaking Wire

    San Francisco, CA
    2 days ago
  • $150k - $170k

     ...development company delivering cloud, AI, data, and enterprise solutions...  ...potential. Job Title: Secure AI Systems Engineer Location: 100% Remote (U.S.) Position Type: Full-time,...  ...processes. Familiarity with adversarial ML, prompt injection, and model abuse... 
    Remote work
    Full time
    H1b
    Local area
    Immediate start
    Visa sponsorship

    Bright Vision Technologies

    United States
    4 days ago
  • $114.6k - $252.1k

    Job Title: Principal AI/ML Engineer (Large Language Model)Job Category:...  ...variety of applications within remote sensing such as tasking...  ...algorithms such as BERTGenerative Adversarial Networks and Variational...  ...higher purpose - to ensure the safety of our nation.An environment... 
    Remote work
    Contract work
    Work experience placement
    Local area
    Flexible hours

    CACI International

    Aurora, CO
    6 hours ago
  • $117.2k - $175.8k

    Sr Data Engineer - GE07BEWe’re determined to make a difference and...  ...seeks energetic and passionate AI Platform Engineers to build...  ...This role can have a Hybrid or Remote work arrangement.  Candidates...  ...Guardrails, responsible AI, adversarial attack mitigation, and red teaming... 
    Remote work
    Full time
    Temporary work
    Work at office
    3 days per week

    The Hartford Financial Services Group

    Chicago, IL
    11 hours ago
  • $180k

    AI Engineer & Researcher - Safety & Societal Impact AI Engineer & Researcher - Safety & Societal Impact xAI’s mission is to create AI systems that can...  ...(ML) Software Engineer, Python - AI Training (Freelance, Remote) San Francisco, CA $130,000 - $160,000 4 months ago San... 
    Remote work
    Full time
    Freelance
    Local area
    Relocation

    xAI

    San Francisco, CA
    4 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Adversarial AI Safety Engineer (Remote). Be the first to apply!