AI/LLM Safety Engineer
Propio
Role Description
We are seeking an AI/LLM Safety Engineer to join our AI team and take ownership of how safely our models and agents behave in production; with a focus on AI Safety, Trust & Safety, and Responsible AI. You will design the evaluations that catch unsafe behavior, build the guardrails that stop it, and lead the red-teaming that finds the gaps before our users—or attackers—do. Agent safety is the primary focus of this role: you will help ensure that as our systems gain the ability to call tools and take actions, they do so within well-defined, well-tested boundaries.
Key Responsibilities
- LLM Safety Evaluation & Red Teaming
- Design and maintain a safety evaluation framework—adversarial prompt sets, scenario-based test suites, and regression suites—so that every model and agent update is validated before it ships.
- Lead structured red-teaming exercises covering jailbreaks, prompt injection, tool misuse, and data exfiltration; document findings and drive each issue through to remediation and closure.
- Guardrails & Runtime Controls
- Build and iterate on guardrail logic, including input/output filtering, tool-boundary constraints, action validation, sensitive-data redaction, and policy prompting.
- Integrate safety checks into CI/CD and runtime so that unsafe behavior is intercepted before it reaches users.
- Agent Safety (primary focus of this role)
- Perform threat modeling for agentic scenarios: tool-call boundaries, sandbox isolation, and least-privilege access, with particular attention to preventing agents from exfiltrating data or executing irreversible actions through chained tool calls.
- Conduct safety reviews of reinforcement-learning (RL) environments and trajectory data, partnering with environment and agent engineering teams to embed safety constraints directly into the environments themselves.
- Monitoring & Observability
- Instrument AI features for safety with structured logging, tracing, and metrics, enabling detection of unsafe patterns and regressions in production.
- Governance & Collaboration
- Prepare evidence for governance reviews—test reports, evaluation summaries, and mitigation validation—aligned with internal Responsible AI standards.
- Collaborate with Product and UX to improve safety interactions (warnings, confirmations, refusal messaging, and feedback collection), and align evaluation goals with the Research and Data teams.
Qualifications
- Bachelor's or Master's degree in Computer Science, Software Engineering, Cybersecurity, or a related technical field—or equivalent practical experience.
- 4+ years building production software, with direct experience working on—or securing—ML/LLM systems.
- Strong software engineering skills with the ability to write production-grade code (primarily Python), beyond scripting or notebook prototyping.
- Solid understanding of LLMs and ML: how models work, prompt engineering, and the safety implications of fine-tuning and RAG (e.g., unsafe retrieval, tool misuse, and data exfiltration).
- A security mindset with demonstrated threat-modeling ability; able to threat-model AI workflows and familiar with the fundamentals of access control, data retention, and incident response.
- Familiarity with the LLM attack surface—prompt injection, jailbreaks, data poisoning, and supply-chain risk—and working knowledge of the OWASP LLM Top 10.
- Hands-on experience with at least one of safety evaluation or red teaming, with the ability to walk through a real finding and how it was remediated.
Preferred Qualifications
- Hands-on experience with industry safety tooling such as garak, PyRIT, promptfoo, Giskard, and NeMo Guardrails, and the ability to articulate the trade-offs between them.
- Visible output in AI safety or security: publications at relevant venues (e.g., the NeurIPS AI Safety Workshop, USENIX Security, or DEF CON AI Village), open-source contributions, or responsible disclosures on frontier models with public write-ups.
- Familiarity with AI governance and compliance frameworks (NIST AI RMF, ISO/IEC 42001, EU AI Act) and the ability to translate compliance requirements into concrete engineering tasks.
- Engineering experience with agents, RL environments, and/or tool use.
- Practical experience with threat-modeling methodologies such as MITRE ATLAS and STRIDE/PASTA.
Company Description
Propio is on a mission to make communication accessible to everyone. As a leader in real-time interpretation and multilingual language services, we connect people with the information they need across language, culture, and modality. We are committed to building AI-powered tools that enhance interpreter workflows, automate multilingual insights, and scale communication quality across industries.
- A tech company is seeking a skilled Prompt Engineer for a remote position in the European Union. The ideal candidate will design, test, and optimize prompts to enhance AI model performance. Responsibilities include collaborating with data scientists, ensuring compliance...SuggestedRemote workFlexible hours
$48 - $72 per hour
Senior Brand Safety Engineering Lead - Ads Join to apply for the Senior Brand Safety Engineering Lead - Ads role at Jobright.ai Senior Brand Safety Engineering Lead - Ads 2 days ago Be among... ...learning models, AI infra, and Grok LLM to enhance content detection accuracy...SuggestedFull timeFor contractorsH1bWork at officeRemote work$320k
...NVIDIA is seeking a Distinguished Engineer to serve as the founding technical leader for our AI Safety & Security Engineering team. Rooted in the firm belief that open... ...Agent systems: Familiarity with agent frameworks or LLM-based tooling.Evaluation: Experience building...SuggestedFull timeRemote work$187.78k - $223.66k
...civil shared experiences for everyone.As a Senior QA Engineer, you will be the first dedicated QA hire for the Safety Engineering Group, which is responsible for... ...proactively surface release risks to stakeholders.Apply AI/LLM-driven solutions to reduce manual QA toil and...SuggestedFull timeWork experience placementH1bWork at officeLocal areaVisa sponsorshipMonday to Friday$152k - $241.5k
...weight models are foundational to American AI leadership and cybersecurity, and that... ...and broad scientific scrutiny. Our AI Safety & Security Engineering team builds and evaluates AI-powered... ...: Experience with agent frameworks, LLM orchestration, ML infrastructure, or...SuggestedFull timeRemote work$196k - $294k
.... As the team behind Next.js, v0, and AI SDK, we create products that help builders... ...define what comes next.As a Software Engineer on our Trust & Safety at Vercel, you’ll build and operate... ...engineering, data analysis, and applied LLM techniques, working closely with...Work at officeRemote workWork from homeWorldwideMonday to FridayFlexible hours$37 - $74 per hour
Job Description Position Title: Senior Safety Methods Engineer Position Description: Protingent Staffing has an exciting contract Senior Safety... ...the forefront of innovation - from Software and Aerospace to AI, Clean Tech, Medical Devices, and Connected Technologies. We...Permanent employmentContract workRemote work- ...organization, apply now.We are currently seeking a AI Safety and Responsible AI Lead to join our team... ....Cross-functional alignment across engineering, product, legal, compliance, model risk,... ...environments.Understanding of LLM-specific risks such as hallucination, bias...Work at officeRemote workFlexible hours
$86.8k - $198k
Software System Safety EngineerThe Opportunity: As a full stack developer, you can resolve... ...clearanceBachelor’s degree in CS, EE, Engineering, or Software EngineeringNice If You Have:... ...and firmware developmentExperience with AI/ML software safety and test and evaluation...Full timeContract workPart timeWork at officeLocal areaRemote work- 10a Labs is seeking a Machine Learning Engineer to design, build, and evaluate advanced ML systems for AI safety and model evaluation. You will work on reinforcement learning, NLP/LLMs, multimodal systems, and classifiers. Collaborating with engineers, analysts, red teamers...Remote job
$92.2k - $141.4k
...online and offline, and through distributors around the world. AI at SharkNinja At SharkNinja, we’re building an AI-native... ...been invented yet, you’ll fit right in.The Senior Product Safety & Compliance Engineer will have a direct partnership with our global product developers...Temporary workLocal areaFlexible hoursShift work$280k - $330k
Senior Machine Learning Engineer - Trust and Safety On-Site: Hybrid 3-4 days a week in-office Salary: $2... ...You will be part of a world-class ML/AI team, developing machine learning models... ...product teams. Create and apply NLP, LLM AI models for customer safety in text,...Work at office3 days per week- A biosecurity innovation firm seeks a Software Engineer to enhance evaluations of frontier AI systems, focusing on security and misuse risk. In this remote role, you'll manage evaluations for new AI models, ensuring thorough analysis and collaboration with research scientists...Remote work
- ...role Moonshot is recruiting a Head of AI Safety to lead the delivery, development, and... ...and safety, product, research, and engineering teams. This is not an engineering or data... ...other AI systems. Understanding of LLM architecture , safety tooling, or trust...Permanent employmentFull timeFlexible hours
$315k
We are looking for Research Engineers to build “gold standard” evaluations for catastrophic risks, in order to understand what AI Safety Level (ASL) to assign to models. Research leads on this team collaborate with engineers in one of our focus areas: CBRN, Cyber, Autonomy...Currently hiringWork at officeImmediate startHome officeVisa sponsorshipRelocation package$210k - $260k
OpenAI is looking for an experienced Analytics Engineer to join the Safety Systems team in San Francisco, CA. You will design data solutions that enhance decision-making and drive strategic initiatives through analytics. The role involves collaborating with various teams...Work at officeRemote work- Mercor is building a remote AI red team to stress test conversational models and surface actionable vulnerabilities. The role emphasizes... ...projects with clear guidelines and wellness resources in a safety-focused environment. You will review outputs on sensitive topics...Remote job
- ...Senior Full-Stack Engineer (Freelance/Contract)NexaFlow is seeking a Senior Full-Stack Engineer to develop a proof-of-concept code review... ...leverages LLMs to automatically review pull requests.Integrate LLM APIs (OpenAI/Anthropic) to identify bugs, security vulnerabilities...Full timeContract workFreelanceRemote work
- Mercor is assembling a remote red team to test and strengthen AI systems through adversarial inputs. You will annotate failures, classify vulnerabilities, and surface systemic risks using rigorous taxonomies and playbooks. The role rewards clear communication of risks...Remote job
- About the Team The Safety Systems team is dedicated to ensuring the... ..., and reliability of AI models and their deployment in... ...About the Role As an Analytics Engineer in Safety Systems, you will play... ...building agentic data tools, LLM‑powered analytics, or other AI...Work at officeRelocation package
$65 per hour
A leading AI consulting firm is seeking an AI Tutor specialized in Coding. This part-time freelance role involves evaluating AI models, creating test cases, and developing automation tools. You must hold a relevant degree, possess advanced English skills, and have experience...Remote jobPart timeFreelanceFlexible hours- ...generation code review tool powered by large language models. We are seeking a Senior Full-Stack Engineer to lead the development of a proof-of-concept prototype that integrates LLM-based analysis into existing CI/CD workflows. Responsibilities: - Design and implement...Full timeContract workRemote work
- Our Trust and Safety RD team is fast-growing and responsible for building machine learning... ...research scientists and machine learning engineers who can take initiative, design and... ...research direction including but not limited to LLM and application in Safety, moderation...
- ...evonik.com/en/about/meet-the-team/RESPONSIBILTIES Provide process safety support to day-to-day manufacturing and projects. Ensure... ...stakeholders to promote process safety. Enhance competencies of site engineers and chemists in process safety. Stay updated on process safety...Full timeLocal areaFlexible hoursShift work
- ...Risk Management, Compliance, Business Process, IT Effectiveness, Engineering, Environmental, Sustainability, and Human Capital. We help... ...ProSidian Consulting at DescriptionProSidian Seeks a Process Safety Engineer | Technical Due Diligence & Engineering Validation For...Full timeContract workTemporary workFor contractorsWork at officeRemote workFlexible hours
- Job DescriptionProcess Safety Engineer BristolFull Time Permanent An exciting opportunity is available for a Process Safety Engineer to Join the team in Bristol on a full-time permanent basis.Why join Rolls-Royce?At Rolls-Royce we are proud to be a business that has truly...Permanent employmentFull timeFlexible hours
$120k - $130k
...Prospect, we are a $990 million division of Bosch, a multinational engineering and electronics organization and the largest privately held... ..., Illinois. (3 days in office per week). The Senior Product Safety Engineer plays a critical role to ensure the safety of new power...Full timeTemporary workWork experience placementWork at officeRemote workFlexible hours3 days per week- Bumble Inc. is seeking a Senior Software Engineer to join the Trust & Safety team in Austin. In this impactful role, you will develop reliable infrastructure that ensures user safety online. You’ll collaborate with engineers and data scientists to create intelligent systems...
$132.4k - $251.6k
...headquartered in Arlington, VA. For more than 70 years, scientists and engineers in a wide ranging disciplines at RTX BBN Technologies have... ..., sonar systems, or marine environment processing.Modern AI Pipelines: Experience implementing machine learning techniques directly...Temporary workWork experience placementWork at officeRemote workWorldwideFlexible hours$260.3k - $312.97k
...safer, more civil shared experiences for everyone.Why Safety?At Roblox, we strive to connect a billion people... ...around the world.Why Text Safety?As a Senior Engineer you will play a key role in advancing large-scale AI systems that strengthen text safety and integrity....Full timeWork experience placementH1bWork at officeLocal areaVisa sponsorshipMonday to Friday
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI/LLM Safety Engineer. Be the first to apply!


