Machine Learning Engineer focused on AI safety solutions, including adversarial testing, model evaluation, robust inference, and monitoring for secure deployment of AI systems.
Curated roles
Browse remote-friendly roles across AI safety, governance, policy, evaluations, compliance and responsible AI teams.
216 active roles found.
Machine Learning Engineer focused on AI safety solutions, including adversarial testing, model evaluation, robust inference, and monitoring for secure deployment of AI systems.
Engagement manager role leading paid campaigns and client work focused on AI risk communications, educating decision makers, and supporting the AI safety ecosystem.
Software engineering role on Perplexity’s Model Behavior team focused on prompt/context engineering, model behavior shaping, failure-mode analysis, and some evaluation work for LLM systems.
Research Scientist role focused on AI evaluation, language model understanding, robustness, red teaming, and alignment for LLM evaluation infrastructure.
Cybersecurity Engineer building infrastructure and tooling for AI security evaluations, adversarial testing, and attack simulations to assess and improve the resilience of AI-powered systems.
Compliance Manager responsible for security/privacy compliance and, importantly, operationalizing AI governance practices aligned with the EU AI Act and NIST AI Risk Management Framework.
Contract infrastructure engineer role building and scaling the Loss of Control Observatory’s data pipelines, LLM-based classification systems, and monitoring dashboard for AI safety monitoring.
Voluntary board member role providing governance, oversight, and strategic guidance for ERA’s frontier AI risk mitigation work, with explicit emphasis on AI safety and AI governance experience.
Research fellowship conducting independent AI governance research to inform high-stakes policy and strategic decisions, including risk management, threat modeling, technical governance, and frontier AI regulation.
ML Research Engineer focused on training, post-training, and evaluating LLMs, including alignment methods and safety/moderation-related datasets and policy systems.
AI Red Team Engineer focused on adversarial testing of LLM-powered systems, including jailbreaks, prompt injection, data leakage, policy bypass, and turning findings into regression tests and reports.
Research role building adversarial agent environments, evaluations, and tooling to study AI system failures, misalignment, and unsafe behaviors.
Research engineer role building an AI safety argumentation platform using ontologies, knowledge graphs, defeasible argumentation, and LLM-assisted pipelines to support AI risk management and governance communications.
Contract role focused on monitoring and enforcing abuse on AI products, including building detection, review, and enforcement systems for high-risk harms.
Engineering fellowship supporting AI abuse detection, red teaming, and related research/engineering work across software, data, and ML concentrations.
Senior public affairs and policy role focused on AI policy narratives, regulator engagement, and shaping the regulatory conversation around enterprise AI.
Lead Cohere’s global external affairs function within Government Affairs and Public Policy, building partnerships and policy narratives around enterprise AI, data governance, safety, and model safety.
Safety-team role focused on model behavior, alignment, and evaluation of large language models, including building evaluation pipelines and synthetic testing environments.
Lead US government affairs and AI policy strategy for Cohere, including advocacy on AI regulation, legislative drafts, and coordination with responsible AI and legal teams.
Part-time mentor role supervising 3-month research projects for aspiring researchers in AI safety, policy, governance, or biosecurity.