Trust & safety / safeguards analyst role at Anthropic focused on enforcing policies against misuse of AI systems, including influence operations, election interference, and surveillance harms.
Curated roles
Browse remote-friendly roles across AI safety, governance, policy, evaluations, compliance and responsible AI teams.
211 active roles found.
Trust & safety / safeguards analyst role at Anthropic focused on enforcing policies against misuse of AI systems, including influence operations, election interference, and surveillance harms.
Enforcement analyst role focused on reviewing flagged activity and enforcing policies to prevent misuse of Anthropic's AI systems for cyberattacks, malware, and related harmful operations.
Operational trust & safety role focused on enforcing child safety policies for Anthropic's AI products, including detection, review, escalation, and reporting for AI-facilitated CSAM/CSEM.
Safeguards enforcement role for Anthropic’s consumer AI products, focused on age assurance, policy enforcement, harmful-use detection, and privacy-preserving safety workflows.
Red team role on Anthropic's Safeguards team focused on adversarial testing of deployed AI systems, model jailbreaking, prompt injection, and detecting novel abuse in frontier AI products.
Co-lead author role updating a technical AI governance paper, involving literature review, expert coordination, synthesis, and publication-ready writing.
Lead policy advocacy for pragmatic AI safety legislation at the state level, including legislative analysis, drafting, and engagement with policymakers and AI companies.
Lead recruitment and implementation support for state AI regulatory roles, with responsibilities spanning AI safety legislation, compliance assessment, and enforcement communications.
Part-time technical fellowship building AI evaluation infrastructure, adversarial evaluation tooling, risk data systems, and governance-related prototypes for an AI risk firm.
Project-based adversarial evaluation and red teaming network for frontier AI systems, with work spanning ML security, prompt injection/jailbreaking, agentic system evaluation, and translating findings into governance and mitigation decisions.
Policy Director role supporting executive operations while building AI policy expertise at a nonprofit focused on pragmatic policies to reduce severe risks from frontier AI models.
Funding call for foundational research on safety and risk in multi-agent AI systems, including emergent dynamics, trustworthy interaction infrastructure, and scalable monitoring and control.
Open call for contractors to support CLTR’s AI policy unit and related AI-bio work, including AI governance, AI safety concepts, frontier-lab security practices, and safeguard approaches for capable models.
State policy and public affairs role focused on OpenAI’s AI policy priorities, including guardrails for advanced AI models, legislative review, and engagement with state and local governments.
EU policy role focused on shaping OpenAI’s approach to AI-related legislative and regulatory issues in Brussels, including engagement with policymakers and support for regulation and industry standards.
Research engineer role focused on building and maintaining AI safety evaluation benchmarks, guardrails, and research on agentic failure modes.
Research scientist role focused on AI behavior failure modes, model evaluations, and safety research for LLM agents, including deception, misalignment, unsafe behavior, and frontier agent pressure-testing.
Senior ML engineer role building product experiences, evaluation systems, and trustworthy language-model workflows for research and high-stakes decision-making; relevant because it explicitly emphasizes careful evaluations, process supervision, and safer AI systems.
Senior applied AI scientist role building trust, guardrails, evaluators, and safety/security classifiers for LLM and agentic applications in production.
Contract research engineer role focused on frontier AI safety evaluations, dangerous behavior elicitation, and safety report writing for loss-of-control and harmful manipulation risks.