Analyst role on Anthropic’s account abuse team focused on enforcement workflows, identity verification, access controls, and policy enforcement for safe AI product use.
A weekly digest of AI safety, governance, policy and responsible AI roles.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
677 active roles found
Analyst role on Anthropic’s account abuse team focused on enforcement workflows, identity verification, access controls, and policy enforcement for safe AI product use.
Research Manager for an AI safety research training programme, supporting teams, technical enablement, proposal evaluation, and participant development.
Senior research staff role leading applied research on AI risks, mitigations, and governance gaps, with stakeholder engagement and decision-support outputs for policymakers and other end users.
Get the best AI safety, governance, policy and responsible AI roles in your inbox.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
Research Engineer/Scientist on SaferAI’s Evaluations team, focused on model evaluations, threat modeling, risk reports, and safety mitigations for frontier AI labs.
Founding Engineer on SaferAI’s Evaluations team, building infrastructure for model evaluations and helping run mitigation-focused safety evaluations for frontier AI labs.
Customer Success role helping enterprises implement AI governance, policy, testing, monitoring, and compliance workflows using ModelOp's AI lifecycle governance platform.
Trust & safety / safeguards analyst role at Anthropic focused on enforcing policies against misuse of AI systems, including influence operations, election interference, and surveillance harms.
Enforcement analyst role focused on reviewing flagged activity and enforcing policies to prevent misuse of Anthropic's AI systems for cyberattacks, malware, and related harmful operations.
Operational trust & safety role focused on enforcing child safety policies for Anthropic's AI products, including detection, review, escalation, and reporting for AI-facilitated CSAM/CSEM.
Safeguards enforcement role for Anthropic’s consumer AI products, focused on age assurance, policy enforcement, harmful-use detection, and privacy-preserving safety workflows.
Red team role on Anthropic's Safeguards team focused on adversarial testing of deployed AI systems, model jailbreaking, prompt injection, and detecting novel abuse in frontier AI products.
Co-lead author role updating a technical AI governance paper, involving literature review, expert coordination, synthesis, and publication-ready writing.
Lead policy advocacy for pragmatic AI safety legislation at the state level, including legislative analysis, drafting, and engagement with policymakers and AI companies.
Lead recruitment and implementation support for state AI regulatory roles, with responsibilities spanning AI safety legislation, compliance assessment, and enforcement communications.
Project-based adversarial evaluation and red teaming network for frontier AI systems, with work spanning ML security, prompt injection/jailbreaking, agentic system evaluation, and translating findings into governance and mitigation decisions.
Lead threat modeling for CBRNe risks in advanced AI systems, building evaluations and safety frameworks to inform deployment decisions and mitigate dual-use misuse.
Policy Director role supporting executive operations while building AI policy expertise at a nonprofit focused on pragmatic policies to reduce severe risks from frontier AI models.
CTO role leading engineering at Syntony, with direct responsibility for evaluation tooling, AI risk mapping, and governance instrumentation across the product suite.
Lead role at DeepMind focused on threat modeling and safety evaluations for CBRNe risks in advanced AI models, supporting the Frontier Safety Framework and deployment decisions.
Funding call for foundational research on safety and risk in multi-agent AI systems, including emergent dynamics, trustworthy interaction infrastructure, and scalable monitoring and control.