Engineer role focused on cyber-relevant model evaluations, safeguard robustness testing, and misuse detection for frontier AI systems.
A weekly digest of AI safety, governance, policy and responsible AI roles.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
95 active roles found for Anthropic
Engineer role focused on cyber-relevant model evaluations, safeguard robustness testing, and misuse detection for frontier AI systems.
Lead enforcement operations to detect and mitigate misuse of Anthropic AI systems for cyberattacks, malware, and offensive exploitation, including managing a team and shaping enforcement strategy.
Senior product management role owning multi-cloud model safeguards, fraud defenses, and compliance posture for Claude deployments across AWS, Google Cloud, and Microsoft.
Get the best AI safety, governance, policy and responsible AI roles in your inbox.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
Leadership role overseeing policy design for Claude consumer harms, including safety policies, evaluations, detection/enforcement systems, and cross-functional mitigation strategy.
Product Policy Manager focused on product risk: assessing AI product launches for safety risks, running product safety reviews and bespoke evaluations, and recommending mitigations and responsible AI deployment policies.
Policy role in Anthropic's Safeguards organization focused on conventional weapons: defining policy boundaries, building threat models and evaluations, and operationalizing enforcement for model misuse prevention.
Senior privacy engineering role at Anthropic focused on building privacy-preserving systems for AI training and inference, translating AI/privacy regulations into technical controls, and conducting privacy reviews and threat modeling for new models and features.
Infrastructure engineer for Anthropic’s Safeguards research team, building tooling, pipelines, and evaluation/scoring workflows that support misuse detection and safety research for frontier AI systems.
Safeguards enforcement analyst for Anthropic’s account abuse team, focused on AI product misuse prevention, compromise response, and policy enforcement workflows.
Analyst role on Anthropic’s account abuse team focused on enforcement workflows, identity verification, access controls, and policy enforcement for safe AI product use.
Trust & safety / safeguards analyst role at Anthropic focused on enforcing policies against misuse of AI systems, including influence operations, election interference, and surveillance harms.
Enforcement analyst role focused on reviewing flagged activity and enforcing policies to prevent misuse of Anthropic's AI systems for cyberattacks, malware, and related harmful operations.
Operational trust & safety role focused on enforcing child safety policies for Anthropic's AI products, including detection, review, escalation, and reporting for AI-facilitated CSAM/CSEM.
Safeguards enforcement role for Anthropic’s consumer AI products, focused on age assurance, policy enforcement, harmful-use detection, and privacy-preserving safety workflows.
Red team role on Anthropic's Safeguards team focused on adversarial testing of deployed AI systems, model jailbreaking, prompt injection, and detecting novel abuse in frontier AI products.
The Engineering Manager for Enterprise at Anthropic will lead a team to build systems that ensure compliance and readiness for enterprise deployment of AI, focusing on security and regulatory requirements.
The Research Engineer for Code RL will advance AI models' coding capabilities, ensuring safety and effectiveness through collaboration with alignment and red teams, focusing on responsible AI development.
Product manager for Anthropic’s Safeguards team, building safety systems, evals, and interventions to mitigate risks from frontier AI models and user misuse.
Engineering manager for review tooling that supports safety investigations, enforcement actions, and privacy-compatible review workflows for Anthropic models and products.
The role involves managing and engineering security risks related to AI systems, focusing on risk assessment, quantification, and building AI-native risk tooling.