Policy role in Anthropic's Safeguards organization focused on conventional weapons: defining policy boundaries, building threat models and evaluations, and operationalizing enforcement for model misuse prevention.
A weekly digest of AI safety, governance, policy and responsible AI roles.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
108 active roles found for Anthropic
Policy role in Anthropic's Safeguards organization focused on conventional weapons: defining policy boundaries, building threat models and evaluations, and operationalizing enforcement for model misuse prevention.
Senior legal counsel role focused on AI safety and security policy, misuse prevention, red-teaming, incident response, and regulatory engagement for Claude and Anthropic.
Legal counsel role focused on AI safety and security, including usage policy enforcement, misuse detection, red-teaming, incident response, and AI safety/security regulatory engagement.
Get the best AI safety, governance, policy and responsible AI roles in your inbox.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
Senior legal counsel role focused on AI safety and security operations, including policy enforcement, misuse detection, red-teaming, incident response, and AI safety/security regulatory engagement.
Policy analyst role focused on cyber product policy for Anthropic AI systems, including enforcement guidance, controlled-access frameworks, model release reviews, and regulatory coordination.
Federal policy partner role focused on AI policy advocacy, legislative/regulatory outcomes, and shaping AI governance through engagement with Senate Republicans and other stakeholders.
Infrastructure engineer for Anthropic's Interpretability team, building secure research environments, data systems, and compute tooling that support interpretability work tied to frontier AI safety and audit pipelines.
Staff+ SRE role on Anthropic's Safeguards ML Infra team, responsible for production infrastructure, deployment, and validation of safety systems and safety classifiers for Claude model launches.
Senior privacy engineering role at Anthropic focused on building privacy-preserving systems for AI training and inference, translating AI/privacy regulations into technical controls, and conducting privacy reviews and threat modeling for new models and features.
Infrastructure engineer for Anthropic’s Safeguards research team, building tooling, pipelines, and evaluation/scoring workflows that support misuse detection and safety research for frontier AI systems.
Tech lead/manager for evals infrastructure at Anthropic, building distributed systems and harnesses for frontier model evaluations that support safety decisions and launch readiness.
Lead Anthropic’s frontier cyber red team, overseeing research on offensive and defensive capabilities of Claude, model safeguarding, and defenses against advanced AI-enabled cybersecurity risks.
Product Manager for Anthropic’s Safeguards team, owning safety systems, evals, detections, interventions, and product UX to mitigate deployment and user risks for frontier AI models.
Safeguards enforcement role focused on AI model behavior, misuse prevention, and evaluation/enforcement workflows for violent extremism and other harmful content.
Safeguards enforcement role at Anthropic focused on fraud, scams, abuse detection, and policy enforcement workflows for AI products.
Safeguards enforcement role at Anthropic focused on detecting ban evasion, linking accounts across identities, and building controls to prevent harmful misuse of AI products.
Safeguards enforcement analyst for Anthropic’s account abuse team, focused on AI product misuse prevention, compromise response, and policy enforcement workflows.
Analyst role on Anthropic’s account abuse team focused on enforcement workflows, identity verification, access controls, and policy enforcement for safe AI product use.
Trust & safety / safeguards analyst role at Anthropic focused on enforcing policies against misuse of AI systems, including influence operations, election interference, and surveillance harms.
Enforcement analyst role focused on reviewing flagged activity and enforcing policies to prevent misuse of Anthropic's AI systems for cyberattacks, malware, and related harmful operations.