Operational trust & safety role focused on enforcing child safety policies for Anthropic's AI products, including detection, review, escalation, and reporting for AI-facilitated CSAM/CSEM.
A weekly digest of AI safety, governance, policy and responsible AI roles.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
108 active roles found for Anthropic
Operational trust & safety role focused on enforcing child safety policies for Anthropic's AI products, including detection, review, escalation, and reporting for AI-facilitated CSAM/CSEM.
Safeguards enforcement role for Anthropic’s consumer AI products, focused on age assurance, policy enforcement, harmful-use detection, and privacy-preserving safety workflows.
Red team role on Anthropic's Safeguards team focused on adversarial testing of deployed AI systems, model jailbreaking, prompt injection, and detecting novel abuse in frontier AI products.
Get the best AI safety, governance, policy and responsible AI roles in your inbox.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
Engineering manager role leading Anthropic’s Safeguards Interventions team, owning AI safety interventions, compliance systems, measurement, and production reliability for model/product deployment.
Software engineer building identity, vetting, access control, and misuse-monitoring systems for approved access to advanced Claude capabilities, with direct responsibility for safeguards and enforcement.
The Engineering Manager for Enterprise at Anthropic will lead a team to build systems that ensure compliance and readiness for enterprise deployment of AI, focusing on security and regulatory requirements.
The Research Engineer for Code RL will advance AI models' coding capabilities, ensuring safety and effectiveness through collaboration with alignment and red teams, focusing on responsible AI development.
Product manager for Anthropic’s Safeguards team, building safety systems, evals, and interventions to mitigate risks from frontier AI models and user misuse.
Policy communications manager for Anthropic focused on external communications about AI security, governance, responsible scaling, and frontier AI regulatory transparency.
Engineering manager for review tooling that supports safety investigations, enforcement actions, and privacy-compatible review workflows for Anthropic models and products.
Senior pre-sales security architect role focused on helping regulated enterprise customers deploy Claude safely and securely, with explicit responsibility for GDPR, EU AI Act, DORA, and NIS2 compliance.
Software engineer building evaluation infrastructure for Anthropic’s safety investigation and abuse detection systems, including datasets, metrics, and pipelines that assess misuse detection and robustness.
The role involves managing and engineering security risks related to AI systems, focusing on risk assessment, quantification, and building AI-native risk tooling.
Lead technical controls assurance within Anthropic Security GRC, defining and validating compliance controls and governance for AI-assisted and AI-performed processes, including autonomous AI operators and AI/ML systems in production.
The role involves managing model launches and building evaluations to measure AI model performance, directly connecting to AI safety and alignment through the development of agentic evaluations.
The role involves building systems for high-quality reinforcement learning data, focusing on AI safety research and ensuring the quality of training data.
Data science role supporting Anthropic's policy and public affairs work by producing analyses, dashboards, and regulatory-trend tracking for policymakers and regulators.
The role involves analyzing user behavior data to provide insights on safety concerns and defining metrics to measure success in deploying safe AI systems.
The Manager of Applied AI Architecture at Anthropic will lead a team to support startups in building AI-native products, focusing on safety, technical excellence, and societal benefits.
The Research Engineer role focuses on advancing reinforcement learning techniques to enhance the safety and capabilities of AI systems, collaborating with teams to ensure effective and safe model development.