AboutTermsRefundsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login

Curated AI safety and governance jobs

Browse jobs

Popular searches

AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs

Filters

Experience

Salary

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

108 active roles found for Anthropic

AN
Anthropic
Safeguards Enforcement Analyst, Child Safety
DetailsApply

Operational trust & safety role focused on enforcing child safety policies for Anthropic's AI products, including detection, review, escalation, and reporting for AI-facilitated CSAM/CSEM.

Added Jul 11, 2026San Francisco, CARemoteAI Compliance & Risk Management$245K - $285K / year
View details
Apply
AN
Anthropic
Safeguards Enforcement Analyst, Age-Appropriate Design
AN
Anthropic
Red Team Engineer, Safeguards
AN
Anthropic
Engineering Manager, Safeguards Interventions
AN
Anthropic
Software Engineer, Identity & Access Controls
AN
Anthropic
Engineering Manager, Enterprise
AN
Anthropic
Research Engineer, Code RL (Reinforcement Learning)
AN
Anthropic
Product Manager, Safeguards Rare Harms
AN
Anthropic
Policy Communications Manager
AN
Anthropic
Engineering Manager, Safeguards Review Tooling
AN
Anthropic
Applied AI Security Architect
AN
Anthropic
Software Engineer, Safeguards Evals
AN
Anthropic
Staff+ Security Engineer, Risk Engineering
AN
Anthropic
Security Controls Assurance Lead
AN
Anthropic
Product Manager, Claude Code Model Performance
AN
Anthropic
Software Engineer, RL Data
AN
Anthropic
Data Scientist, Policy
AN
Anthropic
Data Scientist, Safeguards
AN
Anthropic
Manager of Applied AI Architecture, Startups
AN
Anthropic
Research Engineer, Machine Learning (Reinforcement Learning)

Showing 21–40 of 108 roles

Previous
1
2
3
4
5
6
Next
Previous

Page 2 of 6

Next
Details
Apply

Safeguards enforcement role for Anthropic’s consumer AI products, focused on age assurance, policy enforcement, harmful-use detection, and privacy-preserving safety workflows.

Added Jul 11, 2026San Francisco, CARemoteAI Compliance & Risk Management$245K - $285K / year
View details
Apply
Details
Apply

Red team role on Anthropic's Safeguards team focused on adversarial testing of deployed AI systems, model jailbreaking, prompt injection, and detecting novel abuse in frontier AI products.

Added Jul 11, 2026Friendly US Travel RequiredRemoteAI Safety & Alignment$320K - $405K / year
View details
Apply
Weekly signal

Weekly curated roles

Get the best AI safety, governance, policy and responsible AI roles in your inbox.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

Engineering manager role leading Anthropic’s Safeguards Interventions team, owning AI safety interventions, compliance systems, measurement, and production reliability for model/product deployment.

Added Jul 10, 2026San Francisco, CAAI Compliance & Risk Management$405K - $485K / year
View details
Apply
Details
Apply

Software engineer building identity, vetting, access control, and misuse-monitoring systems for approved access to advanced Claude capabilities, with direct responsibility for safeguards and enforcement.

Added Jun 18, 2026San Francisco, CAAI Compliance & Risk Management$320K - $405K / year
View details
Apply
Details
Apply

The Engineering Manager for Enterprise at Anthropic will lead a team to build systems that ensure compliance and readiness for enterprise deployment of AI, focusing on security and regulatory requirements.

Added Jun 13, 2026San Francisco, CAAI Compliance & Risk Management$405K - $485K / year
View details
Apply
Details
Apply

The Research Engineer for Code RL will advance AI models' coding capabilities, ensuring safety and effectiveness through collaboration with alignment and red teams, focusing on responsible AI development.

Added Jun 12, 2026San Francisco, CAAI Safety & Alignment$500K - $850K / year
View details
Apply
Details
Apply

Product manager for Anthropic’s Safeguards team, building safety systems, evals, and interventions to mitigate risks from frontier AI models and user misuse.

Added Jun 11, 2026San Francisco, CAAI Safety & Alignment$305K - $385K / year
View details
Apply
Details
Apply

Policy communications manager for Anthropic focused on external communications about AI security, governance, responsible scaling, and frontier AI regulatory transparency.

Added Jun 11, 2026San Francisco, CAAI Governance & Policy$265K - $295K / year
View details
Apply
Details
Apply

Engineering manager for review tooling that supports safety investigations, enforcement actions, and privacy-compatible review workflows for Anthropic models and products.

Added Jun 11, 2026San Francisco, CAAI Compliance & Risk Management$405K - $485K / year
View details
Apply
Details
Apply

Senior pre-sales security architect role focused on helping regulated enterprise customers deploy Claude safely and securely, with explicit responsibility for GDPR, EU AI Act, DORA, and NIS2 compliance.

Added Jun 10, 2026EMEAAI Compliance & Risk Management£190K - £230K / year
View details
Apply
Details
Apply

Software engineer building evaluation infrastructure for Anthropic’s safety investigation and abuse detection systems, including datasets, metrics, and pipelines that assess misuse detection and robustness.

Added Jun 9, 2026San Francisco, CAAI Safety & Alignment$320K - $485K / year
View details
Apply
Details
Apply

The role involves managing and engineering security risks related to AI systems, focusing on risk assessment, quantification, and building AI-native risk tooling.

Added Jun 9, 2026San Francisco, CAAI Compliance & Risk Management$405K / year
View details
Apply
Details
Apply

Lead technical controls assurance within Anthropic Security GRC, defining and validating compliance controls and governance for AI-assisted and AI-performed processes, including autonomous AI operators and AI/ML systems in production.

Added Jun 9, 2026San Francisco, CAAI Compliance & Risk Management$345K / year
View details
Apply
Details
Apply

The role involves managing model launches and building evaluations to measure AI model performance, directly connecting to AI safety and alignment through the development of agentic evaluations.

Added Jun 5, 2026San Francisco and SeattleAI Safety & Alignment$305K - $460K / year
View details
Apply
Details
Apply

The role involves building systems for high-quality reinforcement learning data, focusing on AI safety research and ensuring the quality of training data.

Added Jun 3, 2026San Francisco, CAAI Safety & Alignment$320K - $485K / year
View details
Apply
Details
Apply

Data science role supporting Anthropic's policy and public affairs work by producing analyses, dashboards, and regulatory-trend tracking for policymakers and regulators.

Added Jun 3, 2026San Francisco, CAAI Governance & Policy$285K - $380K / year
View details
Apply
Details
Apply

The role involves analyzing user behavior data to provide insights on safety concerns and defining metrics to measure success in deploying safe AI systems.

Added Jun 3, 2026San Francisco, CAAI Safety & Alignment$275K - $370K / year
View details
Apply
Details
Apply

The Manager of Applied AI Architecture at Anthropic will lead a team to support startups in building AI-native products, focusing on safety, technical excellence, and societal benefits.

Added May 27, 2026San Francisco, CAAI Governance & Policy$315K - $380K / year
View details
Apply
Details
Apply

The Research Engineer role focuses on advancing reinforcement learning techniques to enhance the safety and capabilities of AI systems, collaborating with teams to ensure effective and safe model development.

Added May 23, 2026San Francisco, CA | New York City, NYAI Safety & Alignment$500K - $850K / year
View details
Apply