AboutTermsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login

Curated AI safety and governance jobs

Browse jobs

Popular searches

AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs

Filters

Experience

Salary

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

95 active roles found for Anthropic

AN
Anthropic
Cyber Evaluations Engineer
DetailsApply

Engineer role focused on cyber-relevant model evaluations, safeguard robustness testing, and misuse detection for frontier AI systems.

Added Sep 2, 2026San Francisco, CA | Washington, DCRemoteAI Safety & Alignment$300K - $405K / year
View details
Apply
AN
Anthropic
Safeguards Enforcement Lead, Cyber Harms
AN
Anthropic
Product Manager, Multi-Cloud Trust & Safety
AN
Anthropic
Head of Policy Design, Societal Harms
AN
Anthropic
Product Policy Manager, Product Risk
AN
Anthropic
Policy Design Manager, Conventional Weapons
AN
Anthropic
Staff+ Software Engineer, Privacy
AN
Anthropic
Machine Learning Infrastructure Engineer, Safeguards Research
AN
Anthropic
Safeguards Enforcement Analyst, Account Takeover & Credential Abuse
AN
Anthropic
Safeguards Enforcement Analyst, Access Controls & Identity
AN
Anthropic
Safeguards Enforcement Analyst, Integrity & Authenticity
AN
Anthropic
Safeguards Enforcement Analyst, Cyber Harm
AN
Anthropic
Safeguards Enforcement Analyst, Child Safety
AN
Anthropic
Safeguards Enforcement Analyst, Age-Appropriate Design
AN
Anthropic
Red Team Engineer, Safeguards
AN
Anthropic
Engineering Manager, Enterprise
AN
Anthropic
Research Engineer, Code RL (Reinforcement Learning)
AN
Anthropic
Product Manager, Safeguards Rare Harms
AN
Anthropic
Engineering Manager, Safeguards Review Tooling
AN
Anthropic
Staff+ Security Engineer, Risk Engineering

Showing 21–40 of 95 roles

Previous
1
2
3
4
5
Next
Previous

Page 2 of 5

Next
Details
Apply

Lead enforcement operations to detect and mitigate misuse of Anthropic AI systems for cyberattacks, malware, and offensive exploitation, including managing a team and shaping enforcement strategy.

Added Aug 29, 2026Friendly (Travel-Required) | Washington, DC; San Francisco, CA | New York City, NYRemoteAI Safety & Alignment$285K - $330K / year
View details
Apply
Details
Apply

Senior product management role owning multi-cloud model safeguards, fraud defenses, and compliance posture for Claude deployments across AWS, Google Cloud, and Microsoft.

Added Aug 29, 2026San Francisco, CA | New York City, NY | Seattle, WAAI Compliance & Risk Management$305K - $385K / year
View details
Apply
Weekly signal

Weekly curated roles

Get the best AI safety, governance, policy and responsible AI roles in your inbox.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

Leadership role overseeing policy design for Claude consumer harms, including safety policies, evaluations, detection/enforcement systems, and cross-functional mitigation strategy.

Added Aug 29, 2026San Francisco, CAAI Governance & Policy$330K - $395K / year
View details
Apply
Details
Apply

Product Policy Manager focused on product risk: assessing AI product launches for safety risks, running product safety reviews and bespoke evaluations, and recommending mitigations and responsible AI deployment policies.

Added Aug 25, 2026Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NYRemoteAI Compliance & Risk Management$245K - $285K / year
View details
Apply
Details
Apply

Policy role in Anthropic's Safeguards organization focused on conventional weapons: defining policy boundaries, building threat models and evaluations, and operationalizing enforcement for model misuse prevention.

Added Aug 22, 2026Friendly (Travel-Required) | San Francisco, CA | New York City, NY; Washington, DCRemoteAI Governance & Policy$245K - $285K / year
View details
Apply
Details
Apply

Senior privacy engineering role at Anthropic focused on building privacy-preserving systems for AI training and inference, translating AI/privacy regulations into technical controls, and conducting privacy reviews and threat modeling for new models and features.

Added Aug 8, 2026San Francisco, CA | New York City, NY | Seattle, WARemoteAI Compliance & Risk Management$405K - $485K / year
View details
Apply
Details
Apply

Infrastructure engineer for Anthropic’s Safeguards research team, building tooling, pipelines, and evaluation/scoring workflows that support misuse detection and safety research for frontier AI systems.

Added Jul 29, 2026San Francisco, CAAI Safety & Alignment$350K - $500K / year
View details
Apply
Details
Apply

Safeguards enforcement analyst for Anthropic’s account abuse team, focused on AI product misuse prevention, compromise response, and policy enforcement workflows.

Added Jul 14, 2026San Francisco, CARemoteAI Compliance & Risk Management$245K - $285K / year
View details
Apply
Details
Apply

Analyst role on Anthropic’s account abuse team focused on enforcement workflows, identity verification, access controls, and policy enforcement for safe AI product use.

Added Jul 14, 2026San Francisco, CARemoteAI Compliance & Risk Management$285K - $330K / year
View details
Apply
Details
Apply

Trust & safety / safeguards analyst role at Anthropic focused on enforcing policies against misuse of AI systems, including influence operations, election interference, and surveillance harms.

Added Jul 11, 2026San Francisco, CARemoteAI Compliance & Risk Management$285K - $330K / year
View details
Apply
Details
Apply

Enforcement analyst role focused on reviewing flagged activity and enforcing policies to prevent misuse of Anthropic's AI systems for cyberattacks, malware, and related harmful operations.

Added Jul 11, 2026San Francisco, CARemoteAI Safety & Alignment$285K - $330K / year
View details
Apply
Details
Apply

Operational trust & safety role focused on enforcing child safety policies for Anthropic's AI products, including detection, review, escalation, and reporting for AI-facilitated CSAM/CSEM.

Added Jul 11, 2026San Francisco, CARemoteAI Compliance & Risk Management$245K - $285K / year
View details
Apply
Details
Apply

Safeguards enforcement role for Anthropic’s consumer AI products, focused on age assurance, policy enforcement, harmful-use detection, and privacy-preserving safety workflows.

Added Jul 11, 2026San Francisco, CARemoteAI Compliance & Risk Management$245K - $285K / year
View details
Apply
Details
Apply

Red team role on Anthropic's Safeguards team focused on adversarial testing of deployed AI systems, model jailbreaking, prompt injection, and detecting novel abuse in frontier AI products.

Added Jul 11, 2026Friendly US Travel RequiredRemoteAI Safety & Alignment$320K - $405K / year
View details
Apply
Details
Apply

The Engineering Manager for Enterprise at Anthropic will lead a team to build systems that ensure compliance and readiness for enterprise deployment of AI, focusing on security and regulatory requirements.

Added Jun 13, 2026San Francisco, CAAI Compliance & Risk Management$405K - $485K / year
View details
Apply
Details
Apply

The Research Engineer for Code RL will advance AI models' coding capabilities, ensuring safety and effectiveness through collaboration with alignment and red teams, focusing on responsible AI development.

Added Jun 12, 2026San Francisco, CAAI Safety & Alignment$500K - $850K / year
View details
Apply
Details
Apply

Product manager for Anthropic’s Safeguards team, building safety systems, evals, and interventions to mitigate risks from frontier AI models and user misuse.

Added Jun 11, 2026San Francisco, CAAI Safety & Alignment$305K - $385K / year
View details
Apply
Details
Apply

Engineering manager for review tooling that supports safety investigations, enforcement actions, and privacy-compatible review workflows for Anthropic models and products.

Added Jun 11, 2026San Francisco, CAAI Compliance & Risk Management$405K - $485K / year
View details
Apply
Details
Apply

The role involves managing and engineering security risks related to AI systems, focusing on risk assessment, quantification, and building AI-native risk tooling.

Added Jun 9, 2026San Francisco, CAAI Compliance & Risk Management$405K / year
View details
Apply