AboutTermsRefundsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login

Curated AI safety and governance jobs

Browse jobs

Popular searches

AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs

Filters

Experience

Salary

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

108 active roles found for Anthropic

AN
AnthropicNew
Policy Design Manager, Conventional Weapons
DetailsApply

Policy role in Anthropic's Safeguards organization focused on conventional weapons: defining policy boundaries, building threat models and evaluations, and operationalizing enforcement for model misuse prevention.

Added 2 days agoFriendly (Travel-Required) | San Francisco, CA | New York City, NY; Washington, DCRemoteAI Governance & Policy$245K - $285K / year
View details
Apply
AN
AnthropicNew
Safety & Security Counsel, EMEA
AN
AnthropicNew
Safety & Security Counsel, EMEA
AN
AnthropicNew
Safety & Security Counsel
AN
AnthropicNew
Safeguards Policy Analyst, Cyber Harms
AN
Anthropic
External Affairs, US Federal
AN
Anthropic
Software Engineer, Infrastructure, Interpretability
AN
Anthropic
Staff+ Site Reliability Engineer, Safeguards ML Infra
AN
Anthropic
Staff+ Software Engineer, Privacy
AN
Anthropic
Machine Learning Infrastructure Engineer, Safeguards Research
AN
Anthropic
Evals Infrastructure Tech Lead / Manager
AN
Anthropic
Lead, Frontier Red Team (Cyber)
AN
Anthropic
Product Manager, Safeguards (Child Safety)
AN
Anthropic
Safeguards Enforcement Analyst, Violence & Extremism
AN
Anthropic
Safeguards Enforcement Analyst, Fraud & Scams
AN
Anthropic
Safeguards Enforcement Analyst, Ban Evasion & Recidivism
AN
Anthropic
Safeguards Enforcement Analyst, Account Takeover & Credential Abuse
AN
Anthropic
Safeguards Enforcement Analyst, Access Controls & Identity
AN
Anthropic
Safeguards Enforcement Analyst, Integrity & Authenticity
AN
Anthropic
Safeguards Enforcement Analyst, Cyber Harm

Showing 1–20 of 108 roles

Previous
1
2
3
4
5
6
Next
Previous

Page 1 of 6

Next
Details
Apply

Senior legal counsel role focused on AI safety and security policy, misuse prevention, red-teaming, incident response, and regulatory engagement for Claude and Anthropic.

Added 2 days agoDublin, IEAI Compliance & Risk Management€210K - €280K / year
View details
Apply
Details
Apply

Legal counsel role focused on AI safety and security, including usage policy enforcement, misuse detection, red-teaming, incident response, and AI safety/security regulatory engagement.

Added 2 days agoLondon, UKAI Compliance & Risk Management£165K - £205K / year
View details
Apply
Weekly signal

Weekly curated roles

Get the best AI safety, governance, policy and responsible AI roles in your inbox.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

Senior legal counsel role focused on AI safety and security operations, including policy enforcement, misuse detection, red-teaming, incident response, and AI safety/security regulatory engagement.

Added 2 days agoSan Francisco, CAAI Compliance & Risk Management$265K - $335K / year
View details
Apply
Details
Apply

Policy analyst role focused on cyber product policy for Anthropic AI systems, including enforcement guidance, controlled-access frameworks, model release reviews, and regulatory coordination.

Added 3 days agoSan Francisco, CA | Washington, DCAI Governance & Policy$190K - $285K / year
View details
Apply
Details
Apply

Federal policy partner role focused on AI policy advocacy, legislative/regulatory outcomes, and shaping AI governance through engagement with Senate Republicans and other stakeholders.

Added 7 days agoWashington, DCAI Governance & Policy$265K - $295K / year
View details
Apply
Details
Apply

Infrastructure engineer for Anthropic's Interpretability team, building secure research environments, data systems, and compute tooling that support interpretability work tied to frontier AI safety and audit pipelines.

Added 11 days agoSan Francisco, CA | New York City, NYRemoteAI Safety & Alignment$320K - $485K / year
View details
Apply
Details
Apply

Staff+ SRE role on Anthropic's Safeguards ML Infra team, responsible for production infrastructure, deployment, and validation of safety systems and safety classifiers for Claude model launches.

Added 13 days agoFriendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NYRemoteAI Safety & Alignment$405K - $485K / year
View details
Apply
Details
Apply

Senior privacy engineering role at Anthropic focused on building privacy-preserving systems for AI training and inference, translating AI/privacy regulations into technical controls, and conducting privacy reviews and threat modeling for new models and features.

Added 16 days agoSan Francisco, CA | New York City, NY | Seattle, WAAI Compliance & Risk Management$405K - $485K / year
View details
Apply
Details
Apply

Infrastructure engineer for Anthropic’s Safeguards research team, building tooling, pipelines, and evaluation/scoring workflows that support misuse detection and safety research for frontier AI systems.

Added 26 days agoSan Francisco, CAAI Safety & Alignment$350K - $500K / year
View details
Apply
Details
Apply

Tech lead/manager for evals infrastructure at Anthropic, building distributed systems and harnesses for frontier model evaluations that support safety decisions and launch readiness.

Added Jul 23, 2026San Francisco, CAAI Safety & Alignment$500K - $850K / year
View details
Apply
Details
Apply

Lead Anthropic’s frontier cyber red team, overseeing research on offensive and defensive capabilities of Claude, model safeguarding, and defenses against advanced AI-enabled cybersecurity risks.

Added Jul 16, 2026San Francisco, CA | New York City, NYAI Safety & Alignment$485K - $690K / year
View details
Apply
Details
Apply

Product Manager for Anthropic’s Safeguards team, owning safety systems, evals, detections, interventions, and product UX to mitigate deployment and user risks for frontier AI models.

Added Jul 15, 2026San Francisco, CAAI Safety & Alignment$305K - $385K / year
View details
Apply
Details
Apply

Safeguards enforcement role focused on AI model behavior, misuse prevention, and evaluation/enforcement workflows for violent extremism and other harmful content.

Added Jul 15, 2026San Francisco, CARemoteAI Compliance & Risk Management$285K - $330K / year
View details
Apply
Details
Apply

Safeguards enforcement role at Anthropic focused on fraud, scams, abuse detection, and policy enforcement workflows for AI products.

Added Jul 14, 2026San Francisco, CARemoteAI Compliance & Risk Management$245K - $285K / year
View details
Apply
Details
Apply

Safeguards enforcement role at Anthropic focused on detecting ban evasion, linking accounts across identities, and building controls to prevent harmful misuse of AI products.

Added Jul 14, 2026San Francisco, CARemoteAI Compliance & Risk Management$245K - $285K / year
View details
Apply
Details
Apply

Safeguards enforcement analyst for Anthropic’s account abuse team, focused on AI product misuse prevention, compromise response, and policy enforcement workflows.

Added Jul 14, 2026San Francisco, CARemoteAI Compliance & Risk Management$245K - $285K / year
View details
Apply
Details
Apply

Analyst role on Anthropic’s account abuse team focused on enforcement workflows, identity verification, access controls, and policy enforcement for safe AI product use.

Added Jul 14, 2026San Francisco, CARemoteAI Compliance & Risk Management$285K - $330K / year
View details
Apply
Details
Apply

Trust & safety / safeguards analyst role at Anthropic focused on enforcing policies against misuse of AI systems, including influence operations, election interference, and surveillance harms.

Added Jul 11, 2026San Francisco, CARemoteAI Compliance & Risk Management$285K - $330K / year
View details
Apply
Details
Apply

Enforcement analyst role focused on reviewing flagged activity and enforcing policies to prevent misuse of Anthropic's AI systems for cyberattacks, malware, and related harmful operations.

Added Jul 11, 2026San Francisco, CARemoteAI Safety & Alignment$285K - $330K / year
View details
Apply