AboutTermsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login
All jobs

Curated roles

AI Safety Jobs

Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.

391 active roles found.

AI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AN
Anthropic
ML/Research Engineer, Safeguards
DetailsApply

The ML/Research Engineer, Safeguards role at Anthropic focuses on developing systems to detect and mitigate misuse of AI, ensuring safety and compliance in AI systems.

Added May 22, 2026San Francisco, CA | New York City, NYAI Safety & Alignment$350K - $500K / year
View details
Apply
AN
Anthropic
Full-Stack Software Engineer, Reinforcement Learning
AN
Anthropic
Research Manager, Interpretability
AN
Anthropic
Engineering Manager, GPU (ML Accelerator)
AN
Anthropic
Biological Safety Research Scientist
AN
Anthropic
Applied AI Architect, Industries
AN
Anthropic
Applied AI Architect, Applied AI (Digital Natives Business)
AN
Anthropic
Applied AI Architect
Details
AN
Anthropic
Anthropic Fellows Program, ML Systems & Performance
AN
Anthropic
Research Engineer, Model Evaluations
AN
Anthropic
Anthropic Fellows Program, AI Safety

Showing 381–391 of 391 roles

Previous
1
…17
18
19
20
Next
Previous

Page 20 of 20

Next

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

The role involves building platforms and tools for reinforcement learning, focusing on data collection, training observability, and ensuring the safety and reliability of AI systems.

Added May 22, 2026San Francisco, CA | New York City, NYAI Safety & Alignment$300K - $405K / year
View details
Apply
Details
Apply

The Research Manager for Interpretability at Anthropic will lead a team focused on understanding the internal workings of large language models, emphasizing mechanistic interpretability as a means to enhance AI safety.

Added May 22, 2026San Francisco, CAAI Safety & Alignment$350K - $500K / year
View details
Apply
Details
Apply

The Engineering Manager will lead efforts to improve model performance and ensure the safe development of AI systems at Anthropic.

Added May 22, 2026San Francisco, CA | New York City, NY | Seattle, WAAI Safety & Alignment$500K - $850K / year
View details
Apply
Details
Apply

The Biological Safety Research Scientist at Anthropic will design and develop safety systems for AI, focusing on preventing misuse and ensuring responsible AI safety in the biological domain.

Added May 22, 2026San Francisco, CAAI Safety & Alignment$300K - $320K / year
View details
Apply
Details
Apply

The Applied AI Architect role focuses on guiding enterprise customers in integrating AI systems safely and effectively, ensuring alignment with business objectives and technical implementation.

Added May 22, 2026New York City, NY; San Francisco, CA; Seattle, WAAI Safety & Alignment$240K - $315K / year
View details
Apply
Details
Apply

The Applied AI Architect role at Anthropic focuses on integrating AI solutions into enterprise technology stacks while ensuring safety and reliability, and involves developing evaluation frameworks for AI performance.

Added May 22, 2026Munich, GermanyAI Safety & Alignment
View details
Apply
Apply

The Applied AI Architect role at Anthropic focuses on providing technical guidance to enterprise customers for integrating AI solutions, emphasizing safety and reliability in AI systems.

Added May 22, 2026Tokyo, JapanAI Safety & Alignment
View details
Apply
Details
Apply

The Anthropic Fellows Program offers funding and mentorship for research in AI safety and security, focusing on empirical projects aligned with societal benefits.

Added May 22, 2026London, UK; Ontario, CAN;-Friendly, United States; San Francisco, CARemoteAI Safety & Alignment
View details
Apply
Details
Apply

Research engineer role focused on designing and running model evaluations for Claude, including safety properties, agentic behavior, and evaluation infrastructure at scale.

Added May 22, 2026San Francisco, CARemoteAI Safety & Alignment$500K - $850K / year
View details
Apply
Details
Apply

A full-time AI safety fellowship at Anthropic focused on empirical research projects in AI safety areas such as scalable oversight, adversarial robustness, AI control, and mechanistic interpretability.

Added May 22, 2026Berkeley, California or London, UKRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply