AboutTermsRefundsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login

Curated AI safety and governance jobs

Browse jobs

Popular searches

AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs

Filters

Experience

Salary

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

666 active roles found

AN
Anthropic
Senior Software Engineer, Inference
DetailsApply

The Senior Software Engineer, Inference at Anthropic will build and maintain systems for AI model inference, focusing on reliability and societal impacts.

Added May 22, 2026Dublin, IEAI Safety & Alignment€235K - €295K / year
View details
Apply
AN
Anthropic
Security Labs Engineer
AN
Anthropic
Safeguards Policy Analyst, Fraud & Scams
AN
Anthropic
Safeguards Enforcement Analyst, Safety Evaluations
AN
Anthropic
Research Scientist, Interpretability
AN
Anthropic
Research Operations, External Artifacts
AN
Anthropic
Research Lead, Training Insights
AN
Anthropic
Research Engineer, Universes
AN
Anthropic
Research Engineer / Scientist, Frontier Red Team (Cyber)
AN
Anthropic
Research Engineer / Scientist, Alignment Science - London
AN
Anthropic
Research Engineer / Scientist, Alignment Science
AN
Anthropic
Research Engineer, RL Infrastructure (Knowledge Work)
AN
Anthropic
Research Engineer / Research Scientist, Tokens
AN
Anthropic
Research Engineer/Research Scientist, Pre-training
AN
Anthropic
Research Engineer / Research Scientist, Pre-training
AN
Anthropic
Research Engineer, Production Model Post-Training
AN
Anthropic
Research Engineer, Production Model Post-Training
AN
Anthropic
Research Engineer, Pretraining Scaling - London
AN
Anthropic
Research Engineer, Pretraining Scaling
AN
Anthropic
Research Engineer, Pretraining

Showing 601–620 of 666 roles

Previous
1
…30
31
32
…34
Next
Previous

Page 31 of 34

Next
Details
Apply

The Security Labs Engineer role at Anthropic focuses on addressing high-risk security challenges related to AI systems, including adversarial threats and innovative security projects.

Added May 22, 2026San Francisco, CAAI Compliance & Risk Management$405K - $485K / year
View details
Apply
Details
Apply

Policy analyst for Anthropic’s fraud and scams safeguards, responsible for enforcement policies, threat modeling, and classifier-guided detection on AI products and APIs.

Added May 22, 2026San Francisco, CARemoteAI Compliance & Risk Management$245K - $285K / year
View details
Apply
Weekly signal

Weekly curated roles

Get the best AI safety, governance, policy and responsible AI roles in your inbox.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

The Safeguards Enforcement Analyst will ensure AI models meet safety and policy standards through evaluations and mitigations, collaborating with cross-functional teams.

Added May 22, 2026Friendly (Travel-Required) | San Francisco, CA | Washington, DC; San Francisco, CA | New York City, NYRemoteAI Safety & Alignment$230K - $270K / year
View details
Apply
Details
Apply

The Research Scientist, Interpretability role at Anthropic focuses on mechanistic interpretability to enhance the safety and understanding of AI systems.

Added May 22, 2026San Francisco, CARemoteAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The Research Operations Specialist will manage the production of system cards and safety documentation, coordinating contributions from various teams to ensure accuracy and consistency in AI safety claims.

Added May 22, 2026Friendly, United StatesRemoteAI Safety & Alignment$260K - $310K / year
View details
Apply
Details
Apply

The Research Lead for Training Insights at Anthropic will develop strategies for evaluating AI model capabilities, focusing on safety and governance through innovative evaluation methodologies.

Added May 22, 2026Friendly (Travel Required) | San Francisco, CA; San Francisco, CA | New York City, NYRemoteAI Safety & Alignment$850K / year
View details
Apply
Details
Apply

The Research Engineer, Universes role at Anthropic focuses on developing training environments for safe AI systems and includes responsibilities for building evaluations to measure AI capabilities.

Added May 22, 2026Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NYRemoteAI Safety & Alignment$500K - $850K / year
View details
Apply
Details
Apply

Technical research role on Anthropic's Frontier Red Team focused on evaluating and defending against advanced AI-enabled cyber threats, including autonomous capability evaluations, safety defenses, and policy-relevant demonstrations.

Added May 22, 2026San Francisco, CAAI Safety & Alignment$320K - $485K / year
View details
Apply
Details
Apply

The role involves conducting research on AI safety and alignment, focusing on understanding and steering the behavior of powerful AI systems.

Added May 22, 2026London, UKAI Safety & Alignment£260K - £370K / year
View details
Apply
Details
Apply

Research Engineer/Scientist on Anthropic’s Alignment Science team, conducting experimental AI safety research on powerful future systems, safety evaluations, alignment stress-testing, and related safeguards work.

Added May 22, 2026Bay AreaAI Safety & Alignment$350K - $500K / year
View details
Apply
Details
Apply

The Research Engineer will focus on the reliability and integrity of AI training environments and evaluations, ensuring they are stable and high-quality.

Added May 22, 2026San Francisco, CAAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The role focuses on building reliable and interpretable AI systems, emphasizing safety and societal impacts.

Added May 22, 2026New York City, NY; Seattle, WA; San Francisco, CAAI Safety & Alignment$350K - $500K / year
View details
Apply
Details
Apply

The Research Engineer/Research Scientist role at Anthropic focuses on developing large language models with an emphasis on safety, alignment, and societal impacts.

Added May 22, 2026Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NYRemoteAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The role involves research and engineering to develop safe and trustworthy large language models, focusing on multimodal capabilities and ethical implications of AI.

Added May 22, 2026Zürich, CHAI Safety & AlignmentCHF 280K - CHF 680K / year
View details
Apply
Details
Apply

The Research Engineer will enhance AI model safety and alignment through post-training techniques, impacting the quality and capabilities of production models.

Added May 22, 2026Zürich, CHAI Safety & Alignment
View details
Apply
Details
Apply

The Research Engineer will focus on post-training processes to enhance AI model safety and alignment, implementing techniques to improve production model quality.

Added May 22, 2026San Francisco, CA | New York City, NY | Seattle, WAAI Safety & Alignment$350K - $500K / year
View details
Apply
Details
Apply

The Research Engineer will work on training and optimizing large-scale AI models, ensuring their reliability and safety, and addressing production issues.

Added May 22, 2026London, UKAI Safety & Alignment£260K - £630K / year
View details
Apply
Details
Apply

The Research Engineer will work on training and optimizing production pretrained models, ensuring their reliability and efficiency, with a focus on the societal impacts and safety of AI systems.

Added May 22, 2026San Francisco, CAAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The Research Engineer, Pretraining role at Anthropic focuses on developing large language models with an emphasis on safety, alignment, and societal impacts.

Added May 22, 2026London, UKAI Safety & Alignment£260K - £630K / year
View details
Apply