AboutTermsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login

Curated AI safety and governance jobs

Browse jobs

Popular searches

AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs

Filters

Experience

Salary

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

677 active roles found

AN
Anthropic
Safeguards Enforcement Analyst, Safety Evaluations
DetailsApply

The Safeguards Enforcement Analyst will ensure AI models meet safety and policy standards through evaluations and mitigations, collaborating with cross-functional teams.

Added May 22, 2026Friendly (Travel-Required) | San Francisco, CA | Washington, DC; San Francisco, CA | New York City, NYRemoteAI Safety & Alignment$230K - $270K / year
View details
Apply
AN
Anthropic
Research Scientist, Interpretability
AN
Anthropic
Research Engineer, Universes
AN
Anthropic
Research Engineer / Scientist, Alignment Science - London
AN
Anthropic
Research Engineer / Scientist, Alignment Science
AN
Anthropic
Research Engineer / Research Scientist, Tokens
AN
Anthropic
Research Engineer/Research Scientist, Pre-training
AN
Anthropic
Research Engineer, Production Model Post-Training
AN
Anthropic
Research Engineer, Production Model Post-Training
AN
Anthropic
Research Engineer, Pretraining Scaling - London
AN
Anthropic
Research Engineer, Pretraining Scaling
AN
Anthropic
Research Engineer, Pretraining
AN
Anthropic
Research Engineer, Performance RL
AN
Anthropic
Research Engineer, Machine Learning (Reinforcement Learning)
AN
Anthropic
Research Engineer, Knowledge Team
AN
Anthropic
Research Engineer, Interpretability
AN
Anthropic
Research Engineer, Discovery
AN
Anthropic
Research Engineer, Cybersecurity Reinforcement Learning
AN
Anthropic
Model Quality Software Engineer, Claude Code
AN
Anthropic
ML/Research Engineer, Safeguards

Showing 641–660 of 677 roles

Previous
1
…31
32
33
34
Next
Previous

Page 33 of 34

Next
Details
Apply

The Research Scientist, Interpretability role at Anthropic focuses on mechanistic interpretability to enhance the safety and understanding of AI systems.

Added May 22, 2026San Francisco, CARemoteAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The Research Engineer, Universes role at Anthropic focuses on developing training environments for safe AI systems and includes responsibilities for building evaluations to measure AI capabilities.

Added May 22, 2026Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NYRemoteAI Safety & Alignment$500K - $850K / year
View details
Apply
Weekly signal

Weekly curated roles

Get the best AI safety, governance, policy and responsible AI roles in your inbox.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

The role involves conducting research on AI safety and alignment, focusing on understanding and steering the behavior of powerful AI systems.

Added May 22, 2026London, UKAI Safety & Alignment£260K - £370K / year
View details
Apply
Details
Apply

Research Engineer/Scientist on Anthropic’s Alignment Science team, conducting experimental AI safety research on powerful future systems, safety evaluations, alignment stress-testing, and related safeguards work.

Added May 22, 2026Bay AreaAI Safety & Alignment$350K - $500K / year
View details
Apply
Details
Apply

The role focuses on building reliable and interpretable AI systems, emphasizing safety and societal impacts.

Added May 22, 2026New York City, NY; Seattle, WA; San Francisco, CAAI Safety & Alignment$350K - $500K / year
View details
Apply
Details
Apply

The Research Engineer/Research Scientist role at Anthropic focuses on developing large language models with an emphasis on safety, alignment, and societal impacts.

Added May 22, 2026Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NYRemoteAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The Research Engineer will enhance AI model safety and alignment through post-training techniques, impacting the quality and capabilities of production models.

Added May 22, 2026Zürich, CHAI Safety & Alignment
View details
Apply
Details
Apply

The Research Engineer will focus on post-training processes to enhance AI model safety and alignment, implementing techniques to improve production model quality.

Added May 22, 2026San Francisco, CA | New York City, NY | Seattle, WAAI Safety & Alignment$350K - $500K / year
View details
Apply
Details
Apply

The Research Engineer will work on training and optimizing large-scale AI models, ensuring their reliability and safety, and addressing production issues.

Added May 22, 2026London, UKAI Safety & Alignment£260K - £630K / year
View details
Apply
Details
Apply

The Research Engineer will work on training and optimizing production pretrained models, ensuring their reliability and efficiency, with a focus on the societal impacts and safety of AI systems.

Added May 22, 2026San Francisco, CAAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The Research Engineer, Pretraining role at Anthropic focuses on developing large language models with an emphasis on safety, alignment, and societal impacts.

Added May 22, 2026London, UKAI Safety & Alignment£260K - £630K / year
View details
Apply
Details
Apply

The Research Engineer, Performance RL role focuses on advancing AI models' capabilities in safely writing code, collaborating with alignment teams to ensure safety and effectiveness.

Added May 22, 2026San Francisco, CAAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The Research Engineer role focuses on advancing the safety and capabilities of large language models through reinforcement learning, collaborating with alignment teams to ensure safe AI systems.

Added May 22, 2026London, UKAI Safety & Alignment£260K - £630K / year
View details
Apply
Details
Apply

The Research Engineer will redesign how language models interact with external data sources, focusing on safety and societal impacts.

Added May 22, 2026Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NYRemoteAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The Research Engineer, Interpretability role at Anthropic focuses on building infrastructure for interpretability research to enhance AI safety through mechanistic understanding of models.

Added May 22, 2026San Francisco, CARemoteAI Safety & Alignment$315K - $560K / year
View details
Apply
Details
Apply

The Research Engineer, Discovery role at Anthropic focuses on developing infrastructure and evaluation frameworks to support the training and deployment of AI systems aimed at achieving scientific AGI.

Added May 22, 2026San Francisco, CAAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The Research Engineer will work on advancing AI models in secure coding and vulnerability remediation, blending research and engineering in the field of cybersecurity.

Added May 22, 2026San Francisco, CA | New York City, NYAI Safety & Alignment$300K - $405K / year
View details
Apply
Details
Apply

The Model Quality Software Engineer will set technical direction for evaluation systems and research infrastructure, focusing on improving AI model capabilities and ensuring safety in AI systems.

Added May 22, 2026San Francisco, CA | New York City, NYAI Safety & Alignment$405K - $485K / year
View details
Apply
Details
Apply

The ML/Research Engineer, Safeguards role at Anthropic focuses on developing systems to detect and mitigate misuse of AI, ensuring safety and compliance in AI systems.

Added May 22, 2026San Francisco, CA | New York City, NYAI Safety & Alignment$350K - $500K / year
View details
Apply