AboutTermsRefundsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login
All jobs

Curated roles

AI Safety Jobs

Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.

364 active roles found.

AI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AN
Anthropic
Technical Program Manager, Safeguards (Infrastructure & Evals)
DetailsApply

The Technical Program Manager for Safeguards will oversee the operational health of AI safety infrastructure, manage incident responses, and ensure reliability of safety-critical systems.

Added May 22, 2026San Francisco, CA | New York City, NY | Seattle, WAAI Safety & Alignment$290K - $365K / year
View details
Apply
AN
Anthropic
Technical Program Manager, Research
AN
Anthropic
Staff + Sr. Software Engineer, AI Reliability
AN
Anthropic
Staff Software Engineer, Inference
AN
Anthropic
Staff Software Engineer, AI Reliability Engineering
AN
Anthropic
Staff Software Engineer, AI Reliability Engineering
AN
Anthropic
Staff Research Engineer, Discovery Team
AN
Anthropic
Software Engineer, Safeguards Infrastructure
AN
Anthropic
Software Engineer, Safeguards
AN
Anthropic
Senior+ Software Engineer, Research Tools
AN
Anthropic
Safeguards Enforcement Analyst, Safety Evaluations
AN
Anthropic
Research Scientist, Interpretability
AN
Anthropic
Research Engineer, Universes
AN
Anthropic
Research Engineer / Scientist, Frontier Red Team (Cyber)
AN
Anthropic
Research Engineer / Scientist, Alignment Science - London
AN
Anthropic
Research Engineer / Scientist, Alignment Science
AN
Anthropic
Research Engineer / Research Scientist, Tokens
AN
Anthropic
Research Engineer/Research Scientist, Pre-training
AN
Anthropic
Research Engineer / Research Scientist, Pre-training
AN
Anthropic
Research Engineer, Production Model Post-Training

Showing 321–340 of 364 roles

Previous
1
…16
17
18
19
Next
Previous

Page 17 of 19

Next

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

The Technical Program Manager for Research at Anthropic will define and build programs for research teams, focusing on AI safety, alignment, and evaluations.

Added May 22, 2026San Francisco, CA | New York City, NYAI Safety & Alignment$365K - $435K / year
View details
Apply
Details
Apply

The role focuses on enhancing the reliability of AI systems at Anthropic, ensuring safety commitments and robust performance across critical serving paths.

Added May 22, 2026San Francisco, CA | New York City, NY | Seattle, WAAI Safety & Alignment$325K - $485K / year
View details
Apply
Details
Apply

The Staff Software Engineer, Inference at Anthropic will work on building and maintaining systems for AI model inference, focusing on performance optimization and societal impacts.

Added May 22, 2026London, UKAI Safety & Alignment£325K - £390K / year
View details
Apply
Details
Apply

The Staff Software Engineer in AI Reliability Engineering will enhance the reliability of AI systems, focusing on critical service paths and incident response.

Added May 22, 2026London, UKAI Safety & Alignment£325K - £390K / year
View details
Apply
Details
Apply

The Staff Software Engineer in AI Reliability Engineering at Anthropic will focus on improving the reliability of AI systems, ensuring robust performance and safety commitments.

Added May 22, 2026Dublin, IEAI Safety & Alignment€235K - €295K / year
View details
Apply
Details
Apply

The Staff Research Engineer will work on developing AI systems aimed at scientific discovery, focusing on removing bottlenecks in achieving scientific AGI and creating evaluation frameworks for model capabilities.

Added May 22, 2026San Francisco, CAAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The role involves developing systems for safety and oversight of AI models, focusing on misuse prevention and monitoring model behaviors.

Added May 22, 2026London, UKAI Safety & Alignment£255K - £325K / year
View details
Apply
Details
Apply

The Software Engineer, Safeguards role at Anthropic involves developing systems to monitor AI models, prevent misuse, and ensure user safety.

Added May 22, 2026San Francisco, CA | New York City, NYAI Safety & Alignment$320K - $485K / year
View details
Apply
Details
Apply

The Senior+ Software Engineer for Research Tools at Anthropic will develop infrastructure and applications to support AI safety research, focusing on model evaluation and understanding model behavior.

Added May 22, 2026San Francisco, CA | New York City, NYAI Safety & Alignment$300K - $405K / year
View details
Apply
Details
Apply

The Safeguards Enforcement Analyst will ensure AI models meet safety and policy standards through evaluations and mitigations, collaborating with cross-functional teams.

Added May 22, 2026Friendly (Travel-Required) | San Francisco, CA | Washington, DC; San Francisco, CA | New York City, NYRemoteAI Safety & Alignment$230K - $270K / year
View details
Apply
Details
Apply

The Research Scientist, Interpretability role at Anthropic focuses on mechanistic interpretability to enhance the safety and understanding of AI systems.

Added May 22, 2026San Francisco, CARemoteAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The Research Engineer, Universes role at Anthropic focuses on developing training environments for safe AI systems and includes responsibilities for building evaluations to measure AI capabilities.

Added May 22, 2026Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NYRemoteAI Safety & Alignment$500K - $850K / year
View details
Apply
Details
Apply

Technical research role on Anthropic's Frontier Red Team focused on evaluating and defending against advanced AI-enabled cyber threats, including autonomous capability evaluations, safety defenses, and policy-relevant demonstrations.

Added May 22, 2026San Francisco, CAAI Safety & Alignment$320K - $485K / year
View details
Apply
Details
Apply

The role involves conducting research on AI safety and alignment, focusing on understanding and steering the behavior of powerful AI systems.

Added May 22, 2026London, UKAI Safety & Alignment£260K - £370K / year
View details
Apply
Details
Apply

Research Engineer/Scientist on Anthropic’s Alignment Science team, conducting experimental AI safety research on powerful future systems, safety evaluations, alignment stress-testing, and related safeguards work.

Added May 22, 2026Bay AreaAI Safety & Alignment$350K - $500K / year
View details
Apply
Details
Apply

The role focuses on building reliable and interpretable AI systems, emphasizing safety and societal impacts.

Added May 22, 2026New York City, NY; Seattle, WA; San Francisco, CAAI Safety & Alignment$350K - $500K / year
View details
Apply
Details
Apply

The Research Engineer/Research Scientist role at Anthropic focuses on developing large language models with an emphasis on safety, alignment, and societal impacts.

Added May 22, 2026Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NYRemoteAI Safety & Alignment$350K - $850K / year
View details
Apply
Details
Apply

The role involves research and engineering to develop safe and trustworthy large language models, focusing on multimodal capabilities and ethical implications of AI.

Added May 22, 2026Zürich, CHAI Safety & AlignmentCHF 280K - CHF 680K / year
View details
Apply
Details
Apply

The Research Engineer will enhance AI model safety and alignment through post-training techniques, impacting the quality and capabilities of production models.

Added May 22, 2026Zürich, CHAI Safety & Alignment
View details
Apply