AboutTermsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login
All jobs

Curated roles

AI Safety Jobs

Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.

387 active roles found.

AI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
UO
University of Vienna, Faculty of Computer Science
PhD Position, Responsible Machine Learning
DetailsApply

Fully-funded PhD position researching responsible machine learning, with explicit focus on AI safety and robust alignment in multi-agent LLM systems.

Added 4 days agoVienna, AustriaAI Safety & AlignmentSalary not disclosed
View details
Apply
OU
Oxford University, Department of Engineering ScienceNew
Senior Researcher, Interpretability and AI Safety
CA
Compassion Aligned Machine Learning
Research Engineer, Value Persistence Through Reinforcement Learning
MA
Mathematical AI Safety InstituteNew
Mathematicians (x10-100)

Showing 1–20 of 387 roles

Previous
1
2
3
4
…20
Next
Previous

Page 1 of 20

Next

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

Senior research role in Oxford’s Technical AI Governance programme focused on interpretability, evaluations, and AI safety for continuously learning systems, with work on foundation-model experiments.

Added 3 days agoEnds in 10 daysOxford, England, United KingdomAI Safety & Alignment£49.1K - £58.3K / year
View details
Apply
Details
Apply

Research engineer role studying whether values persist after reinforcement learning, using training pipelines, evaluations, and interpretability methods to analyze model behavior and alignment.

Added 4 days agoRemoteRemote - WorldwideAI Safety & Alignment$9.2K / month
View details
Apply
Details
Apply

Mathematical research role focused on AI safety, including theoretical problems and novel mathematical approaches to safety challenges.

Added 2 days agoSan Francisco Bay AreaAI Safety & AlignmentSalary not disclosed
View details
Apply
GR
German Research Center for Artificial IntelligenceNew
Senior Researcher / Postdoc, Multilingual Mechanistic Interpretability
DetailsApply

Postdoctoral researcher role on multilingual mechanistic interpretability, including circuit analysis, controlled validation with backdoored model suites, and work that feeds into safe agentic system design.

Added 2 days agoSaarbrücken, GermanyAI Safety & AlignmentSalary not disclosed
View details
Apply
CG
Coefficient GivingNew
Project Tailwind, Call for Ambitious AI Safety Initiatives
DetailsApply

A remote global funding call for founders to start new organizations tackling critical AI safety problems, including alignment moonshots and frontier capabilities work.

Added 2 days agoRemoteRemote - WorldwideAI Safety & Alignment$200K - $200M / year
View details
Apply
AN
AnthropicNew
Staff+ Software Engineer, ML Inference Path
DetailsApply

Staff+ software engineer role building production ML infrastructure for Claude's safety systems, including safety deployments, monitoring, and productionizing safety research.

Added 2 days agoSan Francisco, CAAI Safety & Alignment$320K - $485K / year
View details
Apply
GR
German Research Center for Artificial IntelligenceNew
Senior Researcher / Postdoc, AI Safety, Ethics, and Agentic Systems
DetailsApply

Postdoctoral research role on safe, controllable agentic AI, combining runtime behavioral control, normative constraints, integrated evaluation, and publication in AI safety and formal methods venues.

Added 3 days agoSaarbrücken, Germany, Darmstadt, GermanyAI Safety & AlignmentSalary not disclosed
View details
Apply
GR
Geodesic ResearchNew
Member of Technical Staff
DetailsApply

Technical AI safety research role focused on AI alignment, robust initialization methods for capable language models, and empirical evaluation of alignment techniques.

Added 3 days agoLondon, UK, Cambridge, UKAI Safety & Alignment£110K - £150K / year
View details
Apply
PA
Parallax
Founding Member, Technical Staff
DetailsApply

Founding technical research role building white-box auditing methods and infrastructure for frontier AI evaluations, with direct focus on interpretability and safety-relevant failure prediction.

Added 4 days agoEurope, UK, London, UKRemoteAI Safety & Alignment$100K - $135K / year
View details
Apply
RE
Resolution
Research Scientist, Philosophy Program
DetailsApply

Research scientist role in Resolution’s philosophy program focused on AI alignment research, including conceptual analysis, empirical hypotheses, and evaluation methods for aligned ASI.

Added 4 days agoBerkeley, CaliforniaAI Safety & Alignment$236K - $930K / year
View details
Apply
AN
Anthropic
Staff+ Site Reliability Engineer, Safeguards ML Infra
DetailsApply

Senior SRE role on Anthropic's Safeguards ML Infra team, focused on deploying, verifying, and operating production safety infrastructure and safety classifiers for frontier model launches.

Added 7 days agoFriendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NYRemoteAI Safety & Alignment$320K - $485K / year
View details
Apply
AN
Anthropic
Product Designer, Evals & Prompts
DetailsApply

Product design role focused on building and maintaining LLM evaluation systems, prompt fixes, and test harnesses for Claude surfaces and model launches, with explicit safety-related evaluation scope.

Added 7 days agoSan Francisco, CAAI Safety & Alignment$305K - $385K / year
View details
Apply
CI
Canadian Institute for Advanced Research
Program Manager, Canadian AI Safety Institute
DetailsApply

Program manager for the Canadian AI Safety Institute research program, supporting AI safety initiatives, research partnerships, proposal calls, and related events.

Added 8 days agoToronto, CanadaAI Safety & AlignmentCA$86.7K - CA$102K / year
View details
Apply
OU
Oxford University, Oxford Martin School
Senior Researcher, Interpretability and AI Safety
DetailsApply

Senior research role focused on interpretability, evaluations, and AI safety for continuously learning foundation models, with some collaboration on AI governance research.

Added 8 days agoOxford, UKAI Safety & Alignment£49.1K - £58.3K / year
View details
Apply
FA
Faculty
Delivery Manager (Safety)
DetailsApply

Delivery Manager role in Faculty’s AI Safety team focused on frontier model evaluations, AI safety red teaming, and delivery of high-impact safety projects for clients, including government and frontier labs.

Added 8 days agoUK - LondonRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
OP
OpenAI
Researcher, Agent Safety, Oversight and System Mitigations
DetailsApply

Research role on agent safety, oversight, evaluations, red-teaming, and system-level mitigations for increasingly capable AI agents operating safely and autonomously.

Added 8 days agoSan Francisco, CAAI Safety & Alignment$380K - $500K / year
View details
Apply
NR
Neo Research
Lead Research Scientist
DetailsApply

Lead research role at an AI safety organization focused on frontier model risks, misalignment, loss of control, harmful manipulation, and rigorous model evaluation research.

Added 9 days agoRemoteRemote - WorldwideAI Safety & Alignment$200K - $250K / year
View details
Apply
KA
kairos-project.org
Research Manager, SPAR Kairos Support mentors and mentees in SPAR, our largest AI safety research fellowship.
DetailsApply

Research manager/generalist for SPAR, an AI safety research fellowship, supporting mentors and mentees, cohort programming, and program operations in the AI safety ecosystem.

Added 9 days agoRemoteRemote - WorldwideAI Safety & Alignment$105K - $200K / year
View details
Apply
AN
Anthropic
Research Manager, Biological Safety
DetailsApply

Hands-on management role leading Anthropic’s biological safety research engineering team, focused on frontier model evaluations, safety classifiers, red-teaming, and deployment safeguards to prevent catastrophic misuse.

Added 10 days agoSan Francisco, CAAI Safety & Alignment$405K - $485K / year
View details
Apply