AboutTermsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login
All jobs

Curated roles

AI Safety Jobs

Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.

393 active roles found.

AI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AN
Anthropic
Product Manager, Safeguards (Generalist)
DetailsApply

Product Manager role focused on Anthropic Safeguards systems for frontier models, including safety-by-design, safety evals, detections, interventions, and mitigation of deployment and user risks.

Added 17 days agoSan Francisco, CA | New York City, NYAI Safety & Alignment$385K - $460K / year
View details
Apply
PU
Princeton University, Laboratory for Artificial Intelligence
Research Fellow, AI Alignment and Safety
GD
Google DeepMind
Research Scientist, Safety Oversight
AN
Anthropic
Product Manager, Safeguards (Account Integrity & Abuse)
OP
OpenAI
Model Policy Manager, Multimodal Safety Model Policy San Francisco
BI
BlueDot Impact
Program Lead, Technical AI Safety Project Sprint
BO
Bosch
AI Research Scientist, Safety
Details
BG
Bosch Group
[EJV] Penetration Tester — AI Safety & Security
BI
BlueDot Impact
Program Lead, Technical AI Safety Course
AL
Alice
Senior Researcher, AI
Details
IF
Institute for Security and Technology
Senior Advisor, Advanced Guardrails for AI-Induced Incidents
CR
Cedar Research
Communications and Outreach
OP
OpenAI
Model Policy Manager, Agentic Safety Model Policy San Francisco
MA
Mathematical AI Safety Institute
Mathematicians (x10-100)
GR
German Research Center for Artificial Intelligence
Senior Researcher / Postdoc, Multilingual Mechanistic Interpretability
CG
Coefficient Giving
Project Tailwind, Call for Ambitious AI Safety Initiatives
AN
Anthropic
Staff+ Software Engineer, ML Inference Path
GR
German Research Center for Artificial Intelligence
Senior Researcher / Postdoc, AI Safety, Ethics, and Agentic Systems
GR
Geodesic Research
Member of Technical Staff
CA
Compassion Aligned Machine Learning
Research Engineer, Value Persistence Through Reinforcement Learning

Showing 21–40 of 393 roles

Previous
1
2
3
4
…20
Next
Previous

Page 2 of 20

Next

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

Research fellowship focused on AI alignment and safety research, including alignment challenges in language models and multimodal systems, with work on robustness, interpretability, and evaluation.

Added 18 days agoPrinceton, NJAI Safety & Alignment$100K / year
View details
Apply
Details
Apply

Research scientist role on DeepMind’s GenAI Safety team focused on production safety oversight for deployed models, including automated evaluations, misuse detection, and monitoring model alignment and misbehavior.

Added 18 days agoSan Francisco Bay AreaAI Safety & Alignment$207K - $300K / year
View details
Apply
Details
Apply

Product Manager role for Anthropic’s Safeguards team focused on building safety systems, safety evals, detections, and interventions to mitigate misuse and deployment risks for frontier AI products.

Added 20 days agoSan Francisco, CA | New York City, NYAI Safety & Alignment$385K - $460K / year
View details
Apply
Details
Apply

Manager role focused on multimodal model safety policy at OpenAI, designing behavioral safety policies, evaluation criteria, and safeguards for frontier AI models.

Added 20 days agoSan Francisco, CAAI Safety & Alignment$266K - $335K / year
View details
Apply
Details
Apply

Program lead for BlueDot Impact’s Technical AI Safety Project Sprint, owning project scoping, participant selection, review, and acceleration for newcomers doing technical AI safety projects.

Added 21 days agoLondon, United KingdomRemoteAI Safety & Alignment$160K - $250K / year
View details
Apply
Apply

Research scientist role focused on AI safety, alignment, and model reliability for autonomous driving and ADAS, including safety validation, edge-case evaluation, and robust system behavior.

Added 21 days agoSan Francisco Bay AreaAI Safety & Alignment$165K - $185K / year
View details
Apply
Details
Apply

Penetration tester on Bosch's AI Safety & Security team, performing safety evaluations and red-teaming of LLMs and agentic AI systems to find prompt injection, jailbreak, and tool-abuse issues.

Added 21 days agoHo Chi Minh, VietnamAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Lead BlueDot Impact’s technical AI safety course, owning curriculum, strategy, admissions, and participant acceleration for a program that trains people entering AI safety.

Added 22 days agoLondon, United KingdomRemoteAI Safety & Alignment$160K - $250K / year
View details
Apply
Apply

Senior AI researcher role focused on post-training evaluation, red-teaming, and RL gym audits for open-weight LLMs, with direct work on security capabilities and defensive alignment against indirect prompt injection.

Added 23 days agoTel Aviv, IsraelAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Senior advisory contractor role focused on advanced AI risks, loss-of-control scenarios, safeguards, incident prevention, and translating technical research into policy and operational guidance.

Added 24 days agoRemoteRemote - WorldwideAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Communications and outreach role at Cedar Research focused on promoting AI safety work through writing, video, outreach, and possibly helping run human studies.

Added 24 days agoLondon, UKRemoteAI Safety & Alignment$60K - $110K / year
View details
Apply
Details
Apply

Manager role focused on agentic safety and model policy for frontier AI systems, turning alignment and misalignment risks into behavioral policies, evaluations, monitoring, and safeguards.

Added 24 days agoSan Francisco, CAAI Safety & Alignment$207K - $335K / year
View details
Apply
Details
Apply

Mathematical research role focused on AI safety, including theoretical problems and novel mathematical approaches to safety challenges.

Added 29 days agoSan Francisco Bay AreaAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Postdoctoral researcher role on multilingual mechanistic interpretability, including circuit analysis, controlled validation with backdoored model suites, and work that feeds into safe agentic system design.

Added 29 days agoSaarbrücken, GermanyAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

A remote global funding call for founders to start new organizations tackling critical AI safety problems, including alignment moonshots and frontier capabilities work.

Added 29 days agoRemoteRemote - WorldwideAI Safety & Alignment$200K - $200M / year
View details
Apply
Details
Apply

Staff+ software engineer role building production ML infrastructure for Claude's safety systems, including safety deployments, monitoring, and productionizing safety research.

Added 29 days agoSan Francisco, CAAI Safety & Alignment$320K - $485K / year
View details
Apply
Details
Apply

Postdoctoral research role on safe, controllable agentic AI, combining runtime behavioral control, normative constraints, integrated evaluation, and publication in AI safety and formal methods venues.

Added Sep 9, 2026Saarbrücken, Germany, Darmstadt, GermanyAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Technical AI safety research role focused on AI alignment, robust initialization methods for capable language models, and empirical evaluation of alignment techniques.

Added Sep 9, 2026London, UK, Cambridge, UKAI Safety & Alignment£110K - £150K / year
View details
Apply
Details
Apply

Research engineer role studying whether values persist after reinforcement learning, using training pipelines, evaluations, and interpretability methods to analyze model behavior and alignment.

Added Sep 8, 2026RemoteRemote - WorldwideAI Safety & Alignment$9.2K / month
View details
Apply