AboutTermsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login
All jobs

Curated roles

AI Safety Jobs

Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.

393 active roles found.

AI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
UO
University of Vienna, Faculty of Computer Science
PhD Position, Responsible Machine Learning
DetailsApply

Fully-funded PhD position researching responsible machine learning, with explicit focus on AI safety and robust alignment in multi-agent LLM systems.

Added Sep 8, 2026Vienna, AustriaAI Safety & AlignmentSalary not disclosed
View details
Apply
RE
Resolution
Research Scientist, Philosophy Program
AN
Anthropic
Staff+ Site Reliability Engineer, Safeguards ML Infra
CI
Canadian Institute for Advanced Research
Program Manager, Canadian AI Safety Institute
FA
Faculty
Delivery Manager (Safety)
Details
OP
OpenAI
Researcher, Agent Safety, Oversight and System Mitigations
NR
Neo Research
Lead Research Scientist
KA
kairos-project.org
Research Manager, SPAR Kairos Support mentors and mentees in SPAR, our largest AI safety research fellowship.
AN
Anthropic
Research Manager, Biological Safety
KA
Kairos
Head of Special Projects
Details
KA
Kairos
Head of Kairos Labs
Details
KA
Kairos
Head of Groups
Details
KA
Kairos
Head of Generator
Details
KA
Kairos
Generalist (Groups)
Details
OP
OpenAI
Researcher, Alignment Interpretability Alignment San Francisco
AN
Anthropic
Safeguards Enforcement Analyst, Conventional Weapons
AN
Anthropic
Cyber Evaluations Engineer
AR
Apollo Research
Research Scientist (Control)
AR
Apollo Research
AI Security & Control Researcher
AR
Apollo Research
AI Red Team Engineer

Showing 41–60 of 393 roles

Previous
1
2
3
4
…20
Next
Previous

Page 3 of 20

Next

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

Research scientist role in Resolution’s philosophy program focused on AI alignment research, including conceptual analysis, empirical hypotheses, and evaluation methods for aligned ASI.

Added Sep 8, 2026Berkeley, CaliforniaAI Safety & Alignment$236K - $930K / year
View details
Apply
Details
Apply

Senior SRE role on Anthropic's Safeguards ML Infra team, focused on deploying, verifying, and operating production safety infrastructure and safety classifiers for frontier model launches.

Added Sep 5, 2026Friendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NYRemoteAI Safety & Alignment$320K - $485K / year
View details
Apply
Details
Apply

Program manager for the Canadian AI Safety Institute research program, supporting AI safety initiatives, research partnerships, proposal calls, and related events.

Added Sep 4, 2026Toronto, CanadaAI Safety & AlignmentCA$86.7K - CA$102K / year
View details
Apply
Apply

Delivery Manager role in Faculty’s AI Safety team focused on frontier model evaluations, AI safety red teaming, and delivery of high-impact safety projects for clients, including government and frontier labs.

Added Sep 4, 2026UK - LondonRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Research role on agent safety, oversight, evaluations, red-teaming, and system-level mitigations for increasingly capable AI agents operating safely and autonomously.

Added Sep 4, 2026San Francisco, CAAI Safety & Alignment$380K - $500K / year
View details
Apply
Details
Apply

Lead research role at an AI safety organization focused on frontier model risks, misalignment, loss of control, harmful manipulation, and rigorous model evaluation research.

Added Sep 3, 2026RemoteRemote - WorldwideAI Safety & Alignment$200K - $250K / year
View details
Apply
Details
Apply

Research manager/generalist for SPAR, an AI safety research fellowship, supporting mentors and mentees, cohort programming, and program operations in the AI safety ecosystem.

Added Sep 3, 2026RemoteRemote - WorldwideAI Safety & Alignment$105K - $200K / year
View details
Apply
Details
Apply

Hands-on management role leading Anthropic’s biological safety research engineering team, focused on frontier model evaluations, safety classifiers, red-teaming, and deployment safeguards to prevent catastrophic misuse.

Added Sep 2, 2026San Francisco, CAAI Safety & Alignment$405K - $485K / year
View details
Apply
Apply

Leads a portfolio of new Kairos programs and ecosystem infrastructure projects aimed at advancing AI safety and reducing risks from advanced AI.

Added Sep 2, 2026RemoteRemote - WorldwideAI Safety & Alignment$140K - $270K / year
View details
Apply
Apply

Founding lead for Kairos Labs, an incubator that will design, launch, and support new AI safety organizations and projects.

Added Sep 2, 2026RemoteRemote - WorldwideAI Safety & Alignment$140K - $270K / year
View details
Apply
Apply

Lead Kairos’s AI safety university groups portfolio, setting strategy, launching new initiatives, and managing the team supporting AI safety fieldbuilding.

Added Sep 2, 2026RemoteRemote - WorldwideAI Safety & Alignment$140K - $270K / year
View details
Apply
Apply

Lead and grow an in-person AI safety residency program for generalists, overseeing recruiting, project matching, advisor coordination, and program strategy for the AI safety ecosystem.

Added Sep 2, 2026Berkeley, CAAI Safety & Alignment$140K - $270K / year
View details
Apply
Apply

Generalist role supporting Kairos’s AI safety university groups program, including the Pathfinder Fellowship, organizer support, events, and program infrastructure.

Added Sep 2, 2026RemoteRemote - WorldwideAI Safety & Alignment$105K - $200K / year
View details
Apply
Details
Apply

Research role focused on mechanistic interpretability and understanding model internals to improve alignment and safety of powerful AI systems.

Added Sep 2, 2026San Francisco, CAAI Safety & Alignment$295K - $500K / year
View details
Apply
Details
Apply

Analyst role focused on enforcing safeguards against misuse of Anthropic's AI systems for conventional weapons and dangerous technology, including model behavior assessment, enforcement workflows, and evals.

Added Sep 2, 2026New York City, NY;-Friendly (Travel-Required) | San Francisco, CA | Washington, DCRemoteAI Safety & Alignment$245K - $330K / year
View details
Apply
Details
Apply

Engineer role focused on cyber-relevant model evaluations, safeguard robustness testing, and misuse detection for frontier AI systems.

Added Sep 2, 2026San Francisco, CA | Washington, DCRemoteAI Safety & Alignment$300K - $405K / year
View details
Apply
Details
Apply

Research scientist role on Apollo Research’s AI control and monitoring team, designing control protocols, evaluation frameworks, and monitors to reduce risks from AI systems and coding agents.

Added Sep 1, 2026London & San FranciscoAI Safety & Alignment$100K - $385K / year
View details
Apply
Details
Apply

Security and control researcher for coding agents, focused on AI threat modeling, red-teaming, and improving monitors and control protocols to reduce catastrophic risks from misaligned or compromised agents.

Added Sep 1, 2026London & San FranciscoAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Dedicated AI red team engineer role focused on red-teaming AI monitors, finding attack surfaces, and improving safety monitoring for coding agents and frontier lab systems.

Added Sep 1, 2026London & San FranciscoAI Safety & AlignmentSalary not disclosed
View details
Apply