AboutTermsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login
All jobs

Curated roles

AI Safety Jobs

Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.

393 active roles found.

AI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
CB
Cambridge Boston Alignment Initiative
Research Fellowship, AI Safety (Fall 2026)
DetailsApply

An intensive, fully-funded ten-week AI safety research fellowship in Cambridge focused on technical AI safety, governance, interpretability, formal verification, model evaluations, and risk management frameworks.

Added Aug 31, 2026Boston metro areaAI Safety & Alignment$15K / month
View details
Apply
AV
AI Verification and Evaluation Research Institute
Security Research Engineer
AN
Anthropic
Safeguards Enforcement Lead, Cyber Harms
TL
Trajectory Labs, PBC
AI Cyber Red Teamer
MY
Mycelium
Technical Advisor, Fine-Tuning and Evals
Details
AR
Apollo Research
AI Security Researcher
GD
Google DeepMind
Research Engineer, AGI Safety and Alignment
GD
Google DeepMind
Research Scientist, Frontier Safety Framework Risk Modelling and Governance
IL
Iliad
Fellowship (Fall 2026)
Details
GD
Google DeepMind
Research Scientist, Safety Oversight
GD
Google DeepMind
Research Engineer, AGI Safety and Alignment
BI
Biohub
Staff Research Scientist, AI Safety
Details
FA
FAR.AI
Engineering Manager (Red Team)
Details
OP
OpenAI
Model Policy Manager Model Policy San Francisco
OF
OpenAI Foundation
Program Officer, AI Model Safety
CM
Carnegie Mellon University, Software Engineering Institute
Senior AI Red Team Engineer
UO
University of Copenhagen, Department of Computer Science
Postdoc, Mechanistic Understanding of AI Reasoning
GD
Google DeepMind
Research Scientist, Safety Oversight
OP
OpenAI
Model Policy (Rodrigo) Model Policy San Francisco
MI
MiniMax
AI Security Researcher
Details

Showing 61–80 of 393 roles

Previous
1
…3
4
5
…20
Next
Previous

Page 4 of 20

Next

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

Security research role focused on AI auditing methodologies, AI system vulnerabilities, and tools/frameworks for evaluating AI safety and security.

Added Aug 31, 2026USA, Mexico, CanadaRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Lead enforcement operations to detect and mitigate misuse of Anthropic AI systems for cyberattacks, malware, and offensive exploitation, including managing a team and shaping enforcement strategy.

Added Aug 29, 2026Friendly (Travel-Required) | Washington, DC; San Francisco, CA | New York City, NYRemoteAI Safety & Alignment$285K - $330K / year
View details
Apply
Details
Apply

Frontier AI red teaming role focused on attacking cyber safeguards, designing attack strategies, and feeding results into safety/alignment evaluations and deployment decisions.

Added Aug 29, 2026RemoteRemote - WorldwideAI Safety & AlignmentSalary not disclosed
View details
Apply
Apply

Part-time technical advisor role for AI safety research on fine-tuning and evaluations, advising on benchmark design, generalization testing, and validity of welfare/capability evals for open-weight models.

Added Aug 29, 2026San Francisco Bay AreaRemoteAI Safety & Alignment$55 - $90 / hour
View details
Apply
Details
Apply

Technical security researcher role focused on red-teaming AI agent systems, securing agent sandboxes and access controls, and building monitoring and controls to reduce catastrophic risks from coding agents.

Added Aug 29, 2026London & San FranciscoAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Research engineer role on DeepMind’s AGI Safety and Alignment Team, focused on alignment methods, adversarially robust control systems, interpretability, and frontier-model safety work.

Added Aug 28, 2026London, UKAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Research scientist role leading Frontier Safety Framework governance research and risk modeling for frontier AI models, including safety evaluations, mitigation assessment, and external safety reporting.

Added Aug 26, 2026New York, NY, London, UK, San Francisco Bay AreaAI Safety & Alignment$207K - $300K / year
View details
Apply
Apply

A 3-month full-time research fellowship in applied mathematics for AI safety, centered on technical AI alignment research and producing a research proposal.

Added Aug 25, 2026San Francisco Bay Area, London, UKAI Safety & Alignment$6K / month
View details
Apply
Details
Apply

Research scientist role focused on monitoring deployed GenAI models for safety, alignment, misuse, and coordinated harms using automated evaluations, classifiers, and large-scale production data.

Added Aug 25, 2026San Francisco Bay AreaAI Safety & Alignment$207K - $300K / year
View details
Apply
Details
Apply

Research engineer role on DeepMind's AGI Safety and Alignment Team focused on alignment methods, adversarially robust AGI control systems, and interpretability for frontier models.

Added Aug 25, 2026San Francisco Bay Area, New York, NYAI Safety & Alignment$174K - $252K / year
View details
Apply
Apply

Staff research scientist role focused on AI safety for biology, including safety evaluations, misuse-resistant safeguards, and frontier safety policy development.

Added Aug 25, 2026New York, NYAI Safety & Alignment$241K - $301K / year
View details
Apply
Apply

Engineering manager for FAR.AI’s red team, leading engineering systems for frontier AI red-teaming, safety evaluations, attacker simulation, and safeguard testing.

Added Aug 25, 2026InternationalRemoteAI Safety & Alignment$170K - $250K / year
View details
Apply
Details
Apply

Senior Model Policy Manager role focused on creating and operationalizing policies, taxonomies, and evaluation criteria to keep frontier AI systems safe, especially for biological and chemical risk.

Added Aug 22, 2026San Francisco, CAAI Safety & Alignment$207K - $295K / year
View details
Apply
Details
Apply

Founding grantmaker for the AI Model Safety program, funding independent evaluations, standards, and foundational safety research to improve frontier model safety.

Added Aug 21, 2026San Francisco, CA, USAI Safety & Alignment$220K - $300K / year
View details
Apply
Details
Apply

Senior AI red team role on SEI's AI Security team focused on adversary emulation against AI-enabled systems, developing offensive techniques to improve defender preparedness and AI system security.

Added Aug 21, 2026Pittsburgh, United States of AmericaAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Two-year postdoctoral research role on mechanistic interpretability and AI reasoning within an ERC-funded project on explainable and robust automatic fact checking. Relevant as adjacent AI safety work because it studies model internals and reasoning behavior.

Added Aug 20, 2026Copenhagen, DenmarkAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Research scientist role on DeepMind's GenAI Safety team focused on production monitoring, automated evaluations, and misuse detection for deployed AI models.

Added Aug 20, 2026San Francisco Bay AreaAI Safety & Alignment$207K - $300K / year
View details
Apply
Details
Apply

Senior Model Policy role at OpenAI focused on designing policies, taxonomies, and evaluation criteria to keep frontier models safe in biological and chemical dual-use scenarios.

Added Aug 20, 2026San Francisco, CAAI Safety & Alignment$207K - $295K / year
View details
Apply
Apply

AI security researcher role focused on LLM security capabilities, benchmarks, risk identification, and security frameworks for model iteration.

Added Aug 19, 2026Beijing, China, Shanghai, ChinaAI Safety & AlignmentSalary not disclosed
View details
Apply