AboutTermsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login
All jobs

Curated roles

AI Safety Jobs

Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.

393 active roles found.

AI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
KA
KairosNew
SPAR Research Manager
DetailsApply

Remote research management role mentoring AI safety fellowship mentors, supporting research leads, team management, and research quality across SPAR.

Added 3 days agoRemoteRemote - WorldwideAI Safety & Alignment$105K - $200K / year
View details
Apply
MA
Mistral AINew
Research Engineer - Eval Platform

Showing 1–20 of 393 roles

Previous
1
2
3
4
…20
Next
Previous

Page 1 of 20

Next

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

Research Engineer on the Eval Platform team building infrastructure for reproducible, scalable model evaluations and dashboards used by researchers to assess model quality.

Added yesterdayParisAI Safety & AlignmentSalary not disclosed
View details
Apply
SI
SimulaNew
Research Engineer, AI Safety Audit and Red Teaming
DetailsApply

Research engineer role focused on building and operating SimpleAudit for AI safety audits, red teaming, and safety evaluations of language models and AI systems.

Added 3 days agoOslo, NorwayAI Safety & AlignmentSalary not disclosed
View details
Apply
FA
FAR.AI
Head of Engineering (Red Team)
DetailsApply

Senior engineering leadership role building the engineering systems, tools, and team for FAR.AI’s frontier AI red-teaming program, including safety evaluations, attacker simulation, and harmfulness evaluation for frontier models.

Added 6 days agoInternationalRemoteAI Safety & Alignment$225K - $350K / year
View details
Apply
EU
European Union, Joint Research Centre
Project Officer, AI Cybersecurity Researcher
DetailsApply

Project Officer role focused on building a secure cyber-range to evaluate AI agents and frontier models in cybersecurity, including red-team/blue-team scenarios and vulnerability testing.

Added 7 days agoIspra, ItalyAI Safety & Alignment€4.4K - €6.4K / month
View details
Apply
GD
Google DeepMind
Senior Technical Program Manager, AGI Safety and Alignment
DetailsApply

Senior TPM role on DeepMind's AGI Safety and Alignment team, driving technical execution for Gemini safety, evaluation infrastructure, monitoring, and risk mitigation across the model lifecycle.

Added 7 days agoSan Francisco Bay Area, London, UKAI Safety & Alignment$256K - $278K / year
View details
Apply
GD
Google DeepMind
Research Scientist, AGI Safety and Alignment
DetailsApply

Research scientist role on DeepMind’s AGI Safety and Alignment Team focused on alignment methods, interpretability, AGI control systems, and frontier safety evaluations.

Added 8 days agoLondon, UKAI Safety & AlignmentSalary not disclosed
View details
Apply
AS
AI Security Institute
Workstream Lead, Agentic AI Risk Modelling and Mitigations
DetailsApply

Senior leadership role leading a research team on agentic AI risk modelling and mitigations, focused on advanced AI systems becoming hard to oversee, correct, or shut down and on turning analysis into practical safety interventions.

Added 10 days agoLondon, UKAI Safety & AlignmentSalary not disclosed
View details
Apply
FA
FAR.AI
Research Scientist, Applied White-Box Methods
DetailsApply

Research scientist role developing and evaluating white-box methods to improve AI system safety and alignment, with emphasis on realistic evaluations, agentic coding, and threat models.

Added 10 days agoBerkeley Office / USAI Safety & Alignment$150K - $250K / year
View details
Apply
GD
Google DeepMind
Technical Program Manager, GenAI Safety
DetailsApply

Technical Program Manager role leading GenAI Safety operations for Gemini and related models, coordinating training, evaluations, red teaming, and safety/alignment initiatives.

Added 11 days agoSan Francisco Bay AreaAI Safety & Alignment$256K - $278K / year
View details
Apply
FA
Faculty
Senior Manager (Safety)
DetailsApply

Senior manager role leading AI safety delivery, including frontier model evaluations, red teaming, safeguard testing, and client advisory work with government and frontier labs.

Added 11 days agoUK - LondonRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
AR
Alignment Research Center
Researcher
DetailsApply

Researcher role at ARC focused on AI alignment research, mechanistic understanding of neural networks, and developing training objectives that encourage honest internal reporting.

Added 13 days agoBerkeleyRemoteAI Safety & AlignmentFrom $300K / year
View details
Apply
AA
AI Alignment Foundation
Technical Grant Writer
DetailsApply

We are looking for an exceptional writer to help turn promising AI alignment research ideas into compelling, fundable proposals.

Added 14 days agoEnds in 17 daysBerkeley, CARemoteAI Safety & Alignment$110K - $150K / per year
View details
Apply
AA
AI Alignment Foundation
Partnerships Lead
DetailsApply

The Partnerships Lead at AIAF will be the bridge to the AI safety funding ecosystem, cultivating strong relationships with funders to update them on our work.

Added 14 days agoEnds in 17 daysBerkeley, CAAI Safety & Alignment$150K - $300K / per year
View details
Apply
LL
Lawrence Livermore National Laboratory
Postdoctoral Researcher, Explainable AI
DetailsApply

Postdoctoral researcher in explainable AI focused on interpreting deep model internals, human-in-the-loop workflows, and rigorous evaluation of interpretability claims; relevant because it explicitly connects interpretability to AI safety and assurance evaluation.

Added 15 days agoLivermore, CAAI Safety & Alignment$143.3K / year
View details
Apply
GD
Google DeepMind
Research Scientist, Gemini Safety and Behaviour
DetailsApply

Research scientist role on Gemini Safety and Behavior focused on LLM safety/security, jailbreak mitigation, red and blue teaming, and safety/alignment work for frontier GenAI models.

Added 15 days agoSan Francisco Bay Area, New York, NYAI Safety & Alignment$207K - $300K / year
View details
Apply
BG
Bosch Group
[EJV] Penetration Tester — AI Safety & Security (with Joining Bonus)
DetailsApply

Penetration tester for Bosch's AI Safety & Security team, running safety evaluations and red-teaming against LLMs and agentic AI systems to find harmful outputs, prompt injection, and tool abuse.

Added 15 days agoHo Chi Minh, VietnamAI Safety & AlignmentSalary not disclosed
View details
Apply
SA
Safe AI Forum
Researcher / Senior Researcher, Misuse Risks
DetailsApply

Senior researcher role focused on extreme misuse risks from frontier AI, including technical safeguards, consensus-building, and translating AI safety work into policy-relevant coordination outputs.

Added 16 days agoSan Francisco Bay Area, New York, NY, USA, UKRemoteAI Safety & Alignment$122.1K - $172.1K / year
View details
Apply
FA
FAR.AI
Research Scientist, Applied White-Box Methods
DetailsApply

Research scientist role on FAR.AI’s Applied White-Box Methods team, developing and evaluating model-internals methods to improve AI safety, alignment, and realistic safety evaluations.

Added 16 days agoBerkeley Office / USAI Safety & Alignment$150K - $250K / year
View details
Apply
AV
AI Verification and Evaluation Research Institute
Research Scientist
DetailsApply

Research Scientist role focused on frontier AI safety and security research, including AI system evaluations, audits, model behavior analysis, robustness, and policy compliance.

Added 17 days agoRemoteRemote - WorldwideAI Safety & Alignment$220K - $600K / year
View details
Apply