AboutTermsRefundsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login
All jobs

Curated roles

AI Safety Jobs

Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.

363 active roles found.

AI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
GO
google.com
Senior Security Engineer, Agentic Red Team Google DeepMind
DetailsApply

Senior security engineer role on DeepMind's Agentic Red Team focused on adversarial testing of AI agents, prompt injection, exploit development, and automated red-teaming frameworks for model safety.

Added Jul 7, 2026Mountain View, CA, USA; New York, NY, USA; Zürich, SwitzerlandAI Safety & Alignment$174K - $253K / year
View details
Apply
TU
Technical University of Darmstadt, Department of Computer Science
Doctoral Researchers, Natural Language Processing and AI, Ubiquitous Knowledge Processing Lab
NR
Neo Research
Research Engineer, Contract
EA
Epoch AI
Researcher, Benchmark Reviews
Details
ME
Meridian
Senior Technical AI Safety Research Manager
AL
AIXI Labs
Machine Learning Research Scientist
Details
JO
jobs.ashbyhq.com
Member of Technical Staff, Security / Engineering / Research SL5 Task Force
AR
Apollo Research
Head of Communications & PR
CO
Cohere
Senior Member of Technical Staff, Safety and Security for Agents
RE
Resolution
Research Scientist
Details
RE
Resolution
Research Engineer
Details
CO
Cohere
Senior Research Engineer - Safety Tooling and Data
EA
Epoch AI
Researcher, Evaluations
Details
SA
SaferAI
Risk Modeling Lead
Details
CO
Cohere
Senior Research Engineer, Model Evaluation
CO
Cohere
Product Manager, Safety Research
Details
CO
Cohere
Senior Research Scientist, Model Evaluation
CO
Cohere
Senior Member of Technical Staff, Safety and Security for Agents
CO
Cohere
Data Annotation Specialist, Safety
Details
GD
Google DeepMind
Research Scientist, Multimodal Alignment, Safety, and Fairness

Showing 101–120 of 363 roles

Previous
1
…5
6
7
…19
Next
Previous

Page 6 of 19

Next

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

Research positions in NLP and AI with a strong emphasis on trustworthy and safe AI, including agent reliability, evaluation science, interpretability, red-teaming, and robustness.

Added Jul 6, 2026Darmstadt, GermanyAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Contract research engineer role focused on frontier AI safety evaluations, dangerous behavior elicitation, and safety report writing for loss-of-control and harmful manipulation risks.

Added Jul 3, 2026Singapore / LondonRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Apply

Remote researcher role producing public reviews of AI benchmarks, evaluating methodologies and implications for AI capabilities; adjacent to AI safety via model evaluations and capability assessment.

Added Jul 3, 2026RemoteRemote - WorldwideAI Safety & Alignment$100K - $200K / year
View details
Apply
Details
Apply

Senior research management role supporting AI safety programmes, including scoping and reviewing research, mentoring researchers, and running red-teaming and proposal development sessions focused on reducing AI risk.

Added Jul 2, 2026Cambridge, UKAI Safety & Alignment£65K - £100K / year
View details
Apply
Apply

Research scientist role at a non-profit AI safety lab focused on theoretical and empirical work on LLM-based agents, loss-of-control risks, and safety mitigations.

Added Jul 2, 2026London, GBAI Safety & Alignment£120K - £160K / year
View details
Apply
Details
Apply

Senior technical role building and securing frontier AI datacenter infrastructure, including threat modeling and red-teaming against nation-state adversaries.

Added Jul 2, 2026Bay AreaAI Safety & Alignment$200K - $350K / year
View details
Apply
Details
Apply

Senior communications leader for an AI safety research organization, responsible for external messaging, media strategy, and translating frontier AI safety research into public-facing materials.

Added Jul 1, 2026San Francisco, New York, Washington DC, LondonRemoteAI Safety & Alignment$150K - $240K / year
View details
Apply
Details
Apply

Senior technical role on Cohere’s Safety for Agents team focused on data generation, post-training algorithms, and evaluation methods to improve safety, trustworthiness, and security of LLMs and agentic models.

Added Jul 1, 2026London, Edinburgh, Paris, Toronto, New York;-friendlyRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Apply

Research scientist role at Resolution focused on technical AI alignment research, including empirical and theoretical work on scalable oversight and related alignment problems.

Added Jul 1, 2026Berkeley, CaliforniaRemoteAI Safety & Alignment$141K - $930K / year
View details
Apply
Apply

Research Engineer role at an ASI alignment lab, building research automation and evaluation infrastructure to support alignment research and frontier-model experiments.

Added Jul 1, 2026Berkeley, CaliforniaRemoteAI Safety & Alignment$141K - $930K / year
View details
Apply
Details
Apply

Senior research engineer role on Cohere’s Safety / Modelling Safety and Trust team building data tooling and pipelines for training and evaluation data to support safer, more reliable models.

Added Jul 1, 2026UK, Europe, or ET timezone restrictions; offices in London, Edinburgh, Paris, Toronto, Montreal, New YorkRemoteAI Safety & Alignment$230K - $535K / year
View details
Apply
Apply

Research role focused on evaluating frontier AI models on real-world tasks, building benchmarks and rubrics, and analyzing model performance; directly relevant to AI evaluations and safety-adjacent capability assessment.

Added Jun 30, 2026RemoteRemote - WorldwideAI Safety & Alignment$115K - $200K / year
View details
Apply
Apply

Lead research on frontier AI risk modeling across cyber, CBRN, and loss of control, setting methodological standards that inform safety cases, evaluations, mitigations, and AI governance.

Added Jun 30, 2026Paris and London; open to San Francisco;possibleRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Senior research engineering role focused on building evaluation methods, benchmarks, datasets, and infrastructure for measuring frontier LLM capabilities; relevant because it directly concerns model evaluations for advanced AI systems.

Added Jun 30, 2026Toronto, San Francisco, New York City, London, Paris, Montreal, and moreRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Apply

Product management role bridging Cohere's safety research and North product, translating model evaluations and red-teaming findings into safety features, guardrails, and evaluation frameworks.

Added Jun 30, 2026orNew York, Toronto, London, Paris, ZurichRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Senior research role focused on creating next-generation evaluation methods, benchmarks, and infrastructure to measure LLM progress and model capabilities.

Added Jun 30, 2026Toronto, San Francisco, New York City, London, Paris, Montreal, and moreRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Senior technical staff role on Cohere’s Safety for Agents team focused on data generation, post-training algorithms, and evaluation methods to improve safety and trustworthiness of LLMs and agents.

Added Jun 30, 2026Toronto, San Francisco, London, Paris, New York City, Montreal, Seoul, Germany;-friendlyRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Apply

Part-time remote contractor role annotating, auditing, and red-teaming LLM outputs to improve model safety, policy alignment, and prevention of unsafe or adversarial outputs.

Added Jun 30, 2026Canada or the United StatesRemoteAI Safety & Alignment$40K - $45K / hour
View details
Apply
Details
Apply

Research Scientist role in Google DeepMind's Frontier AI unit focused on multimodal safety research, AI alignment, behavior assessment, and steering evolving AI systems.

Added Jun 30, 2026Kirkland, Washington, US; Mountain View, California, US; New York City, New York, USAI Safety & Alignment$147K - $211K / year
View details
Apply