AboutTermsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login
All jobs

Curated roles

AI Safety Jobs

Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.

393 active roles found.

AI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
GO
google.com
Senior Security Engineer, Agentic Red Team Google DeepMind
DetailsApply

Senior security engineer role on DeepMind's Agentic Red Team focused on adversarial testing of AI agents, prompt injection, exploit development, and automated red-teaming frameworks for model safety.

Added Jul 7, 2026Mountain View, CA, USA; New York, NY, USA; Zürich, SwitzerlandAI Safety & Alignment$174K - $253K / year
View details
Apply
TU
Technical University of Darmstadt, Department of Computer Science
Doctoral Researchers, Natural Language Processing and AI, Ubiquitous Knowledge Processing Lab
EA
Epoch AI
Researcher, Benchmark Reviews
Details
ME
Meridian
Senior Technical AI Safety Research Manager
AL
AIXI Labs
Machine Learning Research Scientist
Details
JO
jobs.ashbyhq.com
Member of Technical Staff, Security / Engineering / Research SL5 Task Force
CO
Cohere
Senior Member of Technical Staff, Safety and Security for Agents
RE
Resolution
Research Scientist
Details
RE
Resolution
Research Engineer
Details
CO
Cohere
Senior Research Engineer - Safety Tooling and Data
EA
Epoch AI
Researcher, Evaluations
Details
SA
SaferAI
Risk Modeling Lead
Details
CO
Cohere
Senior Research Engineer, Model Evaluation
CO
Cohere
Product Manager, Safety Research
Details
CO
Cohere
Senior Research Scientist, Model Evaluation
CO
Cohere
Senior Member of Technical Staff, Safety and Security for Agents
CO
Cohere
Data Annotation Specialist, Safety
Details
ME
METR
Member of Technical Staff, Evaluation Execution
Details
FA
FAR AI
Senior Programs and Strategy Manager
Details
AL
alignerr.com
AI / Emerging Technology Security Analyst Alignerr

Showing 161–180 of 393 roles

Previous
1
…8
9
10
…20
Next
Previous

Page 9 of 20

Next

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

Research positions in NLP and AI with a strong emphasis on trustworthy and safe AI, including agent reliability, evaluation science, interpretability, red-teaming, and robustness.

Added Jul 6, 2026Darmstadt, GermanyAI Safety & AlignmentSalary not disclosed
View details
Apply
Apply

Remote researcher role producing public reviews of AI benchmarks, evaluating methodologies and implications for AI capabilities; adjacent to AI safety via model evaluations and capability assessment.

Added Jul 3, 2026RemoteRemote - WorldwideAI Safety & Alignment$100K - $200K / year
View details
Apply
Details
Apply

Senior research management role supporting AI safety programmes, including scoping and reviewing research, mentoring researchers, and running red-teaming and proposal development sessions focused on reducing AI risk.

Added Jul 2, 2026Cambridge, UKAI Safety & Alignment£65K - £100K / year
View details
Apply
Apply

Research scientist role at a non-profit AI safety lab focused on theoretical and empirical work on LLM-based agents, loss-of-control risks, and safety mitigations.

Added Jul 2, 2026London, GBAI Safety & Alignment£120K - £160K / year
View details
Apply
Details
Apply

Senior technical role building and securing frontier AI datacenter infrastructure, including threat modeling and red-teaming against nation-state adversaries.

Added Jul 2, 2026Bay AreaAI Safety & Alignment$200K - $350K / year
View details
Apply
Details
Apply

Senior technical role on Cohere’s Safety for Agents team focused on data generation, post-training algorithms, and evaluation methods to improve safety, trustworthiness, and security of LLMs and agentic models.

Added Jul 1, 2026London, Edinburgh, Paris, Toronto, New York;-friendlyRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Apply

Research scientist role at Resolution focused on technical AI alignment research, including empirical and theoretical work on scalable oversight and related alignment problems.

Added Jul 1, 2026Berkeley, CaliforniaRemoteAI Safety & Alignment$141K - $930K / year
View details
Apply
Apply

Research Engineer role at an ASI alignment lab, building research automation and evaluation infrastructure to support alignment research and frontier-model experiments.

Added Jul 1, 2026Berkeley, CaliforniaRemoteAI Safety & Alignment$141K - $930K / year
View details
Apply
Details
Apply

Senior research engineer role on Cohere’s Safety / Modelling Safety and Trust team building data tooling and pipelines for training and evaluation data to support safer, more reliable models.

Added Jul 1, 2026UK, Europe, or ET timezone restrictions; offices in London, Edinburgh, Paris, Toronto, Montreal, New YorkRemoteAI Safety & Alignment$230K - $535K / year
View details
Apply
Apply

Research role focused on evaluating frontier AI models on real-world tasks, building benchmarks and rubrics, and analyzing model performance; directly relevant to AI evaluations and safety-adjacent capability assessment.

Added Jun 30, 2026RemoteRemote - WorldwideAI Safety & Alignment$115K - $200K / year
View details
Apply
Apply

Lead research on frontier AI risk modeling across cyber, CBRN, and loss of control, setting methodological standards that inform safety cases, evaluations, mitigations, and AI governance.

Added Jun 30, 2026Paris and London; open to San Francisco;possibleRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Senior research engineering role focused on building evaluation methods, benchmarks, datasets, and infrastructure for measuring frontier LLM capabilities; relevant because it directly concerns model evaluations for advanced AI systems.

Added Jun 30, 2026Toronto, San Francisco, New York City, London, Paris, Montreal, and moreRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Apply

Product management role bridging Cohere's safety research and North product, translating model evaluations and red-teaming findings into safety features, guardrails, and evaluation frameworks.

Added Jun 30, 2026orNew York, Toronto, London, Paris, ZurichRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Senior research role focused on creating next-generation evaluation methods, benchmarks, and infrastructure to measure LLM progress and model capabilities.

Added Jun 30, 2026Toronto, San Francisco, New York City, London, Paris, Montreal, and moreRemoteAI Safety & AlignmentCA$250K - CA$535K / year
View details
Apply
Details
Apply

Senior technical staff role on Cohere’s Safety for Agents team focused on data generation, post-training algorithms, and evaluation methods to improve safety and trustworthiness of LLMs and agents.

Added Jun 30, 2026Toronto, San Francisco, London, Paris, New York City, Montreal, Seoul, Germany;-friendlyRemoteAI Safety & AlignmentCA$250K - CA$535K / year
View details
Apply
Apply

Part-time remote contractor role annotating, auditing, and red-teaming LLM outputs to improve model safety, policy alignment, and prevention of unsafe or adversarial outputs.

Added Jun 30, 2026Canada or the United StatesRemoteAI Safety & Alignment$40K - $45K / hour
View details
Apply
Apply

Technical role at METR focused on executing and scaling AI evaluations for capabilities, risks, mitigations, autonomy, and alignment.

Added Jun 29, 2026BerkeleyAI Safety & Alignment$285.5K - $503.1K / year
View details
Apply
Apply

Senior programs role at FAR.AI focused on designing and curating AI safety events, convenings, and field-building programming across the AI safety ecosystem.

Added Jun 29, 2026Berkeley, CA / US onlyRemoteAI Safety & Alignment$115K - $175K / year
View details
Apply
Details
Apply

Remote hourly contract role analyzing AI and LLM security scenarios, probing frontier models for vulnerabilities, misuse, and adversarial threats; relevant because it directly involves AI system safety, red-teaming, and misuse prevention.

Added Jun 27, 2026RemoteRemote - WorldwideAI Safety & Alignment$40 - $60 / hour
View details
Apply