AboutTermsRefundsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login
All jobs

Curated roles

AI Safety Jobs

Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.

363 active roles found.

AI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
CM
Carnegie Mellon University, Software Engineering Institute
Associate AI Red Team Engineer
DetailsApply

AI red team engineer role focused on adversary emulation against AI-enabled systems, including the model and its surrounding hardware/software/network stack, to prepare defenders for real-world threats.

Added 6 days agoPittsburgh, United States of AmericaAI Safety & AlignmentSalary not disclosed
View details
Apply
OP
OpenAI
Red Team Specialist - Cyber
OP
OpenAINew
Model Policy Manager Model Policy San Francisco

Showing 1–20 of 363 roles

Previous
1
2
3
4
…19
Next
Previous

Page 1 of 19

Next

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

Cyber-focused AI red team role evaluating model capabilities, safeguards, and abuse risks in agentic systems, with direct responsibility for safety testing and mitigation recommendations.

Added 6 days agoSan Francisco, CA or Seattle, WAAI Safety & Alignment$198K - $320K / year
View details
Apply
Details
Apply

Senior Model Policy Manager role focused on creating and operationalizing policies, taxonomies, and evaluation criteria to keep frontier AI systems safe, especially for biological and chemical risk.

Added 3 days agoSan Francisco, CAAI Safety & Alignment$207K - $295K / year
View details
Apply
OF
OpenAI Foundation
Program Officer, AI Model Safety
DetailsApply

Founding grantmaker for the AI Model Safety program, funding independent evaluations, standards, and foundational safety research to improve frontier model safety.

Added 4 days agoSan Francisco, CA, USAI Safety & Alignment$220K - $300K / year
View details
Apply
CM
Carnegie Mellon University, Software Engineering Institute
Senior AI Red Team Engineer
DetailsApply

Senior AI red team role on SEI's AI Security team focused on adversary emulation against AI-enabled systems, developing offensive techniques to improve defender preparedness and AI system security.

Added 4 days agoPittsburgh, United States of AmericaAI Safety & AlignmentSalary not disclosed
View details
Apply
UO
University of Copenhagen, Department of Computer Science
Postdoc, Mechanistic Understanding of AI Reasoning
DetailsApply

Two-year postdoctoral research role on mechanistic interpretability and AI reasoning within an ERC-funded project on explainable and robust automatic fact checking. Relevant as adjacent AI safety work because it studies model internals and reasoning behavior.

Added 5 days agoCopenhagen, DenmarkAI Safety & AlignmentSalary not disclosed
View details
Apply
GD
Google DeepMind
Research Scientist, Safety Oversight
DetailsApply

Research scientist role on DeepMind's GenAI Safety team focused on production monitoring, automated evaluations, and misuse detection for deployed AI models.

Added 5 days agoSan Francisco Bay AreaAI Safety & Alignment$207K - $300K / year
View details
Apply
OP
OpenAI
Model Policy (Rodrigo) Model Policy San Francisco
DetailsApply

Senior Model Policy role at OpenAI focused on designing policies, taxonomies, and evaluation criteria to keep frontier models safe in biological and chemical dual-use scenarios.

Added 5 days agoSan Francisco, CAAI Safety & Alignment$207K - $295K / year
View details
Apply
MI
MiniMax
AI Security Researcher
DetailsApply

AI security researcher role focused on LLM security capabilities, benchmarks, risk identification, and security frameworks for model iteration.

Added 6 days agoBeijing, China, Shanghai, ChinaAI Safety & AlignmentSalary not disclosed
View details
Apply
1L
10a Labs
Software Engineer, Infrastructure & Platform
DetailsApply

Infrastructure/platform engineer building secure, reproducible systems for advanced AI evaluations, including autonomous model behavior, agentic systems, and loss-of-control risk evaluations.

Added 7 days agoRemoteRemote - WorldwideAI Safety & Alignment$110K - $160K / year
View details
Apply
FA
Faculty
Associate, Safety
DetailsApply

Associate role on Faculty’s AI Safety team supporting frontier model evaluations, AI safety red teaming, and delivery of responsible AI projects for government and industry clients.

Added 8 days agoLondon, UKAI Safety & AlignmentSalary not disclosed
View details
Apply
AL
Alice
GenAI Chemical, Biological, Radiological, Nuclear, and Explosives Cyber Expert
DetailsApply

Remote full-time GenAI red team role focused on adversarial testing of GenAI safety guardrails, jailbreaks, multi-turn attacks, and model evaluations for cyber-CBRNE threat scenarios.

Added 11 days agoRemoteRemote - WorldwideAI Safety & Alignment$150K - $178K / year
View details
Apply
AL
Alice
GenAI Biosecurity Expert
DetailsApply

Remote full-time role focused on evaluating and strengthening GenAI safety guardrails through biological risk assessment, adversarial red-teaming, and model safety alignment.

Added 11 days agoRemoteRemote - WorldwideAI Safety & Alignment$150K - $178K / year
View details
Apply
AN
Anthropic
Software Engineer, Infrastructure, Interpretability
DetailsApply

Infrastructure engineer for Anthropic's Interpretability team, building secure research environments, data systems, and compute tooling that support interpretability work tied to frontier AI safety and audit pipelines.

Added 12 days agoSan Francisco, CA | New York City, NYRemoteAI Safety & Alignment$320K - $485K / year
View details
Apply
AS
AI Security Institute
Research Scientist (Virologist), Chem-Bio
DetailsApply

Senior technical role designing and interpreting evaluations of frontier and specialised AI models for biologically relevant, virology-focused risk and misuse questions, with policy-relevant outputs for government.

Added 13 days agoLondon, UKAI Safety & AlignmentSalary not disclosed
View details
Apply
AS
AI Security Institute
Research Scientist (Biological Models), Chem-Bio
DetailsApply

Technical research role evaluating risk-relevant capabilities of specialised biological AI models, including misuse risks and technical safeguards, with outputs informing government decisions.

Added 13 days agoLondon, UKRemoteAI Safety & AlignmentSalary not disclosed
View details
Apply
OP
OpenAI
Technical Program Manager, AI Safety & Safeguards Technical
DetailsApply

Technical program management role driving AI safety and safeguards initiatives across deployment environments, including model evaluations, mitigations, abuse detection, monitoring, and readiness for high-impact deployments.

Added 13 days agoSan Francisco and Mountain ViewAI Safety & Alignment$257K - $445K / year
View details
Apply
MR
MATS Research
Neel Nanda Stream, MATS Program (Winter 2026)
DetailsApply

A paid MATS research program role involving a ~20 hour AI safety research project focused on pragmatic interpretability or applied safety, with a detailed write-up of findings.

Added 14 days agoSan Francisco Bay Area, London, UKAI Safety & Alignment$23.4K / year
View details
Apply
AN
Anthropic
Staff+ Site Reliability Engineer, Safeguards ML Infra
DetailsApply

Staff+ SRE role on Anthropic's Safeguards ML Infra team, responsible for production infrastructure, deployment, and validation of safety systems and safety classifiers for Claude model launches.

Added 14 days agoFriendly (Travel-Required) | San Francisco, CA | Seattle, WA | New York City, NYRemoteAI Safety & Alignment$405K - $485K / year
View details
Apply
OP
OpenAI
Researcher, Frontier Risk Mitigations
DetailsApply

Research role on OpenAI’s Preparedness team focused on frontier AI safety mitigations, evaluations, red-teaming, and alignment/interpretability methods to make deployed models safer.

Added 15 days agoSan Francisco, CAAI Safety & Alignment$295K - $445K / year
View details
Apply