AboutTermsRefundsPrivacyCookiesContactPost a jobSaved jobs
AI Safety JobsAI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
AI Safety CareersAI Safety CareersCurated jobs in AI safety, governance and frontier AI.
Post a jobEmployer login
All jobs

Curated roles

AI Safety Jobs

Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.

363 active roles found.

AI Governance JobsAI Policy JobsAI Compliance JobsAI Red Teaming JobsRemote AI Safety Jobs
GD
Google DeepMind
Research Scientist, Gemini Safety
DetailsApply

Research Scientist role on the Gemini Safety team focused on improving safety, fairness, adversarial robustness, and evaluation protocols for Gemini models.

Added Jun 30, 2026Mountain View, California, USAI Safety & AlignmentSalary not disclosed
View details
Apply
GD
Google DeepMind
Research Engineer, Human Understanding
ME
METR
Member of Technical Staff, Evaluation Execution
Details
FA
FAR AI
Senior Programs and Strategy Manager
Details
AL
alignerr.com
AI / Emerging Technology Security Analyst Alignerr
OP
OpenAI
Agent Post-Training, Frontier Evals and Environments Research
SA
Singapore AI Safety Hub
Singapore AI Safety Fellowship
OP
OpenAI
Safety Transparency Editor, Safety Systems
BI
BlueDot Impact
Technical AI Safety Project Sprint
FA
FAR AI
Jailbreaking Lead, Red Team
Details
FA
FAR AI
Senior Research Engineer
Details
0L
0labs.ai
AI Security Research Engineer 0Labs
GS
Gray Swan
Machine Learning Researcher
CA
careers.peopleclick.com
Research Scientist Massachusetts Institute of Technology, Computer Science and Artificial Intelligence Laboratory
IR
Irregular
Technical Policy Researcher
IR
Irregular
Cyber Researcher
Details
IR
Irregular
Research Engineer
Details
FA
FAR AI
Technical Project Manager, Red Team
AL
Alignerr
AI Red Team Analyst
Details
ER
ERA
Research Manager / Research Managers, AIxCyber
Details

Showing 121–140 of 363 roles

Previous
1
…6
7
8
…19
Next
Previous

Page 7 of 19

Next

Get weekly AI safety roles

A weekly digest of AI safety, governance, policy and responsible AI roles.

By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.

Details
Apply

Research Engineer in Google DeepMind's Frontier AI unit focused on multimodal human understanding and building defenses against misuse, deepfakes, and impersonation.

Added Jun 30, 2026Los Angeles, California, US; Mountain View, California, USAI Safety & Alignment$174K - $252K / year
View details
Apply
Apply

Technical role at METR focused on executing and scaling AI evaluations for capabilities, risks, mitigations, autonomy, and alignment.

Added Jun 29, 2026BerkeleyAI Safety & Alignment$285.5K - $503.1K / year
View details
Apply
Apply

Senior programs role at FAR.AI focused on designing and curating AI safety events, convenings, and field-building programming across the AI safety ecosystem.

Added Jun 29, 2026Berkeley, CA / US onlyRemoteAI Safety & Alignment$115K - $175K / year
View details
Apply
Details
Apply

Remote hourly contract role analyzing AI and LLM security scenarios, probing frontier models for vulnerabilities, misuse, and adversarial threats; relevant because it directly involves AI system safety, red-teaming, and misuse prevention.

Added Jun 27, 2026RemoteRemote - WorldwideAI Safety & Alignment$40 - $60 / hour
View details
Apply
Details
Apply

Research role focused on frontier model evaluations and environments to improve agent capabilities while steering models toward safe AGI/ASI, in collaboration with safety/alignment partners.

Added Jun 26, 2026San Francisco, CAAI Safety & Alignment$295K - $445K / year
View details
Apply
Details
Apply

A full-time, in-person research fellowship in Singapore focused on technical AI safety and governance, with projects aimed at frontier AI safety practice and policy translation.

Added Jun 26, 2026SingaporeAI Safety & AlignmentSGD 5K / month
View details
Apply
Details
Apply

Editorial role on OpenAI’s Safety Systems team focused on producing and improving public-facing transparency materials about technical safety work, including evaluations, safeguards, red teaming, and deployment decisions for frontier models.

Added Jun 25, 2026San Francisco, CAAI Safety & Alignment$284K - $315K / year
View details
Apply
Details
Apply

Technical AI Safety Project Sprint BlueDot Impact Remote, Global This course teaches how to make a meaningful contribution to AI safety research or engineering through a structured 30-hour project sprint. Includes weekly mentorship from an AI safety expert to scope, refine, and execute your project with rapid feedback. Publishes projects on the course website as a public signal of your technical AI safety skills to collaborators and employers. Provides collaborative check-ins with peers to receive feedback and build your AI safety portfolio. BlueDot Impact is a nonprofit that runs courses which aim to help participants develop the knowledge, ...

Added Jun 25, 2026RemoteRemote - WorldwideAI Safety & AlignmentSalary not disclosed
View details
Apply
Apply

The Jailbreaking Lead will focus on identifying and mitigating vulnerabilities in frontier AI models, leading a red team to enhance AI safety and security through hands-on technical work and collaboration with AI developers and governments.

Added Jun 25, 2026San Francisco, CARemoteAI Safety & Alignment$170K - $250K / year
View details
Apply
Apply

Senior Research Engineer FAR AI San Francisco Bay Area, Remote, Global, Remote, USA $150,000 - $250,000 In this role, you'll accelerate AI safety research by tackling challenging engineering problems and increasing research depth. Lead projects in detecting AI deception, preventing misuse, or building research infrastructure. Mentor team members in technical work to elevate the team's capabilities. Apply software engineering expertise and Python skills to solve complex AI safety challenges. Contribute specialized knowledge in machine learning, high-performance computing, or technical leadership. FAR AI aims to ensure AI systems are trustworth...

Added Jun 25, 2026San Francisco, CARemoteAI Safety & Alignment$150K - $250K / year
View details
Apply
Details
Apply

Back to careers AI Security Research Engineer Remote / Part-time / Competitive The Mission The last major shift in cybersecurity produced CrowdStrike and SentinelOne. The next shift, AI-driven offense, will be bigger, and defense isn't ready. AI is compressing offensive cyber capability. Attacks that required nation-state resources will soon run autonomously, at machine speed. The industry's response so far has been to replay scripted attack simulations and hope for the best. That's not going to hold. 0Labs is building an AI-native platform and service for continuous purple teaming. Teams of agents execute real, adaptive cyber campaigns, then...

Added Jun 25, 2026RemoteRemote - WorldwideAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Machine Learning Researcher Gray Swan Remote, USA, Remote, Global In this role, you'll conduct AI safety research to identify and defend against failure modes in advanced AI systems. Discover emerging failure modes through stress-testing cutting-edge AI models. Help enterprises deploy AI safely and at scale without compromising innovation. Inform official safety evaluations of the world's most advanced AI models. Develop research-backed solutions for emerging AI safety challenges. Gray Swan is an AI security company that develops tools that automatically assess the risks of AI models.

Added Jun 25, 2026RemoteRemote - WorldwideAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Join MIT CSAIL to lead a research program on Provable AI Safety through Self-Proving Models, building theoretical foundations and practical implementations.

Added Jun 25, 2026Cambridge, MAAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Technical Policy Researcher Irregular Tel Aviv, Israel In this role, you'll tackle complex questions of AI development and deployment as a Technical Policy Researcher, shaping the emerging field of AI security. Conduct threat modeling to identify specific ways strong cyber-capable models could cause harm. Develop taxonomies for dangerous AI capabilities and potential mitigations. Create policy proposals for models' refusal policies and track developments in AI policy and research. Write academic papers and blog posts detailing research on evaluation theory and mitigation recommendations. Irregular Labs is a frontier security lab that aims to ...

Added Jun 25, 2026Tel Aviv, IsraelAI Safety & AlignmentSalary not disclosed
View details
Apply
Apply

Cyber Researcher Irregular Tel Aviv, Israel In this role, you'll conduct research on AI security, focusing on protecting models against cyber threats and evaluating AI's cybersecurity capabilities. Develop cyber capabilities evaluation challenges including CTF-style vulnerability tests and network attack simulations. Research methods to mitigate AI misuse risks, model weight theft, and dangerous AI agent capabilities. Publish research findings and deliver results to customers. Advise on product development while exploring frontier questions about AI's potential in cybersecurity. Irregular Labs is a frontier security lab that aims to protect t...

Added Jun 25, 2026Tel Aviv, IsraelAI Safety & AlignmentSalary not disclosed
View details
Apply
Apply

Research Engineer Irregular Tel Aviv, Israel In this role, you'll build systems to evaluate and secure frontier AI models at Irregular. Develop infrastructure and experiments to assess model capabilities and implement agent frameworks. Create robust evaluation pipelines and security-focused testing frameworks. Design challenges to measure models' ability to evade detection by defensive security tools. Build controlled environment frameworks and tools that help understand and mitigate risks related to frontier models. Irregular Labs is a frontier security lab that aims to protect the world in a time of increasingly capable and sophisticated AI...

Added Jun 25, 2026Tel Aviv, IsraelAI Safety & AlignmentSalary not disclosed
View details
Apply
Details
Apply

Technical Project Manager, Red Team FAR AI Remote, Global, San Francisco Bay Area, Remote, USA $125,000 - $190,000 In this role, you'll be the delivery backbone of FAR.AI's red-teaming programme, owning engagements with governments and frontier AI companies. Own the end-to-end red-team hiring pipeline, including sourcing, work trials, and recruiting technical talent. Manage the RFP and opportunity pipeline by scoping engagements, drafting proposals, and supporting negotiations. Conduct analysis on team bottlenecks and ecosystem mapping to improve performance. Support technical writing, event organising, policy work, and grant applications as ...

Added Jun 25, 2026San Francisco, CARemoteAI Safety & Alignment$125K - $190K / year
View details
Apply
Apply

AI Red Team Analyst Alignerr Remote, Global $15 - $75 per hour In this role, you'll conduct red-teaming exercises to uncover AI security weaknesses and deliver findings that improve system safety. Craft adversarial prompts, jailbreak attempts, and edge-case scenarios to challenge AI model guardrails. Evaluate AI outputs for safety violations, bias, and policy compliance. Document vulnerabilities and unexpected behaviours in structured reports for engineering teams. Collaborate with teams to recommend security mitigations and help refine testing protocols. Alignerr is a platform that for people to create data for frontier AI labs.

Added Jun 25, 2026RemoteRemote - WorldwideAI Safety & Alignment$15 - $75 / hour
View details
Apply
Apply

Research Manager / Research Managers, AIxCyber ERA Cambridge, UK £62,000 - £75,000 In this role, you'll lead the execution of ERA's AIxCyber Research Fellowship, supporting research fellows in developing and delivering their projects. Review and evaluate fellowship applications, then match fellows with mentors suited to their research focus. Help fellows scope research projects aligned with their skills and interests. Direct and oversee up to 5 research projects while providing guidance on planning, methodology, and publication. Organise workshops and events to support fellow professional development and strengthen relationships across the AI...

Added Jun 25, 2026Cambridge, UKAI Safety & Alignment£62K - £75K / year
View details
Apply