Lead role focused on AI safety evaluations and security testing for LLMs and agentic AI systems, including red-teaming, prompt injection testing, and producing governance/compliance evidence.
A weekly digest of AI safety, governance, policy and responsible AI roles.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
689 active roles found
Lead role focused on AI safety evaluations and security testing for LLMs and agentic AI systems, including red-teaming, prompt injection testing, and producing governance/compliance evidence.
Lead BlueDot Impact’s technical AI safety course, owning curriculum, strategy, admissions, and participant acceleration for a program that trains people entering AI safety.
State-level AI policy role supporting policymakers with AI governance, policy briefings, pilot projects, and tracking AI policy developments; relevant because it directly focuses on AI policy and governance.
Get the best AI safety, governance, policy and responsible AI roles in your inbox.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
Research manager role at a nonprofit building frontier-AI-specific guidance for insiders, with research operations and publication responsibilities tied to AI policy and governance-adjacent whistleblowing work.
AI governance research and policy role focused on Singapore and Southeast Asia, producing research, policy analysis, and supporting convenings and stakeholder engagement on AI safety and governance.
Program lead for BlueDot Impact’s Frontier AI Governance course, owning strategy, curriculum, participant selection, and talent acceleration for the AI governance field.
Senior AI researcher role focused on post-training evaluation, red-teaming, and RL gym audits for open-weight LLMs, with direct work on security capabilities and defensive alignment against indirect prompt injection.
Senior advisory contractor role focused on advanced AI risks, loss-of-control scenarios, safeguards, incident prevention, and translating technical research into policy and operational guidance.
Communications and outreach role at Cedar Research focused on promoting AI safety work through writing, video, outreach, and possibly helping run human studies.
Manager role focused on agentic safety and model policy for frontier AI systems, turning alignment and misalignment risks into behavioral policies, evaluations, monitoring, and safeguards.
Seoul-based public policy lead for Anthropic focused on AI governance, AI regulation, and AI safety engagement with Korean government and stakeholders.
Mathematical research role focused on AI safety, including theoretical problems and novel mathematical approaches to safety challenges.
Winter research fellowship focused on EU AI law and policy, especially the EU AI Act, with research and advisory work related to AI governance and regulation.
Postdoctoral researcher role on multilingual mechanistic interpretability, including circuit analysis, controlled validation with backdoored model suites, and work that feeds into safe agentic system design.
A remote global funding call for founders to start new organizations tackling critical AI safety problems, including alignment moonshots and frontier capabilities work.
Staff+ software engineer role building production ML infrastructure for Claude's safety systems, including safety deployments, monitoring, and productionizing safety research.
Postdoctoral research role on safe, controllable agentic AI, combining runtime behavioral control, normative constraints, integrated evaluation, and publication in AI safety and formal methods venues.
Technical AI safety research role focused on AI alignment, robust initialization methods for capable language models, and empirical evaluation of alignment techniques.
Research engineer role studying whether values persist after reinforcement learning, using training pipelines, evaluations, and interpretability methods to analyze model behavior and alignment.
Fully-funded PhD position researching responsible machine learning, with explicit focus on AI safety and robust alignment in multi-agent LLM systems.