Fully-funded PhD position researching responsible machine learning, with explicit focus on AI safety and robust alignment in multi-agent LLM systems.
A weekly digest of AI safety, governance, policy and responsible AI roles.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
677 active roles found
Fully-funded PhD position researching responsible machine learning, with explicit focus on AI safety and robust alignment in multi-agent LLM systems.
Senior research role in Oxford’s Technical AI Governance programme focused on interpretability, evaluations, and AI safety for continuously learning systems, with work on foundation-model experiments.
Research engineer role studying whether values persist after reinforcement learning, using training pipelines, evaluations, and interpretability methods to analyze model behavior and alignment.
Staff+ software engineer role building production ML infrastructure for Claude's safety systems, including safety deployments, monitoring, and productionizing safety research.
Three-year fully salaried PhD scholarship researching AI regulation and risk, with emphasis on risk-based regulation, the EU AI Act, and how organizations assess and mitigate AI-related compliance risk.
Postdoctoral research role on safe, controllable agentic AI, combining runtime behavioral control, normative constraints, integrated evaluation, and publication in AI safety and formal methods venues.
Get the best AI safety, governance, policy and responsible AI roles in your inbox.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
Technical AI safety research role focused on AI alignment, robust initialization methods for capable language models, and empirical evaluation of alignment techniques.
Senior staff software engineering role on OneTrust’s AI Governance Engineering team building platform capabilities for AI governance, compliance, safety, observability, controls, and agent runtime management.
Founding technical research role building white-box auditing methods and infrastructure for frontier AI evaluations, with direct focus on interpretability and safety-relevant failure prediction.
Research scientist role in Resolution’s philosophy program focused on AI alignment research, including conceptual analysis, empirical hypotheses, and evaluation methods for aligned ASI.
Security architect role designing physical and operational security for an adversarial testbed supporting AI compute verification, red-teaming, and governance-relevant verification standards.
Senior SRE role on Anthropic's Safeguards ML Infra team, focused on deploying, verifying, and operating production safety infrastructure and safety classifiers for frontier model launches.
Senior compliance leadership role owning Anthropic's sanctions compliance program and supporting export controls, with explicit responsibility for risk governance, policy frameworks, remediation, and AI-related regulatory compliance.
Product design role focused on building and maintaining LLM evaluation systems, prompt fixes, and test harnesses for Claude surfaces and model launches, with explicit safety-related evaluation scope.
Program manager for the Canadian AI Safety Institute research program, supporting AI safety initiatives, research partnerships, proposal calls, and related events.
Senior research role focused on interpretability, evaluations, and AI safety for continuously learning foundation models, with some collaboration on AI governance research.
Delivery Manager role in Faculty’s AI Safety team focused on frontier model evaluations, AI safety red teaming, and delivery of high-impact safety projects for clients, including government and frontier labs.
Research role on agent safety, oversight, evaluations, red-teaming, and system-level mitigations for increasingly capable AI agents operating safely and autonomously.
Legal counsel role on Anthropic's Research Legal team advising on frontier AI model development, AI regulation, model launches, system cards, Responsible Scaling Policy, and capability evaluations.
Lead research role at an AI safety organization focused on frontier model risks, misalignment, loss of control, harmful manipulation, and rigorous model evaluation research.