The role involves pursuing a PhD with a focus on AI safety, generative AI, and agentic AI systems, addressing critical aspects of AI development and deployment.
Curated roles
Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.
393 active roles found.
The role involves pursuing a PhD with a focus on AI safety, generative AI, and agentic AI systems, addressing critical aspects of AI development and deployment.
Research engineer building AI platforms and infrastructure for alignment research, safety research, and evaluation of AI alignment/control approaches at CARMA.
The AI Red Teamer role involves probing frontier AI systems for vulnerabilities, designing attack strategies, and building safety evaluation infrastructure to ensure AI safety.
Research Scientist role studying latent structure and behavior in neural networks, with explicit emphasis on understanding as safety, safety-relevant tools, and internal red teaming.
Founding engineer role building an AI-native cyberdefense platform, including AI systems and evaluations that measure performance against real cyber incidents and frontier AI labs.
Technical lead for ARIA’s multi-agent security programme, supporting research on AI agents operating and coordinating in untrusted environments, with explicit emphasis on AI red-teaming and security research.
Computational biologist role focused on building evaluation frameworks and experiments for biological security problems using frontier AI systems.
Role building evaluation task pipelines and infrastructure for frontier model testing, including systems to prevent models from detecting evaluations.
Product engineer building interfaces, APIs, and workflows that operationalize interpretability research for training, evaluating, debugging, and deploying AI systems; relevant because it supports safety-related understanding of model internals and safer AI systems.
Funding call for research on AI safety in multi-agent systems, including testbeds, agent networks, infrastructure, and oversight/control for frontier-model agents.
PhD/visiting PhD research role focused on AI safety, security, and alignment for advanced autonomous systems, including interpretability, evaluations, situational awareness, and red-teaming.
Lead a research program on Provable AI Safety, focusing on building AI systems that verify their own correctness and collaborating with teams to advance AI safety.
The AI Security Systems Architect will design and develop systems for security testing and evaluation of AI technologies, focusing on adversarial testing methodologies and ensuring the resilience and security of AI systems.
Research scientist role focused on designing and running evaluations of AI harmful manipulation for EU AI Act enforcement, including frontier model evals, red-teaming, and regulator-facing reporting.
Technical project manager for a research division building harmful-manipulation evaluations for the EU AI Office, coordinating delivery, budgets, partners, and hiring for AI safety evaluation work.
The Research Engineer for Code RL will advance AI models' coding capabilities, ensuring safety and effectiveness through collaboration with alignment and red teams, focusing on responsible AI development.
Senior analyst role at OpenAI focused on assessing and mitigating agentic risks across products and platforms, using evaluations, red teaming, investigations, and cross-functional risk coordination.
Product manager for Anthropic’s Safeguards team, building safety systems, evals, and interventions to mitigate risks from frontier AI models and user misuse.
The role involves conducting research in AI safety and alignment, focusing on model evaluations and interpretability techniques, while collaborating with top researchers to address critical AI challenges.
People management and technical leadership role guiding ML research projects for LawZero’s AI safety agenda and safe-by-design AI systems.