Senior hands-on product engineering role building analyst software for verification decisions and AI misuse prevention in a biosecurity-focused trusted-access layer.
Curated roles
Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.
364 active roles found.
Senior hands-on product engineering role building analyst software for verification decisions and AI misuse prevention in a biosecurity-focused trusted-access layer.
A paid, full-time remote fellowship for technical AI alignment research, with dedicated research support and publishable outputs.
Manages end-to-end AI alignment research projects, coordinating technical teams, collaborators, and funders on model safety and interpretability work.
Part-time remote teacher for a 14-week technical AI safety program, leading lectures, AI safety sessions, and participant support on alignment-related topics.
A founder funding request for projects and ventures in AI safety, alignment, governance, standards, and related resilience areas.
Senior virologist role on DeepMind's Responsible Development and Innovation team focused on biology evaluations, LLM safety assessments, harm frameworks, and mitigation strategies for safe model releases.
Research engineer role building multimodal reasoning and vision-language systems to assess media trustworthiness, detect manipulated content, and support information literacy; adjacent to AI safety via misuse prevention and evaluation-related work.
Product Manager for Anthropic’s Safeguards team, owning safety systems, evals, detections, interventions, and product UX to mitigate deployment and user risks for frontier AI models.
Hands-on security engineering role building secure-by-default infrastructure, controls, and assurance for frontier-scale AI safety research and model evaluation workflows.
Operations role supporting ARENA’s technical AI safety bootcamps, including participant selection, programme logistics, and operational improvements for an AI safety training organisation.
Senior virology and biosecurity role at DeepMind focused on evaluating frontier AI models against biological threats, shaping biosecurity mitigation strategies, and embedding safety protocols into model development.
Operations and community management role supporting Meridian’s AI safety research workspace, events, and researcher community in Cambridge.
Research Manager for an AI safety research training programme, supporting teams, technical enablement, proposal evaluation, and participant development.
Founding Engineer on SaferAI’s Evaluations team, building infrastructure for model evaluations and helping run mitigation-focused safety evaluations for frontier AI labs.
Lead OpenAI’s policy and engagement strategy for cyber-focused testing and evaluation of advanced AI capabilities, coordinating with safety and security teams and government partners.
Enforcement analyst role focused on reviewing flagged activity and enforcing policies to prevent misuse of Anthropic's AI systems for cyberattacks, malware, and related harmful operations.
Red team role on Anthropic's Safeguards team focused on adversarial testing of deployed AI systems, model jailbreaking, prompt injection, and detecting novel abuse in frontier AI products.
Part-time technical fellowship building AI evaluation infrastructure, adversarial evaluation tooling, risk data systems, and governance-related prototypes for an AI risk firm.
Project-based adversarial evaluation and red teaming network for frontier AI systems, with work spanning ML security, prompt injection/jailbreaking, agentic system evaluation, and translating findings into governance and mitigation decisions.
Lead threat modeling for CBRNe risks in advanced AI systems, building evaluations and safety frameworks to inform deployment decisions and mitigate dual-use misuse.