Evaluation Engineer role focused on LLM evaluation frameworks, evaluation infrastructure, and production-readiness metrics for enterprise AI systems.
Curated roles
Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.
363 active roles found.
Evaluation Engineer role focused on LLM evaluation frameworks, evaluation infrastructure, and production-readiness metrics for enterprise AI systems.
Safety-team role focused on model behavior, alignment, and evaluation of large language models, including building evaluation pipelines and synthetic testing environments.
Offensive security/security researcher role focused on protecting agentic AI systems through AI safety work, red teaming, and model-security risk assessment.
Lead Anthropic’s frontier cyber red team, overseeing research on offensive and defensive capabilities of Claude, model safeguarding, and defenses against advanced AI-enabled cybersecurity risks.
A 3-month full-time AI safety research fellowship with mentorship, reading groups on AI risks, and a final symposium in Cape Town.
Part-time mentor role supervising 3-month research projects for aspiring researchers in AI safety, policy, governance, or biosecurity.
3-year PhD fellowship researching safety and security evaluation of deployed AI systems, including evaluation, red-teaming, monitoring, and vendor-independent auditing tools for high-stakes deployments.
Security engineer for research infrastructure at an AI alignment nonprofit, building security controls for automated research pipelines, agents, and multi-cloud environments that support frontier alignment research.
Technical communications role translating AI research for policymakers, media, and the public, with explicit responsibility for communicating AI safety and responsible AI research.
Machine Learning Engineer building language-model products, evaluation systems, and trust/transparency features for research and decision-making. Relevant because the role explicitly includes model evaluations and process supervision for more trustworthy AI systems.
Senior hands-on product engineering role building analyst software for verification decisions and AI misuse prevention in a biosecurity-focused trusted-access layer.
Manages end-to-end AI alignment research projects, coordinating technical teams, collaborators, and funders on model safety and interpretability work.
Part-time remote teacher for a 14-week technical AI safety program, leading lectures, AI safety sessions, and participant support on alignment-related topics.
A founder funding request for projects and ventures in AI safety, alignment, governance, standards, and related resilience areas.
Senior virologist role on DeepMind's Responsible Development and Innovation team focused on biology evaluations, LLM safety assessments, harm frameworks, and mitigation strategies for safe model releases.
Research engineer role building multimodal reasoning and vision-language systems to assess media trustworthiness, detect manipulated content, and support information literacy; adjacent to AI safety via misuse prevention and evaluation-related work.
Product Manager for Anthropic’s Safeguards team, owning safety systems, evals, detections, interventions, and product UX to mitigate deployment and user risks for frontier AI models.
Operations role supporting ARENA’s technical AI safety bootcamps, including participant selection, programme logistics, and operational improvements for an AI safety training organisation.
Senior virology and biosecurity role at DeepMind focused on evaluating frontier AI models against biological threats, shaping biosecurity mitigation strategies, and embedding safety protocols into model development.
Operations and community management role supporting Meridian’s AI safety research workspace, events, and researcher community in Cambridge.