AI Red Team Engineer focused on adversarial testing of LLM-powered systems, including jailbreaks, prompt injection, data leakage, policy bypass, and turning findings into regression tests and reports.
Curated roles
Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.
393 active roles found.
AI Red Team Engineer focused on adversarial testing of LLM-powered systems, including jailbreaks, prompt injection, data leakage, policy bypass, and turning findings into regression tests and reports.
Research role building adversarial agent environments, evaluations, and tooling to study AI system failures, misalignment, and unsafe behaviors.
Research engineer role building an AI safety argumentation platform using ontologies, knowledge graphs, defeasible argumentation, and LLM-assisted pipelines to support AI risk management and governance communications.
Contract role focused on monitoring and enforcing abuse on AI products, including building detection, review, and enforcement systems for high-risk harms.
Engineering fellowship supporting AI abuse detection, red teaming, and related research/engineering work across software, data, and ML concentrations.
Evaluation Engineer role focused on LLM evaluation frameworks, evaluation infrastructure, and production-readiness metrics for enterprise AI systems.
Safety-team role focused on model behavior, alignment, and evaluation of large language models, including building evaluation pipelines and synthetic testing environments.
Offensive security/security researcher role focused on protecting agentic AI systems through AI safety work, red teaming, and model-security risk assessment.
A 3-month full-time AI safety research fellowship with mentorship, reading groups on AI risks, and a final symposium in Cape Town.
Part-time mentor role supervising 3-month research projects for aspiring researchers in AI safety, policy, governance, or biosecurity.
3-year PhD fellowship researching safety and security evaluation of deployed AI systems, including evaluation, red-teaming, monitoring, and vendor-independent auditing tools for high-stakes deployments.
Security engineer for research infrastructure at an AI alignment nonprofit, building security controls for automated research pipelines, agents, and multi-cloud environments that support frontier alignment research.
Technical communications role translating AI research for policymakers, media, and the public, with explicit responsibility for communicating AI safety and responsible AI research.
Machine Learning Engineer building language-model products, evaluation systems, and trust/transparency features for research and decision-making. Relevant because the role explicitly includes model evaluations and process supervision for more trustworthy AI systems.
Senior hands-on product engineering role building analyst software for verification decisions and AI misuse prevention in a biosecurity-focused trusted-access layer.
Manages end-to-end AI alignment research projects, coordinating technical teams, collaborators, and funders on model safety and interpretability work.
Part-time remote teacher for a 14-week technical AI safety program, leading lectures, AI safety sessions, and participant support on alignment-related topics.
A founder funding request for projects and ventures in AI safety, alignment, governance, standards, and related resilience areas.
Senior virologist role on DeepMind's Responsible Development and Innovation team focused on biology evaluations, LLM safety assessments, harm frameworks, and mitigation strategies for safe model releases.
Operations role supporting ARENA’s technical AI safety bootcamps, including participant selection, programme logistics, and operational improvements for an AI safety training organisation.