Lead role focused on frontier AI risk intelligence, abuse/misuse detection, risk frameworks, and mitigation guidance to improve safety readiness for OpenAI products.
Curated roles
Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.
363 active roles found.
Lead role focused on frontier AI risk intelligence, abuse/misuse detection, risk frameworks, and mitigation guidance to improve safety readiness for OpenAI products.
Research Engineer role focused on safety oversight for deployed generative AI models, including automated evaluations, misuse detection, and monitoring model behavior at scale.
Research engineer role on DeepMind's Safety Oversight team focused on monitoring deployed model safety and alignment, building classifiers and data pipelines to detect misbehavior, misuse, and coordinated harms.
Cybersecurity Engineer building infrastructure and tooling for AI security evaluations, adversarial testing, and attack simulations to assess and improve the resilience of AI-powered systems.
Research software engineer role on the Safety team building secure infrastructure for sensitive model evaluations, including dangerous-capability domains, to support release decisions for models.
Senior research scientist role focused on AI safety evaluations, red-teaming, and technical research to quantify and mitigate AI risks in high-impact domains.
Research engineer role on DeepMind’s AGI Safety and Alignment Team, focused on alignment methods, adversarially robust control systems, and interpretability for frontier AI models.
Research engineer role on DeepMind’s AGI Safety and Alignment Team focused on alignment methods, adversarially robust control systems, interpretability, and frontier model safety.
Software engineer for xAI’s evals platform, building systems to measure model capabilities and behaviors and turn quality, truthfulness, and safety into measurable criteria.
Tech lead/manager for evals infrastructure at Anthropic, building distributed systems and harnesses for frontier model evaluations that support safety decisions and launch readiness.
Contract infrastructure engineer role building and scaling the Loss of Control Observatory’s data pipelines, LLM-based classification systems, and monitoring dashboard for AI safety monitoring.
Research program management role focused on adversarial model research, safety and robustness evaluations, red-teaming, and identifying failure modes in advanced AI systems.
Research Engineer / Research Scientist role focused on empirical AI safety research for frontier AI systems, including experiments on LLMs, training/evaluation tooling, datasets, and benchmarks.
Research engineer internship focused on AI safety-relevant empirical research on LLMs, including AI security, machine ethics, AI alignment, and benchmarking AI risks.
ML Research Engineer focused on training, post-training, and evaluating LLMs, including alignment methods and safety/moderation-related datasets and policy systems.
AI Red Team Engineer focused on adversarial testing of LLM-powered systems, including jailbreaks, prompt injection, data leakage, policy bypass, and turning findings into regression tests and reports.
Research role building adversarial agent environments, evaluations, and tooling to study AI system failures, misalignment, and unsafe behaviors.
Research engineer role building an AI safety argumentation platform using ontologies, knowledge graphs, defeasible argumentation, and LLM-assisted pipelines to support AI risk management and governance communications.
Contract role focused on monitoring and enforcing abuse on AI products, including building detection, review, and enforcement systems for high-risk harms.
Engineering fellowship supporting AI abuse detection, red teaming, and related research/engineering work across software, data, and ML concentrations.