Research engineer role focused on building and operating SimpleAudit for AI safety audits, red teaming, and safety evaluations of language models and AI systems.
Curated roles
Explore roles in AI red teaming, evaluations, model behavior testing, adversarial assessment and safety measurement.
24 active roles found.
Research engineer role focused on building and operating SimpleAudit for AI safety audits, red teaming, and safety evaluations of language models and AI systems.
Technical Program Manager role leading GenAI Safety operations for Gemini and related models, coordinating training, evaluations, red teaming, and safety/alignment initiatives.
Senior manager role leading AI safety delivery, including frontier model evaluations, red teaming, safeguard testing, and client advisory work with government and frontier labs.
Delivery Manager role in Faculty’s AI Safety team focused on frontier model evaluations, AI safety red teaming, and delivery of high-impact safety projects for clients, including government and frontier labs.
Frontier AI red teaming role focused on attacking cyber safeguards, designing attack strategies, and feeding results into safety/alignment evaluations and deployment decisions.
Associate role on Faculty’s AI Safety team supporting frontier model evaluations, AI safety red teaming, and delivery of responsible AI projects for government and industry clients.
Associate role on Faculty’s AI Safety team supporting delivery of frontier model evaluations, AI safety red teaming, and related client projects for government and industry.
Assistant AI Security Software Engineer focused on AI security research, AI red teaming, adversarial machine learning, and building tools for AI security applications.
Staff software engineer building web applications and backend systems for AI model evaluation, real-time threat detection, and adversarial red teaming for frontier AI deployments.
Senior full-stack software engineer role building AI model evaluation, threat detection, and adversarial red teaming products for frontier labs.
Lead role overseeing GenAI safety research and delivery, including adversarial testing, risk evaluations, red teaming methodology, and client-facing safety deliverables.
Research Scientist role focused on AI evaluation, language model understanding, robustness, red teaming, and alignment for LLM evaluation infrastructure.
Engineering fellowship supporting AI abuse detection, red teaming, and related research/engineering work across software, data, and ML concentrations.
Offensive security/security researcher role focused on protecting agentic AI systems through AI safety work, red teaming, and model-security risk assessment.
Project-based adversarial evaluation and red teaming network for frontier AI systems, with work spanning ML security, prompt injection/jailbreaking, agentic system evaluation, and translating findings into governance and mitigation decisions.
Editorial role on OpenAI’s Safety Systems team focused on producing and improving public-facing transparency materials about technical safety work, including evaluations, safeguards, red teaming, and deployment decisions for frontier models.
Research Scientist role studying latent structure and behavior in neural networks, with explicit emphasis on understanding as safety, safety-relevant tools, and internal red teaming.
Senior analyst role at OpenAI focused on assessing and mitigating agentic risks across products and platforms, using evaluations, red teaming, investigations, and cross-functional risk coordination.
Fellowship focused on adversarial red teaming and model evaluations for AI systems, identifying vulnerabilities, safety risks, and failure modes.
Machine Learning Engineer building AI safety and security systems, including abuse detection, frontier model evaluations, and red teaming workflows.