ML Research Engineer focused on training, post-training, and evaluating LLMs, including alignment methods and safety/moderation-related datasets and policy systems.
A weekly digest of AI safety, governance, policy and responsible AI roles.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
5 active roles found for White Circle
ML Research Engineer focused on training, post-training, and evaluating LLMs, including alignment methods and safety/moderation-related datasets and policy systems.
AI Red Team Engineer focused on adversarial testing of LLM-powered systems, including jailbreaks, prompt injection, data leakage, policy bypass, and turning findings into regression tests and reports.
Research role building adversarial agent environments, evaluations, and tooling to study AI system failures, misalignment, and unsafe behaviors.
Get the best AI safety, governance, policy and responsible AI roles in your inbox.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
Research engineer role focused on building and maintaining AI safety evaluation benchmarks, guardrails, and research on agentic failure modes.
Research scientist role focused on AI behavior failure modes, model evaluations, and safety research for LLM agents, including deception, misalignment, unsafe behavior, and frontier agent pressure-testing.