Job description
Trajectory Labs • Remote
About us
Our mission is to automate AI safety, to pave the way for a future where the vast majority of AI safety work is done by AI models.
Frontier models already solve coding problems that take humans days, but a model that can be hijacked by a malicious email or web page can't be trusted to work on its own. Before AI can do the work that matters, including AI safety research itself, models have to be robust to attack. So we build the safety and alignment evals, red-teaming programs, and RL environments that find these failures and train them out.
Although Trajectory is less than a year old, we have already been featured in Anthropic's Fable 5 & Mythos 5 System Card (the only external team to bypass its cyber safeguards and have it generate a working exploit), led Anthropic's prompt injection evaluation of Auto mode, and been featured in Meta's Muse Spark Safety & Preparedness Report.
About the role
You will probe frontier AI systems for vulnerabilities, design novel attack strategies, and help build the evaluation infrastructure that keeps AI safe.
About you
Essential
- Successful jailbreaks of frontier models (GPT-5, Claude, Gemini, etc.)
- Self-directed — you work independently with minimal guidance
- Creative problem-solving and lateral thinking
Highly valued
- Jailbreaking competition participation
- Python scripting and command-line proficiency
- Expertise with agentic coding tools
- Security consulting or ML security background
Application process
- Submit your application with a resume
- Complete Terminal, The Game
- Interview with the founders
- Offer
Applications reviewed on a rolling basis.
Apply now $ /$