Engineer role focused on cyber-relevant model evaluations, safeguard robustness testing, and misuse detection for frontier AI systems.
Curated roles
Browse remote-friendly roles across AI safety, governance, policy, evaluations, compliance and responsible AI teams.
240 active roles found.
Engineer role focused on cyber-relevant model evaluations, safeguard robustness testing, and misuse detection for frontier AI systems.
Senior regional public policy and partnerships role covering Central and Eastern Europe and the Baltics, with direct responsibility for AI policy engagement, regulatory monitoring, and shaping safe and responsible AI policy environments.
Security research role focused on AI auditing methodologies, AI system vulnerabilities, and tools/frameworks for evaluating AI safety and security.
Founding in-house legal role focused on AI governance, interpreting the EU AI Act and related regimes, owning AI policy and compliance, and leading ISO 42001/27001 audits while supporting product and enterprise deals.
Lead enforcement operations to detect and mitigate misuse of Anthropic AI systems for cyberattacks, malware, and offensive exploitation, including managing a team and shaping enforcement strategy.
Frontier AI red teaming role focused on attacking cyber safeguards, designing attack strategies, and feeding results into safety/alignment evaluations and deployment decisions.
Part-time technical advisor role for AI safety research on fine-tuning and evaluations, advising on benchmark design, generalization testing, and validity of welfare/capability evals for open-weight models.
Lead Talos Network’s AI governance incubator, recruiting and supporting participants to launch new European AI governance initiatives and mapping field gaps in frontier AI governance capacity.
Product Policy Manager focused on product risk: assessing AI product launches for safety risks, running product safety reviews and bespoke evaluations, and recommending mitigations and responsible AI deployment policies.
Engineering manager for FAR.AI’s red team, leading engineering systems for frontier AI red-teaming, safety evaluations, attacker simulation, and safeguard testing.
Policy role in Anthropic's Safeguards organization focused on conventional weapons: defining policy boundaries, building threat models and evaluations, and operationalizing enforcement for model misuse prevention.
Research engineer role building technical prototypes for AI verification mechanisms, with direct links to policymakers and real-world AI governance.
Operations Manager supporting grantmaking, finance, people ops, and cross-entity coordination at a foundation funding European AI policy, AI safety field-building, and non-US AI governance.
Project manager role leading pilot audits for frontier AI systems, focused on building AI auditing methodologies, assessment tools, and industry standards for third-party safety and security audits.
Government affairs and public policy lead covering the Middle East, with direct responsibility for AI policy, regulation monitoring, and advocacy related to responsible AI adoption.
Associate role on Faculty’s AI Safety team supporting delivery of frontier model evaluations, AI safety red teaming, and related client projects for government and industry.
Senior privacy engineering role at Anthropic focused on building privacy-preserving systems for AI training and inference, translating AI/privacy regulations into technical controls, and conducting privacy reviews and threat modeling for new models and features.
Internship focused on ML research engineering for LLM and agent evaluation, red-teaming, guardrails, and model safety/robustness tooling.
One-year senior fellowship placed at the California Department of Technology to advise on AI policy and governance, with explicit focus on frontier AI safety, risk management, transparency, and California AI safety legislation.
Research Scientist role focused on AI safety research, model behavior studies, interpretability, and building reproducible tooling for safe AI deployment.