A weekly digest of AI safety, governance, policy and responsible AI roles.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
Research Scientist Patronus AI Remote, USA, San Francisco Bay Area, New York, NY In this role, you will help solve challenging open research problems facing society’s adoption of AI today, surrounding AI evaluation, language model understanding, and robustness challenges. Develop state-of-the-art systems for AI evaluation and implement algorithms based on NLP advancements. Conduct novel research on redteaming language models, automated evaluation, and alignment. Scope out and lead research projects, including experiment design and understanding results. Patronus AI aims to provide a security and risk management layer for AI by building an aut...
Added Jun 25, 2026San Francisco, CARemoteAI Safety & AlignmentSalary not disclosed
The Jailbreaking Lead will focus on identifying and mitigating vulnerabilities in frontier AI models, leading a red team to enhance AI safety and security through hands-on technical work and collaboration with AI developers and governments.
Added Jun 25, 2026San Francisco, CARemoteAI Safety & Alignment$170K - $250K / year
Senior Research Engineer FAR AI San Francisco Bay Area, Remote, Global, Remote, USA $150,000 - $250,000 In this role, you'll accelerate AI safety research by tackling challenging engineering problems and increasing research depth. Lead projects in detecting AI deception, preventing misuse, or building research infrastructure. Mentor team members in technical work to elevate the team's capabilities. Apply software engineering expertise and Python skills to solve complex AI safety challenges. Contribute specialized knowledge in machine learning, high-performance computing, or technical leadership. FAR AI aims to ensure AI systems are trustworth...
Added Jun 25, 2026San Francisco, CARemoteAI Safety & Alignment$150K - $250K / year
Get the best AI safety, governance, policy and responsible AI roles in your inbox.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
Back to careers AI Security Research Engineer Remote / Part-time / Competitive The Mission The last major shift in cybersecurity produced CrowdStrike and SentinelOne. The next shift, AI-driven offense, will be bigger, and defense isn't ready. AI is compressing offensive cyber capability. Attacks that required nation-state resources will soon run autonomously, at machine speed. The industry's response so far has been to replay scripted attack simulations and hope for the best. That's not going to hold. 0Labs is building an AI-native platform and service for continuous purple teaming. Teams of agents execute real, adaptive cyber campaigns, then...
Added Jun 25, 2026RemoteRemote - WorldwideAI Safety & AlignmentSalary not disclosed
Machine Learning Researcher Gray Swan Remote, USA, Remote, Global In this role, you'll conduct AI safety research to identify and defend against failure modes in advanced AI systems. Discover emerging failure modes through stress-testing cutting-edge AI models. Help enterprises deploy AI safely and at scale without compromising innovation. Inform official safety evaluations of the world's most advanced AI models. Develop research-backed solutions for emerging AI safety challenges. Gray Swan is an AI security company that develops tools that automatically assess the risks of AI models.
Added Jun 25, 2026RemoteRemote - WorldwideAI Safety & AlignmentSalary not disclosed
Head of Operations / Chief Operating Officer Encode Washington, DC, Remote, USA $170,000 - $200,000 In this role, you'll own Encode's compliance, financial leadership, and operational infrastructure across our 501(c)(3) and 501(c)(4) entities. Lead legal and regulatory compliance, including lobbying registrations, political activity rules, and employment law. Manage budgeting, financial planning, audits, and Form 990 filings. Own vendor and counsel relationships with lawyers, accountants, and insurance brokers. Build HR and recruiting systems, manage your operations team, and design organisational structures for scaled growth. Encode is a glo...
Added Jun 25, 2026Washington, DCRemoteAI Governance & Policy$170K - $200K / year
Policy Analyst / Director Encode Washington, DC $110,000 In this role, you'll drive full-stack policy development across the organization, focusing on AI governance and public interest advocacy. Draft legislative language, policy briefs, and talking points for lawmakers. Build coalitions with stakeholders across government, industry, and civil society. Conduct policy research and create communications materials under tight deadlines. Support multiple concurrent projects, including emergency responses and time-sensitive policy opportunities. Encode is a global, youth-led organisation that advocates for policies to ensure responsible AI develop...
Added Jun 25, 2026Washington, DCAI Governance & Policy$110K / year
Policy Advisor, Coalitions Encode Washington, DC $120,000 - $180,000 In this role, you'll build coalitions to drive AI policy priorities forward across federal and state campaigns. Identify advocacy gaps and recruit allies across disparate issue areas to strengthen legislative campaigns. Organize and support grassroots advocates, events, and state-level networks to advance AI regulation priorities. Analyse complex legislation and prepare briefs, talking points, and policy explainers for partners and policymakers. Manage digital infrastructure including coalition lists, email campaigns, and supporter communications. Encode is a global, youth-l...
Added Jun 25, 2026Washington, DCAI Governance & Policy$120K - $180K / year
Technical Policy Researcher Irregular Tel Aviv, Israel In this role, you'll tackle complex questions of AI development and deployment as a Technical Policy Researcher, shaping the emerging field of AI security. Conduct threat modeling to identify specific ways strong cyber-capable models could cause harm. Develop taxonomies for dangerous AI capabilities and potential mitigations. Create policy proposals for models' refusal policies and track developments in AI policy and research. Write academic papers and blog posts detailing research on evaluation theory and mitigation recommendations. Irregular Labs is a frontier security lab that aims to ...
Added Jun 25, 2026Tel Aviv, IsraelAI Safety & AlignmentSalary not disclosed
Cyber Researcher Irregular Tel Aviv, Israel In this role, you'll conduct research on AI security, focusing on protecting models against cyber threats and evaluating AI's cybersecurity capabilities. Develop cyber capabilities evaluation challenges including CTF-style vulnerability tests and network attack simulations. Research methods to mitigate AI misuse risks, model weight theft, and dangerous AI agent capabilities. Publish research findings and deliver results to customers. Advise on product development while exploring frontier questions about AI's potential in cybersecurity. Irregular Labs is a frontier security lab that aims to protect t...
Added Jun 25, 2026Tel Aviv, IsraelAI Safety & AlignmentSalary not disclosed
Research Engineer Irregular Tel Aviv, Israel In this role, you'll build systems to evaluate and secure frontier AI models at Irregular. Develop infrastructure and experiments to assess model capabilities and implement agent frameworks. Create robust evaluation pipelines and security-focused testing frameworks. Design challenges to measure models' ability to evade detection by defensive security tools. Build controlled environment frameworks and tools that help understand and mitigate risks related to frontier models. Irregular Labs is a frontier security lab that aims to protect the world in a time of increasingly capable and sophisticated AI...
Added Jun 25, 2026Tel Aviv, IsraelAI Safety & AlignmentSalary not disclosed
Technical Project Manager, Red Team FAR AI Remote, Global, San Francisco Bay Area, Remote, USA $125,000 - $190,000 In this role, you'll be the delivery backbone of FAR.AI's red-teaming programme, owning engagements with governments and frontier AI companies. Own the end-to-end red-team hiring pipeline, including sourcing, work trials, and recruiting technical talent. Manage the RFP and opportunity pipeline by scoping engagements, drafting proposals, and supporting negotiations. Conduct analysis on team bottlenecks and ecosystem mapping to improve performance. Support technical writing, event organising, policy work, and grant applications as ...
Added Jun 25, 2026San Francisco, CARemoteAI Safety & Alignment$125K - $190K / year
AI Red Team Analyst Alignerr Remote, Global $15 - $75 per hour In this role, you'll conduct red-teaming exercises to uncover AI security weaknesses and deliver findings that improve system safety. Craft adversarial prompts, jailbreak attempts, and edge-case scenarios to challenge AI model guardrails. Evaluate AI outputs for safety violations, bias, and policy compliance. Document vulnerabilities and unexpected behaviours in structured reports for engineering teams. Collaborate with teams to recommend security mitigations and help refine testing protocols. Alignerr is a platform that for people to create data for frontier AI labs.
Research Fellowship, Evolution of AI Capabilities at the Frontier and Thresholds of Human Performance General-Purpose AI Policy Lab Paris, France In this role, you'll extend the Epoch Capabilities Index using Bayesian methods to map AI capability thresholds against human performance. Develop a multidimensional Rosetta Stone framework integrating human performance benchmarks. Identify and compare current AI capabilities against different levels of human expertise. Build statistical projections forecasting when AI systems surpass expert performance. Apply Bayesian modelling (Stan, PyMC, Pyro) and data science analysis throughout. The General Pu...
Added Jun 25, 2026Paris, FranceAI Governance & PolicySalary not disclosed
Research Manager / Research Managers, AIxCyber ERA Cambridge, UK £62,000 - £75,000 In this role, you'll lead the execution of ERA's AIxCyber Research Fellowship, supporting research fellows in developing and delivering their projects. Review and evaluate fellowship applications, then match fellows with mentors suited to their research focus. Help fellows scope research projects aligned with their skills and interests. Direct and oversee up to 5 research projects while providing guidance on planning, methodology, and publication. Organise workshops and events to support fellow professional development and strengthen relationships across the AI...
Added Jun 25, 2026Cambridge, UKAI Safety & Alignment£62K - £75K / year
Senior Software Engineer Irregular San Francisco Bay Area In this role, you'll design, build, and scale production systems that power evaluation and security platforms for frontier AI models. Architect and scale production-grade systems and workflows for AI model evaluation. Build backend services, APIs, and monitoring tools that support large-scale evaluations. Design infrastructure that enables research experiments to run efficiently at scale. Implement agent frameworks and build security challenges to test AI models' evasion capabilities. Irregular Labs is a frontier security lab that aims to protect the world in a time of increasingly cap...
Added Jun 25, 2026San Francisco, CAAI Safety & AlignmentSalary not disclosed
Operations Associate Centre for Long-Term Resilience London, UK £55,000 In this role, you'll support the operations of an AI or Biosecurity policy unit through administrative and project coordination. Coordinate meetings, events, travel logistics, team calendars, and provide executive assistance to the unit Director. Track project progress, activity logging, and provide flexible project management support to policy initiatives. Contribute to grant writing and funding proposals, ensuring consistent language across proposals. Manage operational processes including publication coordination, document filing, information security, budgets, and onb...
Added Jun 25, 2026London, UKAI Governance & Policy£55K / year
Fellowship Manager Singapore AI Safety Hub Singapore $90,000 - $120,000 In this role, you'll design and deliver the East-West AI Safety Fellowship, connecting researchers and government partners across Asia. Lead recruitment for 10-15 fellows and 5-6 mentors, building cohort culture and an active alumni network. Support government partnerships and relationships with universities, research institutes, and AI labs in the region. Develop operational infrastructure, track metrics, and produce impact reports for stakeholders. Contribute to public communications including reports and policy briefs. The Singapore AI Safety Hub is a nonprofit buildin...
Added Jun 25, 2026SingaporeAI Governance & Policy$90K - $120K / year
Project Manager, AI Governance Training AI Safety Asia Jakarta, Indonesia In this role, you'll lead operational delivery of AI governance training programs across ASEAN. Plan and execute training programs for government officials and civil society leaders across the region. Manage program logistics including scheduling, budgets, travel coordination, and materials preparation. Coordinate with trainers, facilitators, and stakeholders to ensure smooth program delivery. Track metrics, collect feedback, and contribute to continuous improvement of training content. AI Safety Asia is a nonprofit dedicated to building Asia as a globally-leading safe ...
Added Jun 25, 2026Jakarta, IndonesiaAI Governance & PolicySalary not disclosed
Join Apple's Responsible AI and Safety team as a Senior ML Researcher leading safety alignment and safeguards development for foundation models powering Apple products.
Added Jun 25, 2026Cupertino, CAAI Safety & Alignment$181.1K - $318.4K / year