The Technical Program Manager for Safeguards will oversee the operational health of AI safety infrastructure, manage incident responses, and ensure reliability of safety-critical systems.
Curated roles
Browse roles focused on reducing risks from advanced AI systems, including alignment research, safety engineering, evaluations and related operations.
364 active roles found.
The Technical Program Manager for Safeguards will oversee the operational health of AI safety infrastructure, manage incident responses, and ensure reliability of safety-critical systems.
The Technical Program Manager for Research at Anthropic will define and build programs for research teams, focusing on AI safety, alignment, and evaluations.
The role focuses on enhancing the reliability of AI systems at Anthropic, ensuring safety commitments and robust performance across critical serving paths.
The Staff Software Engineer, Inference at Anthropic will work on building and maintaining systems for AI model inference, focusing on performance optimization and societal impacts.
The Staff Software Engineer in AI Reliability Engineering will enhance the reliability of AI systems, focusing on critical service paths and incident response.
The Staff Software Engineer in AI Reliability Engineering at Anthropic will focus on improving the reliability of AI systems, ensuring robust performance and safety commitments.
The Staff Research Engineer will work on developing AI systems aimed at scientific discovery, focusing on removing bottlenecks in achieving scientific AGI and creating evaluation frameworks for model capabilities.
The role involves developing systems for safety and oversight of AI models, focusing on misuse prevention and monitoring model behaviors.
The Software Engineer, Safeguards role at Anthropic involves developing systems to monitor AI models, prevent misuse, and ensure user safety.
The Senior+ Software Engineer for Research Tools at Anthropic will develop infrastructure and applications to support AI safety research, focusing on model evaluation and understanding model behavior.
The Safeguards Enforcement Analyst will ensure AI models meet safety and policy standards through evaluations and mitigations, collaborating with cross-functional teams.
The Research Scientist, Interpretability role at Anthropic focuses on mechanistic interpretability to enhance the safety and understanding of AI systems.
The Research Engineer, Universes role at Anthropic focuses on developing training environments for safe AI systems and includes responsibilities for building evaluations to measure AI capabilities.
Technical research role on Anthropic's Frontier Red Team focused on evaluating and defending against advanced AI-enabled cyber threats, including autonomous capability evaluations, safety defenses, and policy-relevant demonstrations.
The role involves conducting research on AI safety and alignment, focusing on understanding and steering the behavior of powerful AI systems.
Research Engineer/Scientist on Anthropic’s Alignment Science team, conducting experimental AI safety research on powerful future systems, safety evaluations, alignment stress-testing, and related safeguards work.
The role focuses on building reliable and interpretable AI systems, emphasizing safety and societal impacts.
The Research Engineer/Research Scientist role at Anthropic focuses on developing large language models with an emphasis on safety, alignment, and societal impacts.
The role involves research and engineering to develop safe and trustworthy large language models, focusing on multimodal capabilities and ethical implications of AI.
The Research Engineer will enhance AI model safety and alignment through post-training techniques, impacting the quality and capabilities of production models.