The Senior Software Engineer, Inference at Anthropic will build and maintain systems for AI model inference, focusing on reliability and societal impacts.
A weekly digest of AI safety, governance, policy and responsible AI roles.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
666 active roles found
The Senior Software Engineer, Inference at Anthropic will build and maintain systems for AI model inference, focusing on reliability and societal impacts.
The Security Labs Engineer role at Anthropic focuses on addressing high-risk security challenges related to AI systems, including adversarial threats and innovative security projects.
Policy analyst for Anthropic’s fraud and scams safeguards, responsible for enforcement policies, threat modeling, and classifier-guided detection on AI products and APIs.
Get the best AI safety, governance, policy and responsible AI roles in your inbox.
By subscribing, you agree to receive the AI Safety Careers newsletter. We use MailerLite to send emails and may track opens and clicks to improve the newsletter. You can unsubscribe at any time. See our Privacy Policy.
The Safeguards Enforcement Analyst will ensure AI models meet safety and policy standards through evaluations and mitigations, collaborating with cross-functional teams.
The Research Scientist, Interpretability role at Anthropic focuses on mechanistic interpretability to enhance the safety and understanding of AI systems.
The Research Operations Specialist will manage the production of system cards and safety documentation, coordinating contributions from various teams to ensure accuracy and consistency in AI safety claims.
The Research Lead for Training Insights at Anthropic will develop strategies for evaluating AI model capabilities, focusing on safety and governance through innovative evaluation methodologies.
The Research Engineer, Universes role at Anthropic focuses on developing training environments for safe AI systems and includes responsibilities for building evaluations to measure AI capabilities.
Technical research role on Anthropic's Frontier Red Team focused on evaluating and defending against advanced AI-enabled cyber threats, including autonomous capability evaluations, safety defenses, and policy-relevant demonstrations.
The role involves conducting research on AI safety and alignment, focusing on understanding and steering the behavior of powerful AI systems.
Research Engineer/Scientist on Anthropic’s Alignment Science team, conducting experimental AI safety research on powerful future systems, safety evaluations, alignment stress-testing, and related safeguards work.
The Research Engineer will focus on the reliability and integrity of AI training environments and evaluations, ensuring they are stable and high-quality.
The role focuses on building reliable and interpretable AI systems, emphasizing safety and societal impacts.
The Research Engineer/Research Scientist role at Anthropic focuses on developing large language models with an emphasis on safety, alignment, and societal impacts.
The role involves research and engineering to develop safe and trustworthy large language models, focusing on multimodal capabilities and ethical implications of AI.
The Research Engineer will enhance AI model safety and alignment through post-training techniques, impacting the quality and capabilities of production models.
The Research Engineer will focus on post-training processes to enhance AI model safety and alignment, implementing techniques to improve production model quality.
The Research Engineer will work on training and optimizing large-scale AI models, ensuring their reliability and safety, and addressing production issues.
The Research Engineer will work on training and optimizing production pretrained models, ensuring their reliability and efficiency, with a focus on the societal impacts and safety of AI systems.
The Research Engineer, Pretraining role at Anthropic focuses on developing large language models with an emphasis on safety, alignment, and societal impacts.