Job description
Meta • San Francisco Bay Area
San Francisco Bay Area
$184,000 - $257,000
-
In this role, you'll design and implement novel safety alignment techniques for large language models and multimodal AI systems at Meta.
-
Create and curate high-quality datasets for safety alignment, including adversarial and borderline prompts.
-
Fine-tune and evaluate large language models to adhere to Meta's safety policies and evolving standards.
-
Build scalable infrastructure and tools for safety evaluation, monitoring, and rapid mitigation of risks.
-
Lead complex technical projects end-to-end with researchers, engineers, and cross-functional partners.
Meta is an American technology company. They have an AI division that focuses on AI Infrastructure, Generative AI, NLP, and Computer Vision, among other topics. You can read concerns about doing harm by working at a frontier AI company in our career review on the topic.
Applications are handled by the employer.