Engineering Manager, Safeguards
Lead the development of internal review and enforcement tooling used by safety investigators and AI systems to detect and respond to potential harms across Anthropic's platforms. Build a scalable, privacy-preserving platform with analytics, sandbox environments, and human-in-the-loop automation, enabling both human reviewers and Claude to act effectively. Partner with policy, legal, and data science teams to ensure systems are trustworthy, compliant, and aligned with evolving safety requirements.