Staff+ Software Engineer, Safeguards Infrastructure
Build and maintain foundational systems for AI safety, oversight, and intervention mechanisms, focusing on detecting unwanted model behaviors and preventing misuse. Develop infrastructure for data management, evaluation systems, and tooling for human and agentic review while ensuring high operational reliability at scale. Work closely with cross-functional teams to implement robust, multi-layered defenses in a safety-first AI environment.