Misuse Red Team - Research Engineer/Research Scientist
This role involves researching and developing novel attack vectors against large language model safeguards to identify vulnerabilities before malicious actors can exploit them. You'll design automated attack systems, evaluate LLM robustness through methods like jailbreaking and data poisoning, and collaborate with leading AI labs to improve defenses. The work directly informs policy and safety practices at frontier AI companies and government levels.