Workplace
hybrid
Employment
Internship
Published
Aug 22, 2026
Closes
No date supplied
The role
About Anthropic
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.
About the role
Anthropic's Safeguards organization builds the policies, evaluations, and enforcement systems that define and hold the limits on how Claude can be used. In this role, you'll own our conventional weapons work.
Defining a crisp boundary between acceptable and harmful requests in this domain is difficult, because the underlying components are dual-use: the same capabilities that serve civilian engineering and research can also contribute to a weapons system. Building the threat models, evaluations, and detection systems that hold that boundary is the core of this role.
Conventional weapons are increasingly defined by software, and the risks reach well beyond firearms: from a model operating a weapons system or writing guidance code, to the weaponization of dual-use platforms. The role spans every weapon class, up to autonomous systems that select and engage targets without human authorization.
You will help define the line between prohibited weapons development and legitimate research and engineering work, and how to operationalize the distinction. This will include translating technical judgment into principles that engineers can implement and enforcement teams can act on.
Key responsibilities
•
Own and maintain Anthropic's conventional weapons policy, defining the boundary between the activities our models should and should not support
•
Build the threat models and evaluations that measure how our models could contribute to weapons development, including the software and autonomy compo
Requirements
Department: Safeguards (Trust & Safety)