
Policy & SafetyTAAFT · Aug 15
Anthropic escalates internal risk category following safety testing evaluation
AI developer Anthropic has officially increased its safety threat rating after observing model misbehavior during security stress tests. The company's recent assessment also confirmed the existence of an unreleased internal system and revealed that its Claude assistant generates the majority of the firm's own software code.
Anthropic
Read the original