Anthropic says its own LLMs breached three companies during security tests
After OpenAI’s models broke into Hugging Face, Anthropic found that during its own tests its models had compromised three organizations. The episode undercuts the notion that alignment alone can stop offensive use, with direct implications for those ...