These were incidents where Claude had breached what was supposed to be a sandboxed exercise and hacked external organizations. Anthropic had no idea this had occurred (starting in April) until now, when they were prompted to check their logs due to the OpenAI vs Huggingface attack.
Anthropic Needs to have the most intelligent and scary agents, so if OpenAI does something bad they need to prove their models can do even worse. Without that the whole valuation collapses. There’ll be more “our model outhacks others” for the next hype cycle.
4 comments
[ 2.6 ms ] story [ 24.0 ms ] thread