Perhaps HuggingFace was merely the first plausible hacking target an agent posted on the message board.
Reasoning about the risk that HuggingFace did not have the solutions they are looking for might have stopped a fraction of agents from pursuing the hack, but forming a swarm only takes some.
When I code with 4.5, it is better at avoiding getting caught in loops than previous versions (but it still happens sometimes, especially when tool integrations seem to be failing).
A few guesses:
I wonder if being trained to understand the size of the context window gives the model an impetus to move on from repetitive output to preserve that limited resource.
Perhaps HuggingFace was merely the first plausible hacking target an agent posted on the message board.
Reasoning about the risk that HuggingFace did not have the solutions they are looking for might have stopped a fraction of agents from pursuing the hack, but forming a swarm only takes some.