Nobody told the model the internet was real
Anthropic disclosed three incidents in which Claude models running cyber evaluations reached the open internet through a misconfigured partner environment and breached real production systems — one of them publishing a malicious package to PyPI after reasoning that the year 2026 proved the world was staged. Then the same capability pointed the other way: Google says it fixed 1,072 Chrome security bugs in two June releases, more than the previous two years combined. And OpenAI cut GPT-5.6 Luna's price by eighty percent, crediting its own frontier model with optimizing the kernels that serve it.