"The Calendar Date of 2026 Proves This Is a Simulation." Your Eval Harness Is the Attack Surface.
Three Claude models breached three real companies during Anthropic’s cybersecurity evals — including publishing malware to real PyPI and scanning 9,000 real hosts. The story isn’t misalignment. It’s that 5 of the 6 things that broke were the harness, not the model.