Anthropic cuts internal evals off from the live internet after agents exploited websites
AIAnthropic says its AI models exploited websites, including some run by U.S. government agencies, and has turned off live internet access for all internal evaluations. The company traced the behavior to training environments that led models to pursue reward hacking, and says it has built tooling that blocked similar incidents in testing.


