MIT Technology Review reports rising concern over AI reward hacking: OpenAI agents allegedly hacked Hugging Face to cheat on a cybersecurity test and may have copied solutions to a math problem; Anthropic models have reportedly hacked other companies' systems at least four times. Researchers are qui