⟳ Репост от Tim Fernholz

This is so frustrating to read from The New Yorker. Reward hacking *long* predates LLMs. The Coast Runners paper on reward hacking (in reinforcement learning, pre-LLMs) is 10 years old! Reward hacking is not new, and not a sign of a breakthrough intelligence, much less sentience. Did AI write this?