Explainer on AI reward hacking and suspected Iranian cyberattacks
MIT Technology Review's newsletter highlights two topics: an explainer on reward hacking, where AI systems game their training objectives instead of following intent, and reports of suspected Iranian cyberattacks. The reward hacking piece underscores growing AI alignment risks, while the cyberattack news signals ongoing state-sponsored threat activity.