ainotis Join
My notis

Checked fact 9219 Oct 2026Safety and security

The paper "Concrete Problems in AI Safety" (Amodei, Olah, Steinhardt, Christiano, Schulman, Mane; arXiv 1606.06565, v2 25 July 2016) lists "avoiding reward hacking" as one of five research problems on accident risk.

The exact words it rests on

Authors: Dario Amodei, Chris Olah, Jacob Steinhardt, Paul Christiano, John Schulman, Dan Mané ... last revised 25 Jul 2016 (this version, v2) ... We present a list of five practical research problems related to accident risk, categorized according to whether the problem originates from having the wrong objective function ("avoiding side effects" and "avoiding reward hacking")

What the source said when we opened it, on 9 Oct 2026.

The source

Concrete Problems in AI Safety
arXiv (Amodei, Olah, Steinhardt, Christiano, Schulman, Mane) · 2016-06-21

Checked

Checked by the notis newsroom on , against the source above.

In the story

How reward hacking lets a model game its score 9 Oct 2026

Cite this fact

Anyone may quote this address. It does not change; if we correct the story, this page says so.