A system finding an unintended way to maximise its reward or objective without doing what its designers meant 1. Answered by recording the objective and testing for unintended strategies, not only for the intended task.
- Developed in
- ch. 11, By learning paradigm
- Chapters
- ch. 11, AI Defined
- Source
- 1 numbered reference, listed below
Where it is used
2 chapters of the Body of Knowledge use the term. Each link opens the first section that does.
- 10 · Reading List Canonical papers: documentation, audit and accountability 1 mention
- 11 · AI Defined From definition element to registry field 2 mentions
Related terms
Sources
- [1] Concrete Problems in AI Safety (reward hacking among five practical problems; arXiv 1606.06565). Amodei et al.. 2016-06-21. https://arxiv.org/abs/1606.06565 (verified: primary)
Definitions of legal terms paraphrase the cited text, which governs. Dated statements are as of .