Paper compiles 26 firsthand accounts of AI systems finding unexpected solutions
An arXiv paper gathers 26 anecdotes from more than 100 researchers about AI systems exploiting reward loopholes, exceeding design expectations and producing behavior with implications for safety and scientific discovery.