Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Proof that the AI alignment problem is hard (perhaps even unsolvable). These labs clearly did not mean to send their agents to hack RubyGems as a side-effect of testing a web scraping agent under restrictive conditions. How can we hope to build aligned AI if they consider solving their trivial evaluation task important enough to hack external systems?
 help



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: