A reminder of a recent discussion here that goes into a lot more detail about why reinforcement learning works well for specialized domains like Go but is having a very hard time generalizing to more "real-world" types of tasks: https://news.ycombinator.com/item?id=16383264