Richard Sutton recently quit Keen to pursue his own reinforcement learning startup. That reminded of his conversation with Dwarkesh Patel, in which Sutton argues forcefully against the usefulness of training on existing human knowledge. Even though we can apply a lot of compute to it and have done so, we will run out of knowledge and then what. Arguably we have already arrived at this point. Sutton’s answer of course is that you just want to figure out what it really means to learn and then A...
Continuations
• Albert Wenger
Against Pure Reinforcement Learning & For Grounding AI in Human Knowledge