TMLR: Outcome-Based Reinforcement Learning to Predict the Future (openreview.net) 4 points by bturtel 9mo ago ↗ HN
1 comment
[ 3.6 ms ] story [ 22.7 ms ] thread