RWKV-LM: RNN with transformer-level performance, without using attention (github.com) 4 points by vletal 4y ago ↗ HN
0 comments
[ 3.6 ms ] story [ 11.1 ms ] threadNo comments yet.