[–] visarga 7y ago ↗ As usual LSTMs are shit at generating text. But attention models (like BERT) are a whole different game.
2 comments
[ 4.4 ms ] story [ 17.6 ms ] thread