GPT-4 Trained on 10x More Data, Can Write 60k Word Books (twitter.com) 3 points by baptiste313 3y ago ↗ HN
[–] minimaxir 3y ago ↗ This misinformed GPT-4 take is new levels of comically bad because it correlates number of model hyperparameters with the quantity of training data.
1 comment
[ 2.5 ms ] story [ 14.8 ms ] thread