1 comment

[ 3.5 ms ] story [ 15.5 ms ] thread
I am curious whether this theory can explain some phenomena of large language models.