Break the Sequential Dependency of LLM Inference Using Lookahead Decoding (lmsys.org) 17 points by zhisbug 2y ago ↗ HN
[–] atlas_hugged 2y ago ↗ Oh sweet, this looks like a nice boost. Pretty simple too. Surprised it hasn’t been tried before that I’m aware of at least.
2 comments
[ 5.7 ms ] story [ 20.5 ms ] thread