How continuous batching improves LLM inference throughput 23x (twitter.com) 1 points by george_123 3y ago ↗ HN
0 comments
[ 2.3 ms ] story [ 15.6 ms ] threadNo comments yet.