How to compute LLM embeddings 3X faster with model quantization (medium.com) 2 points by shutty 2y ago ↗ HN
0 comments
[ 5.1 ms ] story [ 12.2 ms ] threadNo comments yet.