EfficientQAT: LLM Quantization, gets a 2-bit llama2-70B outperform regular 13B (old.reddit.com) 21 points by jackbravo 2y ago ↗ HN
0 comments
[ 2.8 ms ] story [ 11.5 ms ] threadNo comments yet.