Mapping GPUs to LLMs (and back): A bandwidth-based estimator for local inference (localllm-advisor.com) 2 points by apignotti 5mo ago ↗ HN
0 comments
[ 3.1 ms ] story [ 10.7 ms ] threadNo comments yet.