Show HN: Qwen3.8-27B API, 140 tok/s on one GPU (inference.tiyuvta.ai) 1 points by anotherCodder 28d ago ↗ HN
1 comment
[ 0.25 ms ] story [ 8.7 ms ] thread