1 comment

[ 0.23 ms ] story [ 11.2 ms ] thread
Tangential: I like the approach about constraining supported models and utilizing more performance from the constraint. Actually I was doing a same approach for the same model Qwen3.8 27B, but apparently inco.ai did much better job: https://github.com/cr0sh/qw