The Llama.cpp Fork That Enables Qwen 3.8 27B Large Contexts for 16GB VRAM GPU (github.com) 4 points by dazhbog 13d ago ↗ HN
[–] dazhbog 13d ago ↗ A video I found covering the improvementshttps://www.youtube.com/watch?v=n_ggLjIgRcM
[–] akshay_akula 13d ago ↗ Will this perform better than lmstudio-community/Qwen3.8-27B-MLX-4bit on my m5 max 48gb memory mac?
3 comments
[ 0.20 ms ] story [ 2.1 ms ] threadhttps://www.youtube.com/watch?v=n_ggLjIgRcM