The Llama.cpp Fork That Enables Qwen 3.8 27B Large Contexts for 16GB VRAM GPU(github.com)5 points by dazhbog 40 days ago | 2 comments