This person runs Qwen 3.8 27B locally with SGLang in NVFP4 quantisation and reports it working really well and impressively capable. They set the maximum context at 180K on purpose, to avoid pushing deep into the range where they feel output quality degrades.
VERIFIABLE SOURCE
View source Reported anonymously by an r/LocalLLM contributor · score 4