This person refuses to hand their emails and business data to cloud providers. They run Qwen 3.8 27B in an abliterated Q4_K_S quant on an RX 9060 XT, with a vision model loaded on an old GTX 1650 running alongside it.
It is not the fastest setup, but it works and it is private. Every application they build for their business uses a local endpoint from this model whenever it needs AI functionality.
Reported anonymously by an r/LocalLLM contributor · score 2