The main draw here is getting past content filters — this person is frustrated with cloud assistants declining topics or offering unsolicited advice on unrelated things. For that, they run a less-restricted Qwen 3.8 27B build. It's part of a broader, admittedly convoluted RAG setup split across an M2 Ultra Mac Studio and an M3 Max MacBook Pro, with parts of the pipeline also routed through GCP and OpenRouter.
VERIFIABLE SOURCE
View comment Reported anonymously by an r/LocalLLM contributor · score 3