This daily local user runs Qwen 3.8 27B on an RTX 4070 for Python and TypeScript work, and is upfront that it's not as good as Claude for complex refactoring — but calls it good enough for about 80% of what they actually do. Privacy matters more than expected once client data enters the picture, since it removes the need to think about what's being sent to a third party. The feature they value most, though, is iteration speed: no rate limits and no token costs mean they can run ten completions in the time it takes to write one carefully worded prompt for Claude.
Reported anonymously by an r/LocalLLM contributor · score 1