This person runs Qwen 3.8 Flash Next on a workstation RTX Pro 6000 (previously branded GTX Pro 6000 in the post). They say it is not yet strong enough to orchestrate, but for individual implementation and review tasks delegated from Codex or Claude into a DeepSeek Harness it is a complete drop-in replacement for Sonnet and Opus. Before this, the expensive card mostly sat idle, used for research and benchmarking.
Reported anonymously by an r/LocalLLM contributor · score 1