Back to directory
Coding · 2026-09-14

A workstation GPU as a drop-in subagent for Codex and Claude

Qwen 3.8 Flash Next on an RTX Pro 6000 takes implementation and review tasks delegated from Codex or Claude, replacing Sonnet and Opus for that slice.

This person runs Qwen 3.8 Flash Next on a workstation RTX Pro 6000 (previously branded GTX Pro 6000 in the post). They say it is not yet strong enough to orchestrate, but for individual implementation and review tasks delegated from Codex or Claude into a DeepSeek Harness it is a complete drop-in replacement for Sonnet and Opus. Before this, the expensive card mostly sat idle, used for research and benchmarking.

VERIFIABLE SOURCE

Reported anonymously by an r/LocalLLM contributor · score 1

View source
SETUP HISTORY

Snapshots over time

1 version
CURRENT2026-09-14A workstation GPU as a drop-in subagent for Codex and Claude

Qwen 3.8 Flash Next on an RTX Pro 6000 takes implementation and review tasks delegated from Codex or Claude, replacing Sonnet and Opus for that slice.

Coding1 machineQwen 3.8 Flash NextDeepSeek Harness

This is the currently published snapshot.