Beyond acting as the main model behind their Hermes agent, this person's Qwen 3.8 27B — running on a 32GB Radeon Pro AI R9700 — also serves as an MCP tool wired into Codex. Their setup has Codex send a proposal to the local model before starting a task and then have the local model review the final output before anything is marked complete; each side gets one chance to rebut the other, and if they can't reach agreement, the task simply stops for a human to break the tie rather than letting either model force it through.
Reported anonymously by an r/LocalLLM contributor · score 2