Running a 4-bit Qwen 3.8 quant with a 200K context window on a 24GB RTX 3090 and 96GB of system RAM, this person builds complete apps — web and desktop OS apps across several languages — through OpenCode, leaving it running until it needs a decision from them. They lean on a 'planning with files' skill to keep track of progress across context compaction, plus several MCP servers to keep active context low, and run reasoning effort on the highest setting even though it burns more tokens weighing options before committing.
VERIFIABLE SOURCE
View comment Reported anonymously by an r/LocalLLM contributor · score 3