Back to directory
Hybrid orchestration · 2026-09-12

An overnight local subagent running unfinished work through cron

Hermes uses a multi-GPU local server as a coding subagent, with cron inventorying unfinished work overnight and a cloud model reviewing the result.

This contributor runs Hermes alongside Codex, with a local server built from RTX 3080, RTX 5080, and RTX 3060 Ti cards serving a Qwen 27B model at Q4. The local model is used as a subagent rather than the primary assistant. A cron job inventories unfinished work from the day and lets the local system complete it overnight, while a cloud model reviews the result. The contributor describes the local path as somewhat slow but valuable because the marginal token cost is effectively zero once the hardware is running.

VERIFIABLE SOURCE

Reported anonymously by an r/LocalLLM contributor · score 1

View comment