This person has run a home lab setup for about a month and keeps a model stable for days at a time. They built a finance-tracking web app for themselves and their family with a Hermes agent and GLM 5.3 Flash, reporting 15-20 tok/s normally and 12-15 tok/s across two concurrent sessions, with up to four requests at a 260K context.
The app is designed to include an MCP so the local agent can add expenses and update bill costs. It receives forwarded emails for credit card charges, and the owner can also send it expenses through Telegram or Discord or add them manually. They built the whole app in one weekend of vibe coding and question whether the subscription services doing this are on the way out.
Reported anonymously by an r/LocalLLM contributor · score 2