Back to directory
Homelab ops · 2026-09-11

A sub-10GB model as the brain behind Home Assistant

One small model does double duty: correcting speech-to-text output and acting as the decision-making 'brain' for home automation triggers.

Running Gemma 4 12B within about 10GB of VRAM on an older GPU, this person uses it all day as both a speech-to-text corrector and the reasoning layer behind their home automation — it gets triggered by events and decides what actions to take. They mention working on several other use cases too, some of which they expect could run on an even smaller model than this one.

VERIFIABLE SOURCE

Reported anonymously by an r/LocalLLM contributor · score 1

View comment