ABOUT
A wiki about
the practice.
vram.wiki documents what people actually do with local language models: the hardware, the stack, the task and the result.
This project was born from an r/LocalLLM thread in September 2026. Every setup comes from a community answer, is rewritten and anonymized, and stays linked to its source comment.
This is a community-built snapshot.
The people who shared those answers are what make the dataset useful. If you see your setup here, thank you — and please correct or extend it if anything is missing.
Our most important rule
When a detail is not stated, it stays unknown. No VRAM guessed from a GPU name, no invented usernames, no promises turned into facts.