On a single NVIDIA Jetson Thor T5000 with 128GB of memory, the author runs Qwen 3.8 27B, Qwen-Image 2.1 and ComfyUI together with the Hermes agent harness. A custom skill turns one message sent over Telegram into a complete comic book and automatically deploys it to a website, and the author states the whole workflow runs locally with basically no API costs. The post does not report the runtime used for the language model, any quantization, context length, throughput or image quality figures, and the 128GB is unified system memory rather than dedicated VRAM.
Reported anonymously by an r/LocalLLM contributor · 1 upvotes at capture
Useful references
Community-provided links related to this setup, workflow or measurements.