r/LocalLLaMA
· Communities
Script to monitor llama cpp and analyze memory usage
My goal has always been to be productive with commodity hardware. So far my workhorses have been the MoE editions of gemma 4 and Qwen 3.6 on an old desktop with a single 9060XT with 16GB ram. The problem has always been that every source is vague about Vram/ram requirements. Models are trained at 16 bits, many guides s