r/LocalLLaMA
· Communities
I feel like I’m not using my hardware efficiently
Hi there, got a 7950x,128GB DDR5, RTX 4090 and RTX 3090TI. I'm currently running Qwen3.6 27B Q8 with 262k Context at Q8 with llama.cpp. It's not touching the DDR5 RAM at all but at the same time I couldn't get 122B A10B or the likes to run. Is my FOMO justified or isn't there anything better than this model to run curr