r/LocalLLaMA
· Communities
Qwen3.8 2.4T UD-Q1_0 – 178 token generation – 11 min 38s – 0.25 tokens/sec
https://preview.redd.it/nqr9nj5028jh1.png?width=1516&format=png&auto=webp&s=47af40a0737429dc65f2e455b87b2db81725cb14 So I wanted to see... is it possible/viable to run this perhaps once in a while some hard task.. yea... no. Even with dual 5090's and 3 3090's and 96GB system ram.. still far exceeds my vram + ram by dou