r/LocalLLaMA
· Communities
Those who use many layers in CPU/RAM and some in GPU – what are your specs and speeds?
I am trying to figure out if it's worth upgrading my RAM, but I've noticed that some MoE models don't seem to do well with many layers shared from VRAM --> CPU/RAM. This may be something on my end; a software config or perhaps my specific hardware config. This made me curious as to how many are doing this. I'm thinking