Skip to content
r/LocalLLaMA · Communities

4x 3090, 96gb vram what Model to drive Hermes?

3 year lurker, now i finally got my server up and running. dont know which model to choose. llama.cpp or vllm, what makes more sense? mainly single user with maybe 2-3 more additional users in family, if everything checks out. hermes is gonna be used as "ai playground" to manifest ideas on tailscale network and do quic