r/LocalLLaMA
· Communities
PR for running Ternary-Bonsai-8B-Q2_0.gguf in llama.cpp with CUDA support just got merged
Time to see what it's capable of submitted by /u/413205 [link] [comments]
Time to see what it's capable of submitted by /u/413205 [link] [comments]