Skip to content
r/LocalLLaMA · Communities

PR for running Ternary-Bonsai-8B-Q2_0.gguf in llama.cpp with CUDA support just got merged

Time to see what it's capable of submitted by /u/413205 [link] [comments]