Skip to content
r/LocalLLaMA · Communities

model: add openPangu-2.0-Flash (92B-A6B) with MLA-latent cache, DSA/SWA, mHC, and multi-head MTP by joelfarthing · Pull Request #2065 · ikawrakow/ik_llama.cpp

openPangu-2.0-Flash - 92B A6B & 512K Context Length Model : https://huggingface.co/openpangu/openPangu-2.0-Flash/blob/main/README_EN.md GGUF for ik_llama folks : https://huggingface.co/ji-farthing/openPangu-2.0-Flash-ik-llama-GGUF submitted by /u/pmttyji [link] [comments]