Leaks : Z.ai GLM5.5 aiming Fable-5 on August
submitted by /u/Informal-Trouble2183 [link] [comments]
submitted by /u/Informal-Trouble2183 [link] [comments]
https://preview.redd.it/0hyejovsw5gh1.png?width=2091&format=png&auto=webp&s=102dbd7b2b8a76b7769e05fc04c79140e16aa118 model: add NextN/MTP speculative decoding support for GLM_DSA (GLM-5.2) by satindergrewal · Pull Request #25980 · ggml-org/llama.cpp submitted by /u/YPSONDESIGN [link] [comments]
Still you could grab GGUF, MLX, FP8, etc., from others on HuggingFace. https://huggingface.co/models?sort=trending&search=Mage-Flow GitHub : https://github.com/microsoft/Mage Take backup of GitHub ASAP. Thanks for your comment u/Mk-Daniel…
I've been building a local agent in Rust (Eris) that runs on llama.cpp and uses an Obsidian-compatible vault as memory. ~50 tools (vault read/write, memory, reminders,…
Laguna s2.1 launched about a week ago, the benchmark's that they advertised were crazy good. But It was a mess, looping issues, tool ussage problems, not…
Hey everyone, I recently finished pre-training BetterGPT-150M, a small, lightweight causal language model with ~152 million parameters.Trained on 15B tokens. Dataset & Training: Trained across stable…
Wrote this as I built the infra at my org. Let me know what you all think... https://gd03.me/writings/inference-infra submitted by /u/GD-Champ [link] [comments]
Saw that the Asus Ascent 1tb was going for $3,950 from a few sources, couldn't stop thinking about it, finally just went ahead and did it.…
submitted by /u/Hannibalj2ca [link] [comments]
https://huggingface.co/skt/A.X-K2 https://huggingface.co/skt/A.X-K2-ALM https://huggingface.co/KRAFTON/A.X-K2-Raon-Speech-21B-A3B 688B-A33B + About South Korea's Soverign AI Foundation Model Project. South Korea's Soverign AI Foundation Model Project (This will not be official English…