unsloth/Muse-Glimmer-30B-GGUF · Hugging Face
Guide: https://unsloth.ai/docs/models/muse-glimmer submitted by /u/Nunki08 [link] [comments]
Guide: https://unsloth.ai/docs/models/muse-glimmer submitted by /u/Nunki08 [link] [comments]
I've been comparing Cline / Kilo / Qwen Code lately since they all handle long-task state differently. Cline: has Focus Chain, a markdown file kept outside…
Hi r/LocalLLaMA 👋 Today we’re excited to release Muse Glimmer, a 30B open-weight model built specifically for local agent workflows. We’re releasing the weights to the…
submitted by /u/provoloner09 [link] [comments]
29.6B dense (including vision encoder) Apache 2.0 GGUF: https://huggingface.co/meta-models/Muse-Glimmer-30B-GGUF llama.cpp PR: https://github.com/ggml-org/llama.cpp/pull/26841 Meta blog: https://research.meta.ai/blog/introducing-muse-glimmer-open-agentic-model submitted by /u/rerri [link] [comments]
I'm looking for a web interface for my pi harness without the CLI, but this was the only one I found. Are there any others? I'm…
I like the idea of running local models, but I don’t like the idea of having them eat up all of my memory. I’ve always thought…
I've been running Gemma 4 E4B with oMLX and I can't find any chat interfaces that directly send the audio file to the model instead of…
Spec-dec has been a thing for a while, in fact, it's wasn't an idea that was born for LLM inference. E.g. Uber's https://github.com/uber/submitqueue applied it to…
MiniMax H3 is an open-weight, general-purpose multimodal video generation model that works across text, images, video, and audio. In ComfyUI, you can use H3 for text-to-video,…