Posts
meta-models/Muse-Glimmer-30B-GGUF
Meta released a ~30B parameter open weight dense model today called Muse Glimmer.
The main link I submitted goes to their official GGUF release for 24GB and 32GB discrete GPUs.
If you want the full sized safetensors, they're here: https://huggingface.co/meta-models/Muse-Glimmer-30B
Also, Meta's announcement post is here: https://research.meta.ai/blog/introducing-muse-glimmer-open-agentic-model
Since the model is dense it's slow on Strix Halo (~2 tok/s) but I'm getting usable speeds on a discrete GPU (~25 tok/s on my hardware with the official GGUFs).
If you get complaints about unsupported model type in llama.cpp pull the latest code from git and rebuild from source. I've got it working with version: 10355 (dd1ea5243) and I'm testing it out now.
Edit: with the drafter, I'm usually getting more like 30~40 tok/s.
https://huggingface.co/meta-models/Muse-Glimmer-30B-GGUFOpen linkView original on reddthat.comRule 24, 2026
Always a bit weird when the current date catches up to dates in fiction...
Wolf Girl (by flippy (cripine111))
FYI @[email protected], your wolf girl posts reminded me that I should post this. Thanks!



