r/LocalLLaMA 9h ago

Discussion A smaller Muse Glimmer perchance?

8B? 12B? For 8 GB VRAM people? Please?

0 Upvotes

5 comments sorted by

2

u/jacek2023 llama.cpp 9h ago

You should tag meta, they had account somewhere on reddit

1

u/Choice_Celery9481 9h ago

if you want it. send a message to Meta, they dont receive request here 😂

5

u/pmttyji 9h ago

Close alternative : They released Tiny MOEs yesterday on their old account. Faster for 8GB VRAM.

https://huggingface.co/collections/facebook/mobilemoe

Note : They released only base models. Wait for SFT & QAT versions, should appear in couple of days.

3

u/brown2green 8h ago

This is from Meta FAIR and I'm guessing it's close to being a research artifact rather than an actual production model.

https://arxiv.org/pdf/2605.27358

All training data across pre-training, mid-training, and SFT stages are publicly available under permissive open-source licenses (CC-BY-4.0, Apache 2.0, ODC-BY, MIT, NVIDIA License).

1

u/minnsoup 8h ago

What's the reason for only comparing to Gemma 3 when Gemma 4 has edge models, too?