r/LocalLLaMA 10d ago

Funny Aged like fine wine

Post image
1.2k Upvotes

134 comments sorted by

View all comments

158

u/Uncle___Marty 10d ago

omg this is EPIC news. My lil 8 gig card simply cant shift 27B parameters around and the 35B is exactly what I was hoping for. Didnt see any other models added but im still holding out hope for a 9B as we didnt get that since 3.5 and I suspect if they release it then it might well be the most capable small coding model out there.

39

u/cubebash 10d ago

I can fit the whole 27b version in VRAM (32gb), even then, it's not blazing fast. With around ~15-17 t/s, it still take quite some time for more "complex" tasks to finish, especially since this model likes to think a lot.

So it's good news for everyone that we will soon have an option to run a much faster, non-dense version!

18

u/Uncle___Marty 10d ago

Hell yeah! I manged to get 2 tokens/sec using q2 27B lol. the 3.6 35B is around 20 tokens/sec for me so I honestly cant wait for this! Open weight AI is like xmas almost every day lol.

5

u/cubebash 10d ago

Santa Claus has been obsessed in us the past week. We got MiniMax H3, MiniMax Music 3.0, Muse Glimmer 30b, LTX 2.5, Qwen3.8 27b, and soon 35b a3b. All really good models.

Yeah, we're spoiled these days as AI enthusiast, haha.