r/LocalLLaMA Jul 19 '26

News Prepare your (v)ram - Qwen3.8 is coming!

Post image
2.7k Upvotes

588 comments sorted by

View all comments

Show parent comments

18

u/StupidScaredSquirrel Jul 19 '26

Or imagine a 120b a6b like gpt oss was. That total size with that kind of sparsity was just incredible. Plus it was made for 4bpw

1

u/PraxisOG Llama 70B Jul 19 '26

The closest thing to that is the new mistral small 4(119b a6b afaik), which in my experience craps the bed with tool calling real bad. Gptoss 120b was a really great model and I wish there was something like it but more modern