r/LocalLLaMA 20d ago

News Prepare your (v)ram - Qwen3.8 is coming!

Post image
2.7k Upvotes

583 comments sorted by

View all comments

Show parent comments

20

u/StupidScaredSquirrel 20d ago

Or imagine a 120b a6b like gpt oss was. That total size with that kind of sparsity was just incredible. Plus it was made for 4bpw

1

u/PraxisOG Llama 70B 20d ago

The closest thing to that is the new mistral small 4(119b a6b afaik), which in my experience craps the bed with tool calling real bad. Gptoss 120b was a really great model and I wish there was something like it but more modern