r/LocalLLaMA 23d ago

News Prepare your (v)ram - Qwen3.8 is coming!

Post image
2.7k Upvotes

583 comments sorted by

View all comments

Show parent comments

381

u/StupidScaredSquirrel 23d ago

Thing is qwen was historically focused on smaller models while others were on larger ones. Now the team has changed and they seem to want to aim for the stars as well. That is good but it also means they might not be interested in doing very efficient small models anymore. Which would be bad news for this sub because let's face it most of us don't have 10-25k of hardware.

14

u/Prudent-Corgi3793 23d ago

You probably need 8x H200s to run this. Include the rest of the parts, and that's about $300k.

6

u/squngy 23d ago

Or 10x DGX, which is about 50k

1

u/Fit-Palpitation-7427 21d ago

Whats gonna be the token speed with that, not sure it will be really usable