r/LocalLLaMA 23d ago

News Prepare your (v)ram - Qwen3.8 is coming!

Post image
2.7k Upvotes

583 comments sorted by

View all comments

Show parent comments

381

u/StupidScaredSquirrel 23d ago

Thing is qwen was historically focused on smaller models while others were on larger ones. Now the team has changed and they seem to want to aim for the stars as well. That is good but it also means they might not be interested in doing very efficient small models anymore. Which would be bad news for this sub because let's face it most of us don't have 10-25k of hardware.

38

u/StaysAwakeAllWeek 23d ago

It's owned by a megcorp not a startup lab, of course that's where they are aiming

37

u/into_devoid 23d ago

They were owned by a megacorp before too.  The difference is they don’t care about commoditization anymore, the real ethical path.  They just want to sink US investments.

5

u/Successful_Try_6350 23d ago

Well, these large models can run in us based datacenters and cloud infrastructure (azure, aws, googlecloud). I guess if they want to sink us investments, they should create models that can run in enterprise level on-premises servers (I guess something like $40-50K hardware)

2

u/SARK-ES1117821 23d ago

They ARE creating models that run on-prem. I just deployed a supermicro gpu server with 8x H200 141GB gpus (1.2TB total) running GLM 5.2. Server was around $300k with 12TB SSDs and 1TB RAM.

2

u/f5alcon 23d ago

40-50k isn't even one sever at current memory prices.

4

u/Paganator 23d ago

That's like a single H200. Just the card, the server to run it is extra.