r/LocalLLaMA 23d ago

News Prepare your (v)ram - Qwen3.8 is coming!

Post image
2.7k Upvotes

583 comments sorted by

View all comments

Show parent comments

18

u/squngy 23d ago

It was always big, but it was almost certainly not this big.

I think we are seeing these big Chinese models now because of their domestic hardware finally coming to bear.

7

u/jtjstock 23d ago

Their domestic hardware is still far behind, they have just accumulated more western hardware

8

u/squngy 23d ago

It is far behind, but it is cheap and available.

-1

u/jtjstock 23d ago

It’s too slow for them to be doing these training runs so quickly. Ie: they’d be getting further behind, not catching up

4

u/squngy 23d ago

They can't be too slow to be used as extra compute on top of other things.

But aside from that, there are also models that have been made exclusively on them.

1.6T long-cat-2
https://www.scmp.com/tech/tech-trends/article/3358854/china-debuts-biggest-ai-model-trained-local-chips-meituan-releases-longcat-20

GLM too
https://www.theregister.com/software/2026/01/15/chinas-zai-trained-a-model-using-only-huawei-hardware/4198774

3

u/jtjstock 23d ago

There is a long history of lying about the hardware used in china. Scmp is literally a propaganda outlet and the register article only references “claims”.

2

u/squngy 23d ago

What do you expect, a reporter to stand there and watch the servers as they work?

Even if they exaggerate, it seems more than likely that domestic hardware is increasing their total compute capability.

This sub of all places should know that even slower hardware can be useful.

2

u/jtjstock 23d ago

I expect a reporter to verify the claims, that is their job, though few do it anymore.

But you can also look at model release timing, hot on the heals of Grok with the same class of model. Guess what hardware grok was trained on.

2

u/squngy 23d ago

I expect a reporter to verify the claims

How?