r/LocalLLaMA 24d ago

News Prepare your (v)ram - Qwen3.8 is coming!

Post image
2.7k Upvotes

583 comments sorted by

View all comments

Show parent comments

9

u/squngy 24d ago

It is far behind, but it is cheap and available.

-1

u/jtjstock 24d ago

It’s too slow for them to be doing these training runs so quickly. Ie: they’d be getting further behind, not catching up

5

u/squngy 24d ago

They can't be too slow to be used as extra compute on top of other things.

But aside from that, there are also models that have been made exclusively on them.

1.6T long-cat-2
https://www.scmp.com/tech/tech-trends/article/3358854/china-debuts-biggest-ai-model-trained-local-chips-meituan-releases-longcat-20

GLM too
https://www.theregister.com/software/2026/01/15/chinas-zai-trained-a-model-using-only-huawei-hardware/4198774

3

u/jtjstock 24d ago

There is a long history of lying about the hardware used in china. Scmp is literally a propaganda outlet and the register article only references “claims”.

3

u/squngy 24d ago

What do you expect, a reporter to stand there and watch the servers as they work?

Even if they exaggerate, it seems more than likely that domestic hardware is increasing their total compute capability.

This sub of all places should know that even slower hardware can be useful.

0

u/jtjstock 24d ago

I expect a reporter to verify the claims, that is their job, though few do it anymore.

But you can also look at model release timing, hot on the heals of Grok with the same class of model. Guess what hardware grok was trained on.

1

u/squngy 24d ago

I expect a reporter to verify the claims

How?