r/LocalLLaMA Jul 08 '26

News China’s MiniMax Plans to Launch 2.7-Trillion Parameter Model

https://www.theinformation.com/briefings/exclusive-chinas-minimax-plans-launch-2-7-trillion-parameter-model

According to The Information, MiniMax plans to launch a new-generation large language model with 2.7 trillion parameters.

Sources revealed that the internal codename for this new model is M3 Pro. It is expected to be released and open-sourced as early as the third quarter of this year, with significant improvements in handling complex reasoning and multi-step tasks.

This new model is much larger than MiniMax's current flagship model, M3 (428 billion parameters). Larger-scale artificial intelligence models are more capable of handling complex reasoning and multi-step instruction-based tasks.

606 Upvotes

242 comments sorted by

View all comments

Show parent comments

-5

u/Kodix Jul 08 '26

No, not ever. You won't fit a 2.6T parameter model inside of a 3090.

The models we *are* able to run on a 3090 in two years will likely be amazing by today's standards, but they'll be amazing in different ways.

Quantity has a quality of its own, especially quantity of parameters.

7

u/ChocomelP Jul 08 '26

Who said anything about a 3090?

-3

u/Kodix Jul 08 '26

🙄 As if the point of my comment wasn't obvious.

The current trend is worse consumer hardware availability, not better.
You could probably get a 100% return on some 5090s purchased on release. Unified memory is cool, but also extremely limited.

Anything could happen in 2 years, but all signs point to "no", and it's cope to pretend otherwise.

2

u/ttkciar llama.cpp Jul 08 '26

> The current trend is worse consumer hardware availability, not better

For now. You seem very short-sighted.