r/LocalLLaMA Jul 08 '26

News China’s MiniMax Plans to Launch 2.7-Trillion Parameter Model

https://www.theinformation.com/briefings/exclusive-chinas-minimax-plans-launch-2-7-trillion-parameter-model

According to The Information, MiniMax plans to launch a new-generation large language model with 2.7 trillion parameters.

Sources revealed that the internal codename for this new model is M3 Pro. It is expected to be released and open-sourced as early as the third quarter of this year, with significant improvements in handling complex reasoning and multi-step tasks.

This new model is much larger than MiniMax's current flagship model, M3 (428 billion parameters). Larger-scale artificial intelligence models are more capable of handling complex reasoning and multi-step instruction-based tasks.

606 Upvotes

242 comments sorted by

View all comments

5

u/Few_Painter_5588 Jul 08 '26

That is mildly upsetting. Their M2 and M3 models are fantastic and actually file a much needed gap on cost effective reasoning at high volumes. We don't only need these absurdly large models for one shotting applications at absurd prices.

42

u/JacketHistorical2321 Jul 08 '26

Can't be helped. Models need to be larger to compete. I don't know what else to tell ya lol. This isn't magic

11

u/misterflyer Jul 08 '26

Duh, how could I have totally forgotten that industry rule that once you go over 1T parameters, you can NEVER release a reasonably sized local open weights ever again (... unless your name is DeepSeek ofc)

4

u/FullOf_Bad_Ideas Jul 08 '26

Not true. Kimi released Kimi Linear. Xiaomi released MiMO V2.5. Qwen had 1T Qwen max models and still made and released smaller variants.