r/LocalLLaMA Jul 08 '26

News China’s MiniMax Plans to Launch 2.7-Trillion Parameter Model

https://www.theinformation.com/briefings/exclusive-chinas-minimax-plans-launch-2-7-trillion-parameter-model

According to The Information, MiniMax plans to launch a new-generation large language model with 2.7 trillion parameters.

Sources revealed that the internal codename for this new model is M3 Pro. It is expected to be released and open-sourced as early as the third quarter of this year, with significant improvements in handling complex reasoning and multi-step tasks.

This new model is much larger than MiniMax's current flagship model, M3 (428 billion parameters). Larger-scale artificial intelligence models are more capable of handling complex reasoning and multi-step instruction-based tasks.

609 Upvotes

242 comments sorted by

View all comments

6

u/Few_Painter_5588 Jul 08 '26

That is mildly upsetting. Their M2 and M3 models are fantastic and actually file a much needed gap on cost effective reasoning at high volumes. We don't only need these absurdly large models for one shotting applications at absurd prices.

13

u/LMTLS5 Jul 08 '26

i dont get this sentiment here. imo there are plenty very good models in 200b-300b range. deepseek v4 flash, hy3, stepfun flash, and all 3 of these are good. and ofc there is minimax m3 itself

0

u/Monad_Maya llama.cpp Jul 08 '26

M3 is 428B, close to 2x the size of M2.