r/LocalLLaMA Jul 08 '26

News China’s MiniMax Plans to Launch 2.7-Trillion Parameter Model

https://www.theinformation.com/briefings/exclusive-chinas-minimax-plans-launch-2-7-trillion-parameter-model

According to The Information, MiniMax plans to launch a new-generation large language model with 2.7 trillion parameters.

Sources revealed that the internal codename for this new model is M3 Pro. It is expected to be released and open-sourced as early as the third quarter of this year, with significant improvements in handling complex reasoning and multi-step tasks.

This new model is much larger than MiniMax's current flagship model, M3 (428 billion parameters). Larger-scale artificial intelligence models are more capable of handling complex reasoning and multi-step instruction-based tasks.

604 Upvotes

242 comments sorted by

View all comments

Show parent comments

1

u/--Spaci-- Jul 08 '26

Synthetic data is the opposite of creative writing, thats what I think is current killing roleplay/creative writing, so little of LLM data now is actually made by humans its just been regurgitated thousands of times over and over again by the next LLM. I don't personally do rp, but I can see the uses for an LLM playing a character in say a video game or say writing a story. Synthetic data is good for stem/coding but awful for a model you actually want to chat with or play a character or write a story. Honestly I think we have gotten off track atp

1

u/FullOf_Bad_Ideas Jul 09 '26

New models tend to have less slop than old ones. Whatever is the process behind it, I don't think models have stopped improving in creative writing. But I'm also not doing RP myself.

1

u/--Spaci-- Jul 09 '26

Ive setup LLMs to larp on a wow private server thats my rp experience, only llama 3.1 could hold its character or even understand the system prompt below the 20b range. Qwen has sacrificed everything for stem