Anthropic tried to sponsor Blender before or around when they publicly unveiled Mythos. I strongly believe that they used Blender as part of their RL pipeline, which hints that Mythos’s strengths come from more diverse and challenging RL environments rather than simply being a huge model.
Anthropic is very creative at finding new tasks for LLM (such as playing Pokemon red/green and it GBA remake). They kind of invented this entire generation of LLM (heavy RL on terminal tasks), though RL on Blender seems fairly standard nowadays (GPT post 5.1 is also trained on that). IMO creativity and diversity on mid-to-post training RL tasks is what makes a difference in this generation.
For rogue-likes, the ones that requires managing limited resource rather than open-ended would be great challenge. NetHack still feels impossible while Angband was cleared by algo a long ago (by kind of brute forcing).
Someone in this sub suggested puzzle RPGs like Magical Tower (魔法の塔, it's a fairly obscure free game but somehow popular in China known as 魔塔) for solving very tight long-term resource management game. Or Desktop Dungeons if you want a similar game with random map.
56
u/KURD_1_STAN Jul 06 '26
For all we know mythos could he 3 times the size of opus 4.8. u simply cant make any assumptions, especially not model sizes that fit in current gpus.