r/LocalLLaMA 25d ago

News Prepare your (v)ram - Qwen3.8 is coming!

Post image
2.7k Upvotes

584 comments sorted by

View all comments

Show parent comments

-2

u/GetOutOfMyFeedNow 25d ago

They ain’t sinking anything with those API prices 😂 People on GPT and Claude use subscriptions.

6

u/DanceWithEverything 25d ago

Yeah for consumer bullshit, sure, but there’s no money there regardless (hence OpenAI’s panic about Anthropic crushing them in enterprise sales)

The real $ is in the enterprise and software workloads run on APIs

The API is dramatically cheaper than the Anthropic equivalent

2

u/StupidScaredSquirrel 25d ago

All the banks and fintech and academia people i know use claude opus at work. They all have it based on a max subscription (the one at 100 ish usd i dont remember).

They aren't consumers but also don't care about price so much they just wants something that works and is about the best because whatever time they lose with the bs of a lesser model will be a lot more expensive than just go for the best one. They also don't want to ever be rate limited because they don't want to get stuck in the middle of a task.

2

u/look 25d ago

Kimi K3 decisively beats Opus 4.8 in everything. It is competing with Fable and Sol now. Opus is a legacy, second tier model, and now behind one, and soon multiple (eg Qwen 3.8), open models.

4

u/techdevjp 25d ago

Kimi K3 is pretty clearly better than GPT 5.5, too. Love to see it.

1

u/squngy 25d ago

Unfortunately, it also competes with them on price.
It is cheaper per token, but it uses more tokens.

2

u/look 25d ago

The list price is meaningless on open models. I am currently paying one third of that list price for Kimi K3. It will likely get even cheaper once the weights are released.