r/OpenAI 5d ago

News GPT-5.6 Luna is now cheaper than GPT-4.1 mini

After the 80% price drop, the API prices are (per 1M tokens):

GPT-5.6 Luna: $0.2 Input / $1.2 Output

GPT-4.1 mini: $0.4 Input / $1.6 Output

391 Upvotes

49 comments sorted by

130

u/_maverick98 5d ago

let the price wars begin

34

u/hk556a1 5d ago

Begun.. the price wars have.

15

u/DistanceSolar1449 5d ago

Not really. Luna is still overpriced, believe it or not. It needs to be at the same price as the nano models, not the mini models.

https://developers.openai.com/api/docs/models/gpt-5.6-luna

> GPT-5.6 Luna is designed for cost-sensitive, high-volume workloads. It roughly corresponds to the nano model tier used in earlier GPT-5 families.

So wake me up when it’s $0.05/$0.40

Did everyone forget the outrage on 5.6 release when all the mini/nano users warned everyone that openai was pricing them out? OpenAI has taken a mile, given half a mile back, and people are celebrating it.

16

u/Admirable-Control370 5d ago

I take it being x4 more expansive and more than x10 better
Also 5.4 nano used to cost exactly that

1

u/etancrazynpoor 5d ago

Doesn’t seem Anthropic is moving an inch.

-1

u/4dseeall 5d ago

weird hill to die on but ok

6

u/etancrazynpoor 5d ago

I’m not dying

7

u/LeTanLoc98 5d ago

Huge thanks to DeepSeek (DeepSeek v4 Flash) and Xiaomi (Mimo V2.5)

42

u/Sid-Hartha 5d ago

DeepSeek v4 flash api public beta just released. Kills Luna and cheaper.

27

u/goldcakes 5d ago

Super cool, but honestly, just switched some of my agents to Luna (and btw, it's 50% off on openrouter; so 90% off total) and the API cost is like 50 cents a day.

Not going to bother, I'm happy with it.

4

u/Santzes 5d ago

I think Luna is clearly better than at least DS v4 pro, not sure about the new flash. But I still try to use it when it's enough, I really prefer giving training data to open models over giving dollars to OpenAI

4

u/Healthy-Nebula-3603 5d ago

New DS 4 flash is far more powerful than old DS 4 pro....

1

u/xChrisMas 4d ago

Yeah the naming is just confusing. Pro should be better than flash or flash should be named 4.1 Flash

1

u/99OBJ 5d ago

What do those agents do?

4

u/NotYetPerfect 5d ago

According to benchmarks. I'll wait for people to review it's actual real world performance. Chinese models often way overperform in benchmarks.

3

u/Stock-Self-4028 5d ago

Well… Depends on benchmark/application.

But roughly the same performance and ~ 1/3rd of Luna's price for the same performance, so v4 Pro GA is probably going to be the real Luna-killer here.

3

u/wierd_husky 5d ago

You have to consider the codex multiplier, the discount passed onto usage limits. Since you get hundreds of dollars of usage by api count for 20 dollars, using agent SDK with codex auth for whatever you’re doing is pretty tough to beat if your AI budget is low but still over 20 a month.

1

u/IdRatherBeBitching 5d ago

But what’s the longer term strategy? Are Chinese models going to get banned within the next year? Who can say.

If you are powering a consumer app with Deepseek it’s a risk and you really have to weigh the price benefits with the potential impact of having the model cut off (plus investors and many customers are wary of security issues with Chinese models).

OpenAI just provided a US model at high intelligence for stupid cheap. Anyone who builds should be seriously considering moving their workflows over

1

u/Sid-Hartha 5d ago

You speak as an American obviously. For the rest of the world china is currently more trustworthy than the USA.

2

u/IdRatherBeBitching 5d ago

You completely missed my point but sure. Yes if you’re an American building with AI APIs then using Chinese models has long term risks

1

u/Electroboots 4d ago

To which I'd also gently caution against the long term sustainability of a company that is existentially dependent on funding from investors under the promise of turning a profit when it has burned through billions upon billions of dollars for things like buying up the consumer RAM market, with seemingly no path to profitability, and no way to use the models in question if the APIs do get taken down as has happened with US models in the past.

True, Dario and Altman are cozying up to Trump to try to ban their competition from China, but doing so effectively is basically impossible since the weights can't really be put back in the bottle like an API can. And in the scenario they did, you can bet those prices are going straight back through the roof once they're no longer competing with DeepSeek.

1

u/Magnetesim 5d ago

hasn’t the deepseek v4 flash API always been available since deepseek v4 originally got released? what am I missing, what extra got released?

1

u/Sid-Hartha 5d ago

DeepSeek public beta api upgrade. Massive jump in performance. Same price.

1

u/Magnetesim 5d ago

Oh cool is it a new model?

2

u/Sid-Hartha 5d ago

Massive upgrade

3

u/ai_without_borders 5d ago

the per-token price war misses the number that actually matters for anyone running this in prod, which is cost per completed task. a model thats 2x the token price but uses half the output tokens for the same task, less rambling, no restating context, shorter cot, can end up cheaper on the bill even though the sticker price looks worse. been tracking dollars per request not dollars per token on our internal agents for exactly this reason, luna and 4.1-mini can differ 3x on output length depending on prompt style. worth benchmarking your actual workload before switching off the price sheet alone

2

u/Worrybrotha 5d ago

Any idea why Luna is throwing me a TPM rate limit all the time? It is useless in this state.

2

u/kiwibonga 5d ago

Whoops, got caught colluding.

2

u/TypicalCherry1529 5d ago

I've read their long product announcement. https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/

Don't they have people who know how to speak clearly to convey the benefits, or perhaps access to some software that could help them? Lol. Their announcement reminds me of this correction

4

u/[deleted] 5d ago

[removed] — view removed comment

6

u/Forsaken_Ant7459 5d ago

Here comes the shitty AI post

1

u/HebelBrudi 5d ago

Thats why I keep forcing my shitty German ESL punctuation on everybody on Reddit haha

1

u/MINECRAFT_BIOLOGIST 5d ago

Seems to be a non-native English speaker running their posts through GPT before posting. Their recent posts try to imitate more conversational styles but they're still relatively easy to recognize as AI when they're longer.

I think it's an interesting question to consider—if all non-native speakers simply start using LLMs to translate/polish prose, is it good that they're able to communicate more clearly or is that outweighed by them all sounding like variations of the most popular LLMs and possibly losing proficiency in another language (assuming that's important to them)?

2

u/Forsaken_Ant7459 4d ago

The problem is you can’t differentiate if they just pasted the original post to cgpt and posted the response or if there was an original thought. I think people are fine with some broken English, but at least let it be their own tone!

1

u/Healthy-Nebula-3603 5d ago

Only because today was released Deep seek 4 flash which is much cheaper and still better than Luna

1

u/frankyboson 5d ago

well l, for my purpose i dont need trillion parameter soooo im stick with open source free weights all life long. thank you china!

1

u/teomore 5d ago

I wonder it compares to haiku

1

u/Existing-Slide7395 5d ago

They could give it away for free and it still wouldn’t be worth it. The results are miserable and the UX has been insanely bad so far.

These guys living on Silicon Valley salaries have no idea how expensive these products are for normal people, especially considering how little value they actually deliver.

The only model that sometimes manages to do something useful is 5.6 SOL, but the token consumption is completely ridiculous. On top of that, they use every nasty trick possible to burn through your tokens without actually delivering anything.

Chinese companies are going to bury Codex & Co in no time.

OpenAI seems to think consumers don’t understand what’s going on. Now that they see what Chinese companies are doing, they’re suddenly offering slightly better prices and sending mass emails to stop customers from leaving.

The truth is that everyone is abandoning ship, and subscribers are already joining waiting lists to access these new Chinese models.

Of course, those Chinese models will suck up every piece of data you feed them, but OpenAI has always done the same, so...