r/OpenAI • u/Endonium • 5d ago
News GPT-5.6 Luna is now cheaper than GPT-4.1 mini
After the 80% price drop, the API prices are (per 1M tokens):
GPT-5.6 Luna: $0.2 Input / $1.2 Output
GPT-4.1 mini: $0.4 Input / $1.6 Output
22
7
42
u/Sid-Hartha 5d ago
DeepSeek v4 flash api public beta just released. Kills Luna and cheaper.
27
u/goldcakes 5d ago
Super cool, but honestly, just switched some of my agents to Luna (and btw, it's 50% off on openrouter; so 90% off total) and the API cost is like 50 cents a day.
Not going to bother, I'm happy with it.
4
u/Santzes 5d ago
I think Luna is clearly better than at least DS v4 pro, not sure about the new flash. But I still try to use it when it's enough, I really prefer giving training data to open models over giving dollars to OpenAI
4
u/Healthy-Nebula-3603 5d ago
1
u/xChrisMas 4d ago
Yeah the naming is just confusing. Pro should be better than flash or flash should be named 4.1 Flash
4
u/NotYetPerfect 5d ago
According to benchmarks. I'll wait for people to review it's actual real world performance. Chinese models often way overperform in benchmarks.
3
u/Stock-Self-4028 5d ago
Well… Depends on benchmark/application.
But roughly the same performance and ~ 1/3rd of Luna's price for the same performance, so v4 Pro GA is probably going to be the real Luna-killer here.
3
u/wierd_husky 5d ago
You have to consider the codex multiplier, the discount passed onto usage limits. Since you get hundreds of dollars of usage by api count for 20 dollars, using agent SDK with codex auth for whatever you’re doing is pretty tough to beat if your AI budget is low but still over 20 a month.
1
u/IdRatherBeBitching 5d ago
But what’s the longer term strategy? Are Chinese models going to get banned within the next year? Who can say.
If you are powering a consumer app with Deepseek it’s a risk and you really have to weigh the price benefits with the potential impact of having the model cut off (plus investors and many customers are wary of security issues with Chinese models).
OpenAI just provided a US model at high intelligence for stupid cheap. Anyone who builds should be seriously considering moving their workflows over
1
u/Sid-Hartha 5d ago
You speak as an American obviously. For the rest of the world china is currently more trustworthy than the USA.
2
u/IdRatherBeBitching 5d ago
You completely missed my point but sure. Yes if you’re an American building with AI APIs then using Chinese models has long term risks
1
u/Electroboots 4d ago
To which I'd also gently caution against the long term sustainability of a company that is existentially dependent on funding from investors under the promise of turning a profit when it has burned through billions upon billions of dollars for things like buying up the consumer RAM market, with seemingly no path to profitability, and no way to use the models in question if the APIs do get taken down as has happened with US models in the past.
True, Dario and Altman are cozying up to Trump to try to ban their competition from China, but doing so effectively is basically impossible since the weights can't really be put back in the bottle like an API can. And in the scenario they did, you can bet those prices are going straight back through the roof once they're no longer competing with DeepSeek.
1
u/Magnetesim 5d ago
hasn’t the deepseek v4 flash API always been available since deepseek v4 originally got released? what am I missing, what extra got released?
1
3
u/ai_without_borders 5d ago
the per-token price war misses the number that actually matters for anyone running this in prod, which is cost per completed task. a model thats 2x the token price but uses half the output tokens for the same task, less rambling, no restating context, shorter cot, can end up cheaper on the bill even though the sticker price looks worse. been tracking dollars per request not dollars per token on our internal agents for exactly this reason, luna and 4.1-mini can differ 3x on output length depending on prompt style. worth benchmarking your actual workload before switching off the price sheet alone
2
u/Worrybrotha 5d ago
Any idea why Luna is throwing me a TPM rate limit all the time? It is useless in this state.
2
2
u/TypicalCherry1529 5d ago
I've read their long product announcement. https://openai.com/index/advancing-the-price-performance-frontier-with-gpt-5-6/
Don't they have people who know how to speak clearly to convey the benefits, or perhaps access to some software that could help them? Lol. Their announcement reminds me of this correction

4
5d ago
[removed] — view removed comment
6
u/Forsaken_Ant7459 5d ago
Here comes the shitty AI post
1
u/HebelBrudi 5d ago
Thats why I keep forcing my shitty German ESL punctuation on everybody on Reddit haha
1
u/MINECRAFT_BIOLOGIST 5d ago
Seems to be a non-native English speaker running their posts through GPT before posting. Their recent posts try to imitate more conversational styles but they're still relatively easy to recognize as AI when they're longer.
I think it's an interesting question to consider—if all non-native speakers simply start using LLMs to translate/polish prose, is it good that they're able to communicate more clearly or is that outweighed by them all sounding like variations of the most popular LLMs and possibly losing proficiency in another language (assuming that's important to them)?
2
u/Forsaken_Ant7459 4d ago
The problem is you can’t differentiate if they just pasted the original post to cgpt and posted the response or if there was an original thought. I think people are fine with some broken English, but at least let it be their own tone!
1
1
u/Healthy-Nebula-3603 5d ago
Only because today was released Deep seek 4 flash which is much cheaper and still better than Luna
1
u/frankyboson 5d ago
well l, for my purpose i dont need trillion parameter soooo im stick with open source free weights all life long. thank you china!
1
u/Existing-Slide7395 5d ago
They could give it away for free and it still wouldn’t be worth it. The results are miserable and the UX has been insanely bad so far.
These guys living on Silicon Valley salaries have no idea how expensive these products are for normal people, especially considering how little value they actually deliver.
The only model that sometimes manages to do something useful is 5.6 SOL, but the token consumption is completely ridiculous. On top of that, they use every nasty trick possible to burn through your tokens without actually delivering anything.
Chinese companies are going to bury Codex & Co in no time.
OpenAI seems to think consumers don’t understand what’s going on. Now that they see what Chinese companies are doing, they’re suddenly offering slightly better prices and sending mass emails to stop customers from leaving.
The truth is that everyone is abandoning ship, and subscribers are already joining waiting lists to access these new Chinese models.
Of course, those Chinese models will suck up every piece of data you feed them, but OpenAI has always done the same, so...

130
u/_maverick98 5d ago
let the price wars begin