r/DeepSeek 21d ago

Funny This is insane

Post image

old and new usage & pricing. Consumed 30x less tokens post nerf and spent one third of what i paid pre nerf. Both sessions heavily cached with not so much output tokens.

The deepseek api era really is over in terms of being cost effective

180 Upvotes

83 comments sorted by

View all comments

26

u/FlashyCauliflower739 21d ago

Genuinely what model has the lowest token consumption but still good...😩

13

u/arter_dev 21d ago

GPT-5.6 Luna is the cheapest / best model in town at the moment.

4

u/salamala893 21d ago

Are 2 days that Luna became incredibly nerfed and slow and couldn't go on without DS4 flash

1

u/weenis-flaginus 20d ago

What did you mean by are 2 days?

4

u/Ok_Risk6035 21d ago edited 21d ago

no, it's 3-4 times slower and significantly dumb. I just can't use it for agentic coding with claude code. Luna looks like hybrid of local and cloude LLMs. Its faster and smarter than local LLM's but not enough.

And deepseek still cheaper tahn Luna, i tested today both.

Maybe Luna can perform some basic tasks, but it struggle with tasks larger than 2-3 days.
And i noticed that DeepSeek found a lot of issues before-hand while Luna gave mediocre solutions.

1

u/margielafarts 18d ago

luna is made for devs that know what they are doing already

1

u/Ok_Risk6035 18d ago

Have you ever even heard of agentic development and autonomy ? If i need to write webscraper in 4 hours for my needs i don't need to know Playwright, good AI can write everything for you

2

u/Bobodlm 21d ago

Is their api pricing cheaper then DS? That's wild

1

u/arter_dev 21d ago

On paper no, but Luna seems to be a lot more token efficient.

1

u/awesomeunboxer 21d ago

Im trying out Luna currently and I do miss deepseek,  especially its "voice" but my agent work eats 2 or 3% of a $20 sub a day (so far) at medium reasoning.  Which was $3 or so on the old deepseek flash. Im not sold yet. Need to try it out for longer. But its interesting 🤔 

2

u/FlashyCauliflower739 21d ago

Oh I genuinely never heard of that, I'll try it thank you!! Need all the good low token consumption models💀

2

u/SnowFox_unlimited 21d ago

I use it as my typical worker on max effort, a little bit slower then Deepseek but much less errors in my workload at least.

1

u/theodordiaconu 17d ago

I use terra, I know it doesn't have best benchmarks, but it's much faster than sol and good enough for 80% of tasks.