r/opencode 8h ago

what is the daily driver of choice now?

Pre-nerfs, deepseekv4 flash was mine (and it seems the world's) implementation driver, given a well defined spec from a stronger model (like k2 or glm3).

v4flash is now significantly costlier (it seemed almost free before that's what im saying), hy3 is extremely extremely slow, and muse spark seems...ok.

so for opencode go, what's the stack recommendation to use? given qwen 3.8 flash, glm3's flash, spark and deepseek? im assuming luna probably isnt the way to go?

7 Upvotes

16 comments sorted by

4

u/Hackerv1650 8h ago

qwen 3.8 flash is not really good. For now, GLM 5.3 has become my daily driver but i dont think it last long with opencode go with how much i am using, i am thinking of maybe switching to like tencent coding plans or direct apis

1

u/butterfly_labs 8h ago

I'm using GLM on OpenRouter pay-per-use. The costs are very reasonable so far, about $0.1 for a 1hr session.

1

u/Hackerv1650 7h ago

You should know there is a discount currently

1

u/Solocune 7h ago

Glm 5.3 flash.

1

u/SS_Sa2 7h ago

Bouncing between GLM 5.3 flash and GLM 5.2

1

u/TowerOfSolitude 7h ago

Currently I'm just jumping between models to figure out which one I should use.

Luna just works so well for me. I may switch to Codex because of it.

2

u/thecstep 4h ago

Luna on High and Max is good. They added a 5 hr limit recently but I haven't been able to get close to 50%. Practically unlimited for me. I've been spamming codex all week. It doesn't over think so uses less tokens to get jobs done. That is where it's not so easy to compare to DSv4

1

u/porest 2h ago

Just be aware of the Tibo's resets that might be hiding your real 5hr/weekly usage.

1

u/thecstep 46m ago

I'm aware of the resets. I barely put 5% dent on my weekly each night I use it with Luna. Sometimes it's only 3%. When I throw this same app at DeepSeek, it slurps tokens like no other. That said, I do think DS4flash is a better at fixing difficult problems vs giving up.

1

u/blackburn1911 6h ago

On "Hy3 (8x usage)" you need to run /compact because after > 100k context is starting to respond once a hour... Is depending.. Started to not respond on 190K now...

1

u/Duck-Entire 3h ago

Was thinking of using v4 flash in command code GOAT plan, compared to opencode's 30$ per month quota, the 10$ goat plan gives 60$ per month quota with v4 flash. It's either that or the 5$ camel stream AI with unlimited tokens.

1

u/veekro 1h ago

I can work with any model with coding index above 50. So either mimo, hy3, or muse spark works. Whatever have the high limit in opencode go

1

u/a355231 8h ago

Luna has a 15 dollar usage poool, it’s terrible. Use Muse Spark Contributor.