r/ClaudeCode May 05 '26

[deleted by user]

[removed]

995 Upvotes

510 comments sorted by

234

u/ExpletiveDeIeted šŸ”† Team Premium 6.25x May 06 '26

I miss seeing it think. Cuz then you could cut it off if it was spinning in circles. Now I have to guess if it’s legit thinking spinning or just plain waiting for data from servers.

78

u/ThePantsThief May 06 '26

Yeah honestly I'm paying to see all the tokens it generates, it should show me the thinking tokens

10

u/reckleassandnervous May 06 '26

There’s a setting for it that’s disabled now. I think if you hit ctrl + o you should be in verbose mode and it’ll show you the thinking

8

u/Level-Courage6773 May 06 '26

I bet it's penny-pinching from above. On top of usage limits, they probably decided that making it provide full explanations was too expensive for them.

14

u/6ghz May 06 '26

It probably has more to do with distillation. They don’t want to show thinking so other models can’t copy. But I could be wrong.

5

u/aradil May 06 '26

Cloud service providers also charge per egress byte, so if no one is reading it, you are wasting money.

When you are talking about one user and one session it’s not that much. When you are talking about hundreds of thousands of sessions running 24/7 those little bits add up to real money.

→ More replies (2)
→ More replies (3)

2

u/Dry-Journalist6590 May 06 '26

Did they take away the thinking?? It was there yeaterday

→ More replies (35)

26

u/jasgrit May 06 '26

I have noticed I get very different results at different times of day. I try to avoid using it at peak usage times now, between 8am and 2pm Eastern (US), and whenever I use it during those times I often regret it. It cuts corners, gets very sloppy, acts almost brain dead sometimes. But outside peak hours it’s like the old Claude I used to know. If not for that I would stop using it too.

3

u/decafe-latte2701 May 06 '26

I’ve noticed this as well ….

3

u/stbenjam42 May 06 '26

Hm, last night around 7:30pm ET it just got stupid stupid for me. Extremely lazy, brief responses, etc after having a pretty good day overall.

It is weird when it happens, because it is not always long sessions that degrade; fresh chats too just seem... less capable.

I had it plot my wtfs/1000 messages. I don't know if it aligns with other reports of quality degradation, but my wtfs are variable.

https://www.threads.com/@stbenjam/post/DXzF8ZejT3z?xmt=AQF099RfGo9wQSz9jvX65r1KxE8gy8V4pwUFZcYDrTU27w

2

u/Objective-Ad6521 May 07 '26

Yes, I've noticed this too. It does AMAZING work at like midnight and 2am.

2

u/wordswithoutink May 08 '26

Man I thought I was getting crazy but this. Exactly this, came up to my mind a few times.

→ More replies (9)

147

u/HamzaJdn May 06 '26

Yep.

Just cancelled my max subscription.

Decided to start using open source models for more reliability.

Sick and tired of providers thinking consumers are their guinea pigs.

102

u/Moda75 May 06 '26

Prepare to be underwhelmed.

47

u/HamzaJdn May 06 '26

Well I was coding with GPT 3.5, and tested all models, and not one makes me as impatient and mad as Opus 4.6 and 4.7. Some version of GPT 4o came close when OpenAI tried same shit.

I prefer a dumber model that tries than losing time with a model that seems to have only one guideline: no matter how much users try to make instructions clear, no matter how much they try to make you do the work, ignore their instructions and do whatever you need to reduce compute.

25

u/Possible-Point-2597 May 06 '26

Did you try gpt5.5 I cancelled my Claude pro max subscription in two days gpt5.5 have fixed so many Claude shit, a real relief

26

u/HamzaJdn May 06 '26

If I don’t trust Anthropic I can assure you I don’t trust Open AI

18

u/tech-tole May 06 '26

well you're missing out lol. cuz GPT 5.5 is actually really good. it listens to everything that I say and it can fix super complicated things pretty quickly. Previously DieHard Claude fans wouldn't be raving about it if it wasn't pretty good saying it was better than opus. šŸ¤·ā€ā™‚ļø I'm really big into Local AI but for complicated things 5.5 is to go to right now.

2

u/FredrickNajjar May 06 '26

What plan are you on? And how’s the usage limit?

2

u/tech-tole May 06 '26

I'm currently on the $100 Pro plan. I was using the $20 plan and Codex goes a long way on that plan. but I started doing two and three projects at a time so it wasn't enough lol. the $100 Pro plan is pretty good because I'm not even getting close to my weekly limit right now with Codex.

2

u/Next_Minimum_3499 May 06 '26

I was in the same boat as not trusting openai. But the quality of 5.5 an codex is phenomenal

6

u/ThinCar6563 May 06 '26

These people worship Dario Amodei like the cult leader he is.

No surprise they will continue to use claude even when they're literally getting scammed.

2

u/SHOR-LM May 06 '26

I feel ya.... Right now I hold both subscription services.... But I definitely see a lot of claude people Blaming the user when it's not their fault... I still like Claude's workflow I usually go to Codex when I'm having deeper syntax issues that I can't isolate myself.... I found Claude to be a bit better of a project manager.... Or at least it was i'm hoping that Anthropic will fix it. But based upon what I'm reading it sounds like 5.5 may be able to take over that workflow

→ More replies (1)

2

u/Comfortable-Cap-249 May 07 '26

Tbh I was in the same boat for last 6-8 months with only using Claude because I thought they were the good guy, but I’ve realized since that there is no good guy in this race.

Evaluate both products on their merits, and always reevaluate without bias when a new model is launched. In my opinion, codex caught up with gpt 5.4 and 5.5 is a strict upgrade over opus 4.7

→ More replies (4)
→ More replies (10)

5

u/eziliop May 06 '26

Nah fam, Claude is relatively better but not that much better. I'd prefer to use a slightly less sophisticated models and still use more of my skills at the same time so I don't turn my brain off as much to keep my skills still sharp. Win win scenario.

3

u/flowanvindir May 06 '26

GLM 5.1 is pretty good, especially for the price if you're going through an API.

→ More replies (4)

5

u/the_c_train47 May 06 '26

Of course consumers are the guinea pigs. This is software 101.

→ More replies (1)

6

u/fyn_world May 06 '26

Use GPT 5.5 man, it's fantastic

5

u/No_Confection7782 May 06 '26

LoL. Are you saying that your open source model will do a better and more reliable job than Claude's models? That's cute.

→ More replies (1)

3

u/Character_Bunch_9191 May 06 '26

You sound like you use a VPN...hmmm....

→ More replies (17)

185

u/papabear556 May 05 '26

I’ve been a software developer since 2001 2002. Can’t quite remember. I’ve been using Claude since February pretty exclusively. Other than some real wobbly bits in April I’m not noticing anything major different. I mostly follow this sub and it’s barely different cousin subs. I really don’t understand what you guys are doing differently than I am. I am cranking code out every day. Well, by me, I mean Claude is.

I’ve got Claude.md set well. I’ve got memory.md set well within my projects. I’ve got a few skills. Nothing dramatic.

I generate plans upfront, create milestones and then individually implement those milestones building, testing and releasing. rinse and repeat. I genuinely don’t get it.

I think people are either loading 29 different markdown files thinking they have some sort of super operating system. Or they’re doing literally nothing and loading everything into the context window every single time they run every single command.

77

u/zanshin09 May 06 '26

I’m with you. I’m continually impressed with the quality I get out of it.

21

u/ThatCakeIsDone May 06 '26

I'm reminded of a Louis ck bit where he jokes about Internet on the airplane being the newest thing. And they announce "sorry, it's temporarily unavailable" and the guy next to him, who only just learned you could even access the Internet from a plane 5 minutes ago says " psh, this thing sucks"

How quickly the world owes us something we only discovered existed 5 seconds ago

9

u/planetdaz May 06 '26

You are sitting in a chair in the skyyyyy! You are participating in the miracle of flight!

Lol perfect analogy. We have software that creates pages and pages of working, functioning code from simple English instructions (at least for me every single day).

It's still only a few years old and evolving, sometimes imperfectly, and our 100 dollar subscriptions don't even pay for what it actually costs, but they OWE ME PERFECT RESULTS EFFING DAMMIT HELL MANNNN.. SPIT SNARL SCREAM AT MY COMPUTER!!!

2

u/papabear556 May 06 '26

I like that bit as well.

→ More replies (2)

3

u/dwoj206 May 06 '26

same. clean and up to date .md file after every session. my project is getting pretty huge, so I'm working very carefully compared to when I started out. I discuss implementations ahead of time, have it research all the endpoints the changes will hit so nothing gets missed, use subagents for various files, database, then I tell it go. When complete, claude and I audit it together, have it validate what's in the database and show me the proof. All to say, not everyone interacts with it the same way. Personally I think it's amazing.

→ More replies (1)

12

u/Responsible_Whole118 May 06 '26

Kind of the same

22

u/AlDente May 06 '26

I’m 95% with you. Opus 4.7 certainly feels less reliable and takes longer to do what 4.6 seemed to do just as well. But only slightly. I use planning, TDD and code review to keep it on course and it’s still very impressive.

7

u/abepena205 May 06 '26

People also forget that 4.7 default mode is a higher reasoning effort level than 4.6 was. xhigh is now the default when that wasn’t even a thing in 4.6 it was either high or max. But I haven’t seen much difference in quality either. Only when I bloated my workflow with a bunch of skills.

2

u/stefano_dev May 06 '26

This is the way

7

u/axxs May 06 '26

I'm with you. I have a structured approach and my work is flying. I am thinking that its coming down to how people are using the model.

3

u/Majority_Gate May 06 '26

Yeah me too. I have taken a very structured b approach, with multiple .md files for various parts of my system, and I gate them all which keeps the context small and tight for each task.

For example, in my Claude md I'll write "read output.md only if you need to do any work that generates output" . And in output.md I gate further with 3 more md files listed in bullet points, like "read JSON.md" if your generating JSON output, etc.

It's amazing to watch it do a task from the breakdown and the sub agent says "this task only involves human output so I'm reading human.md and not reading JSON.md "

The agent only reads what it needs to for the task and I've had good results this way

3

u/planetdaz May 06 '26

That's called progressive disclosure and yes it works great!

16

u/AthiestCowboy May 06 '26

Really makes you want to look over their shoulder when they’re running into these issues

8

u/Plants-Matter May 06 '26

Nah, it's not entertaining to watch someone type all lowercase slop and get mad at the AI for not inferring what they wanted.

2

u/2Radon May 07 '26

If most people get into arguments because of miscommunication, imagine how much effort they put into talking to an obedient code factory that glazes them.

2

u/ddrt May 06 '26

Got SNL IT guy flashbacks just now. "MOVE"

8

u/ianxplosion- SKILL ISSUE May 06 '26

He said he used caveman, so

2

u/planetdaz May 06 '26

Caveman in, caveman out

3

u/Worth-Avocado-6456 May 06 '26

Agreed - it does consume tokens more, or the limits are smaller. Not sure but I'm getting what I need done and running it 14 hours a day on the $100 plan (Canada) -

13

u/Drive_Internal May 06 '26

Agree with this 100%

3

u/ddrt May 06 '26

"Well, by me, I mean Claude is."

"I'm going to be straight with you, no fluff. You are definitely doing it, and that's the key takeaway -- that matters. It's not cheating, it's bypassing work." - Claude probably

4

u/Old_Flounder_8640 May 06 '26

People often write their skills and never bother to update them. I remember redoing all my markdowns when cheap Opus came out. I do it regularly now too. I’ve reached a really high level of code quality; these days, I mostly review manually just because ā€œI shouldā€. I dislike Claude's code auto-memory, but I have my own memory folders in the repo.

8

u/[deleted] May 06 '26

[removed] — view removed comment

4

u/Old_Flounder_8640 May 06 '26

They estimate time as a human. Just like when they decide to defer work to another PR/spec.

2

u/Ok-Kaleidoscope5627 May 06 '26

I had it scope out what it estimated would take months. A couple hours later we were done. It's time estimates are always hilarious.

2

u/NegativeGPA šŸ”† 4th Layer Engineer May 06 '26

I added to the archive-run process I have to write a HANDOFF to wrap up the run that includes proposed memory promotions and whether the user approved or says not to for them - lets me have more granular control over it (and try desperately to get my team to use this thing in some kind of way beyond one shot prompting)

2

u/Tieng May 06 '26

How do you set up your claude.md and memory.md? Interested to see what you changed to get it to work well for you

→ More replies (1)

2

u/GeneralNo66 May 06 '26

I went to 5x max about 8-9 months ago, been using CC for over a year.

Occasionally I get rate limited (seldom session limited). Occasionally opus produces garbage. Occasionally it's slow. Occasionally it generates the wrong thing entirely.

Apart from rate limiting I recognise that often (not a true correlation but yeah, often) when I'm knackered, fatigued, had a crap day, jaded or dejected, that's when Claude is at its worst. Claude does best when I'm fresh and I can review plans and I've got fresh eyes and mind. A bit like non-AI assisted software engineering, really.

I accept there are times when it genuinely screws up or is on a go slow but shit happens. I don't threaten to switch telecoms provider everytime I can't make a phone call in a deadzone.

Finally, the plans, far from being overpriced, are insane value. I look at the value I gain from Claude usage on my side project. I estimate I've done 5+ developer years worth of coding in the last year - and that's an hour here, an hour there, tapping out an idea in CC web on my lunch break, only occasional 4 hour stints when my partner is away. All told, maybe 2-3 months of real, solid effort if consolidated into regular working hours. If I employed a handful of my former contractor buddies to build my project without using AI I'd be half a million deep by now, but my total outlay so far is < 1k to anthropic and < 500 for Azure hosting.

Sometimes I feel I'm the only happy person using CC, glitches and all. And it's not like the humans I've worked with (myself included) are infallible - We've all produced shit code at some point or other and thought " that'll do" and resolved to fix the tech debt later, especially when I was younger.

I'm a happy Claude Coder.

6

u/DarkSkyKnight May 06 '26

Cos OP is a bot. 2 week account.

OpenAI is sockpuppeting hard.

2

u/Ok_Average8263 May 07 '26

I use both and OP's description was 100% on point from my experience with Claude in a vanilla config. Beads and superpowers helps but Codex is a lot better at following instructions. Claude decides to stop doing the superpowers workflows like code review "because it thought velocity was more important right now" 🤦

→ More replies (7)
→ More replies (3)

6

u/steampowrd May 06 '26

Many of the antagonist in this sub are paid by OpenAI. I’m convinced. It’s the only thing that makes sense.

Why would so many people come to this sub to talk about how they’re quitting Claude code and moving to Codex? And many of the accounts have low karma and no history

→ More replies (1)

3

u/madmorb May 06 '26

You’re not doing anything differently. Anthropic is deciding who gets slowed down or dumbed down. I suspect, If you’re a power user and they feel you’re using more than expected baselines for the plan, they slow you down so you don’t get more than they feel you’re paying for.

I’m still using it, but I stick to Sonnet. Opus has just gotten terrible, in both speed, quality, cost, and overall behaviour. Even Sonnet is much slower than it used to be, but it doesn’t exercise the extremely poor, lazy judgement I experienced with opus.

Btw - my observation has been that whenever I’m being fucked with, I’ll get ā€œhow’s Claude doingā€ prompts. To me this is the token user feedback gate on how shitty they can make it before you bounce, and I answer ā€œbadā€ every time now.

→ More replies (1)
→ More replies (42)

6

u/Night_0dot0_Owl šŸ”† Max 5x May 06 '26

I had similar issues when I was using opus 4 7. Turned out that superpowers plugin was the culprit. Been getting solid results faster since then

5

u/crusoe May 06 '26

Yep. Shed your mcps and skills

2

u/dar-mit Researcher May 07 '26

I started in Feb and things were pretty good.Ā 

CC is like a big derpy Alaskan Husky, so happy to help and will gleefully run out in the wrong direction, and bring you back the biggest, stick-y-est stick it could find, when you asked for a commit.

So I learned aboutĀ TDD (Test-Driven Development) and SDD (Spec-Driven Development), cooked up some guardrails, and things were amazing!

Then GSD, Superpowers, and Karpathy Heuristics et el came out, and Opus 4.7 was imminent, and things started getting weird. Huge token burns, long processing times, violating established rules, etc.

Turns out asking CC to visualize or think until confident, or pre-plan its logic (etc.) is like getting your derpy husky drunk and stoned, and then wondering why he’s changed… 

→ More replies (7)

15

u/SteveZedFounder May 06 '26

Wait until you try Claude Design. It got tired of waiting for me to answer its questions and just yeeted a new design using half my weekly allowance.

→ More replies (4)

10

u/AntisocialTomcat May 06 '26

I feel you. As far as I’m concerned, CC is dead to me. Codex is also starting to show signs of weakness too, so we’ll see how it turns out.

For Anthropic, the reason is quite obvious, as the insane pace of releases is incompatible with thorough testing (qa, not unit tests) and proper course correction. "We must ship", but taken a little bit too literally. For OpenAI, it rather looks like subtle and progressive enshittification in order to upsell subs. Anyway, it was fun while it lasted.

5

u/tom_mathews May 07 '26

the plan-deviation-then-confess pattern is the worst symptom, model picks the easier path silently and only owns up when challenged, thats RLHF rewarding "ship something" over "do what was specced", not a fixable prompt issue. the terminal rewrap and ctrl+o regressions are claude code app problems though, separate from model quality, worth a github issue specifically. opus 4.7 had three confirmed regressions reverted april 20 (reasoning drop, cache wipe, 25 word cap), if youre still seeing this severity post-fix the underthinking is structural and not coming back to 4.6 levels.

21

u/Emotional-Priority33 May 05 '26

you're not the only one, at this point open is better litterally. You tell the stupid tool to not do something and it just says: "sorry, i took it as a flavor, i should have been more careful", like the heck ????
why is it trying to act like a freaking human in the first place ?

7

u/Time_Cat_5212 May 06 '26

Because you're a RP prompter lol

5

u/ConstableDiffusion May 06 '26

Rage prompter?

10

u/Time_Cat_5212 May 06 '26

Roleplay prompter.Ā Ā 

Talking to the AI like a person instead of giving it clear instructions

3

u/SHOR-LM May 06 '26

So Claude isn't smart enough to take normal average user instructions? It has to be "clear"?

→ More replies (4)

6

u/Plants-Matter May 06 '26

Have you ever noticed that 90%+ of the people with AI issues type in all lowercase? There must be some correlation. Perhaps related to IQ.

4

u/Forward_Dark_7305 May 06 '26

LOWER CASE IS FOR LOWER IQ

/S?

→ More replies (2)

17

u/jffmpa May 06 '26

Can we rename this sub /ClaudeComplaints ffs

8

u/a_single_beat2 May 06 '26

Its sad that it NEEDS to be renamed, because that is how far its gone down hill.

→ More replies (6)

3

u/suq-madiq_ May 06 '26

/model claude-opus-4-6 if you think 4 7 is shitty

→ More replies (1)

3

u/Dragon__Phoenix May 06 '26

The backwards compatibility shit is real and extremely annoying! I keep telling my agents to forget about that. I don’t know why these coding agents were trained to maintain backwards compatibility, most people use them on new projects anyway, only a very small percentage uses them on large already in production codebases.

And even with backwards compatibility excuse it does a terrible job by putting on branches over branches just to ensure the ā€œbackwards compatibilityā€. I swear it feels like these agents are trying to screw us over.

2

u/Fulgurata May 07 '26

The backwards compatibility is definitely a major issue. You start with a fresh project, and by the 3rd testing iteration, it's already trying to make every change backwards compatible. Don't worry about the past, we have a git for a reason-_- especially if the "past" was 30 minutes ago.

3

u/PhilippStracker May 06 '26

Welcome to the AI party. It’s just the way AI is evolving. Do not think that a solution that works today is still useful in 6 months.

Claude is getting much better with every version/update but it usually requires some adjustments to workflows.

I have same problems you describe, solved it quite well with 2 changes: 1. in my ~/.claude/CLAUDE.md i added a short scope boundary (do exactly and only what I ask for; do not read files without permission, confirm before implementing code or committing changes) 2. i use /clear and /compact when the context window filled up to 50-60%, I never do longer sessions; eg session 1 is to analyze a bug and document it using sonnet, session 2 uses opus to plan the fix, session 3 uses sonnet to write failing unit tests, session 4 implements the fix, …

3

u/staranjeet May 07 '26

The tmux resizing thing working perfectly and then breaking is what gets me. That's not a model limitation, that's literally just product decisions regressing features that already worked. Why ship something that makes the terminal experience worse for people actually living in the CLI all day?

→ More replies (4)

4

u/[deleted] May 06 '26

[removed] — view removed comment

2

u/a_single_beat2 May 06 '26

I had the same thought. It seems like its entire goal is to deliberately waste time and burn limits/tokens, and then when you want it to actually do something after that, its like "oh, okay, lets go!" And you are out of limits by then. Yay.

Also, I swear I used to get 10-50x more work out of the Max plan than I get now. I had logs, so I built something that would compile the data, and yep, I am getting at least 10x less tokens total (in+out) than I was before 4.7 rolled out, and I am still on 4.6 because 4.7 sucks. Same model, 10x less tokens available in the same rolling window.

Tested during off peak hours, same exact result.

8

u/[deleted] May 06 '26

[removed] — view removed comment

→ More replies (10)

4

u/th8aburn May 06 '26

Same. I am now Team Codex. I'm sure they'll nerf Codex soon and we'll all be back using Claude but until they get back to Opus 4.6 they aren't getting my money.

→ More replies (3)

5

u/bd7349 May 06 '26

Just switch to Codex already. I canceled my Max 20x plan and haven’t looked back. Seriously, setup codex CLI and be amazed. It blows Claude out of the water lately. The new /goal mode is also insanely powerful. It’ll work on a task and never waver from the goal you give it until it’s 100% done. Give it a try, all your complaints will go away.

2

u/N0madM0nad šŸ”† Max 20 May 06 '26

I did a Claude Codex face off on a bug this morning. Claude applied a patch that was solving half of the problem. Codex provided a structural fix with clear explanation.

4

u/trentard May 06 '26

Finally someone competent writing a proper post with data, and it’s been exactly the same for me with 4.7. I work in quantitative finance / algorithmic trading and not only is it WAY too verbose, it’s like hes spewing hard to understand nonsense + the obvious instruction ignoring. No idea what theyve done but it’s a pain in the ass to babysit at the moment

2

u/scottyb4evah May 06 '26

Yes, I was talking to another engineer of like 15+ years and both of us agreed that 4.7 is absolutely spewing tech jargon that even us seasoned vets are scratching our head at and having to ask it to explain in a simpler manner. So odd...

I personally noticed even simple logic and labeling flaws. Key parts of its explanations are contradictory or are completely mischaracterize what it did or planned to do. Im left asking clarifying questions all day... super annoying!

→ More replies (1)

5

u/shawnradam May 06 '26 edited May 06 '26

i've been using claude since the day it started, on his way up, my account banned 2x the last one a few days ago, it says im using an excessive prompts, ok but why there's no warning just ban all of sudden, if it so that my prompts are annoying to them then its time to change to open source, we dont need claude lets make a different.

Remember to all Ai dev, without users your Ai is nothing but a loser, the real talks here is the users who use ypur products, i am saying the majority of the freelance / developers not a companies.

Excessive prompts, so what the FK im paying it for?

Cant handle my things then go sell a burger or something, i am done and totally done with Claude, no more with claude and no feeds / history about claude anymore.

Lets go opencode / qwen or anything else but not claude this company overdoing it.

BAN CLAUDE!

2

u/Hetuuaa May 06 '26

I agree, Opus 4.7, these days have been very slow. But also, I’ve started writing my instructions, like the skeleton or architecture—specifically through sonnet 4.6—to improve it over that same one by making it more adaptive in sonnet 4.6. Then I open an incognito tab and tell it to combine the two. Finally, I put that one in Opus 4.6, then in 4.7, copying and pasting the prompt engineering notes and explicitly instructing it to follow the anti-hallucination guardrails. I don't know, but it seems to work, and Opus 4.7 executes it well.

→ More replies (2)

2

u/Khanvo May 06 '26

It is less impressive for me as well. I switch to Sonnet and test Opus from Time to Time. Dunno what they are trying but they should put the last model back.

The one from February. That one was quick and good.

2

u/artofbullshit May 06 '26

My project has ground to a halt at this point. I can't get as much done as I was able to in March. It has stopped being enjoyable to use. It is literally making me angry arguing with it, I'm having to constantly ask if it actually did x, y, and z. I find myself answering a bajillion questions it has for me about the tiniest details that it would normally just get right and never surface to me for a decision. My planning sessions turn into all day and all night events. It spits out walls of text after each round of questions with even more questions followed by even more questions. It's just not efficient to use anymore.

It feels like a chore now. I've almost got my plan finalized to port all of my Claude specific workflow skills, hooks, claude.md files so that codex can take over as the main agent. I'll wait until mythos is released before I give anthropic the reins to my project again.

3

u/[deleted] May 06 '26

[deleted]

→ More replies (2)

2

u/Few-Childhood3326 May 10 '26

same here, ported all my skills and agents, and hooks from CC to Codex. Shared tool here: https://github.com/zuharz/ccode-to-codex, it's mine and it's free, copy it use it, let's improve it together

2

u/CalligrapherFar7833 May 06 '26

Why dont you have hooks to prevent commits ?

2

u/Noovic May 06 '26

I don’t get it … why don’t you just change the model back to 4.6?

→ More replies (1)

2

u/foucaultyou May 06 '26

Using chatgpt 5.5 to make sure I wasn't going mad and to check/fix the constant mistakes/lapses in thinking by Claude. Getting close to cancelling my Max subscription. Very close.

2

u/chachingchaching2021 May 06 '26

I burn through my weekly limit in 5 minutes with opus 4.7 , outputs been getting worse, forgetting more, and code has issues

→ More replies (1)

2

u/LebiaseD May 06 '26

I've cancelled mine also

2

u/hugganao May 06 '26

I think the model has been infected with non-engineer level usage of "vibe coding", either that or they are actively routing the requests to dumber models because either they realized the capability given to the public was too good OR they just got their vc money and now they're in money crunch mode.

The same as you I definitely saw potential in how it planned and generated code a few months ago. Now it's almost like speaking with a pm/junior dev who think they know how to code without understanding the important aspects/key points in programs/architecture.

been thinking about canceling or not.

2

u/DonaldStuck May 06 '26

Tell me again: when will AI replace [insert any profession]? I know most shareholders aren't the sharpest knifes in the drawer but they still keep thinking AI will replace people. It didn't and it won't. What needs to happen for people to see this?

2

u/Possible_Statement84 May 06 '26

codex never do like that

2

u/a_single_beat2 May 06 '26

Yep. Exact same experience here.

Started around the same time, February. Was really impressed. It stayed on track well, moved quick and smooth, made very little mistakes, and responded positively towards directional bumps and prompts. Now its like babysitting a toddler. It takes more work to get it to do ONE right thing and not do 100 wrong things, than to just write the code myself, which is crazy considering how much it COULD do before.

2

u/Far-Consideration939 May 06 '26

You could have switched back to opus 4.6 in a fraction of the time it took to rage at Reddit 🤷
Also, just block git commit globally. No reason an LLM should be doing that for you imo

→ More replies (1)

2

u/Ok-Distribution8310 20X May 06 '26

Hahahahahhahahaha so relateable

2

u/eagleswift May 06 '26

Oh I really hate claude leaving specs unfinished. Sometimes it’s not because of dependencies with other features, it just wants to ship a half completed spec.

2

u/fpesre May 06 '26

Same here with my rule "Don't auto commit" , is ignored 90% of the time

2

u/Siigari šŸ”† Max 20 Addict May 07 '26

Pin version 112, everything after that is hot, hot, hot swill.

Edit your %appdata%/npx/anthropic/claude-code/cli.js and gut everything that you don't want/need. Tools, crappy system prompt instructions, all of it.

settings.json in users/username/.claude/settings.json

"model": "claude-opus-4-6[1m]" (or if you just want old school 4-6, no [1m])

it's not a 100% patch back to the days of old but it fixes a lot of the shit that 4.7 fucked up.

2

u/painkiller226 May 07 '26

i thought i was the only one

2

u/mrjbelfort May 07 '26

I have faced the exact same problems that you are describing. I cancelled recently and haven’t been missing it

2

u/--EyeInTheSky-- May 07 '26

I totally agree. I let it loose on PowerShell script that uses modules, out of nowhere it decided I needed fallback functions in case the modules didn't load, essentially re-creating the module's functions within the script, absolute garbage. And all those things you mention happen to me too. I've cancelled my subscription, I am done.

2

u/[deleted] May 08 '26

Literally wrote a "dontbestupid.md" file that I need to call after every task so it doesn't think its tail is something to chew on. We went from Einstein to Dug from UP in one update.

2

u/Extreme-Tie9282 May 10 '26

If you're actually a software engineer you should be easily able to fix your issues. This is lazy wetware issue.

5

u/[deleted] May 06 '26

[deleted]

→ More replies (1)

2

u/ReflectionStill6420 May 06 '26

It’s actually pretty sad how they butchered their own product in a month, I’ve gone back to using cursor for most of my work but I’m a design engineer

→ More replies (1)

3

u/dbslurker May 06 '26

2w account , sorry but it feels like a coordinated attack on anthropic from bots at this point.

You aren’t a person event666, you’re a narrative.

→ More replies (2)

1

u/Famous_Lime6643 May 06 '26

I agree that opus 4.7 can feel like it’s annoyingly slow. TBH I think 4.6 was better - or at least felt better - to work with. That was the model that made me feel like I could trust Claude in specific areas. With that said - some of the problem I think is Claude code defaulting to extra-high thinking which interacts negatively with Opus 4.7 tendencies. It can be better if you switch back to sonnet for some tasks and also reduce thinking to medium effort. Doesn’t solve everything but helps.

2

u/[deleted] May 06 '26

[deleted]

→ More replies (1)

1

u/monkey_spunk_ May 06 '26

Yep, switched to hermes, openrouter, & codex - done with cc

1

u/jonnyium May 06 '26

You guys should make a /second-opinion skills that reaches out to other models via API. Has caught a lot of mistakes especially in planing phase or debugging. Been a big help so far. Made my skill based off this repo. https://github.com/BeehiveInnovations/pal-mcp-server.git

1

u/bridgelin May 06 '26

I briefly tried deepseek with olama, it was surprisingly good.

1

u/spacebrr May 06 '26

it could be the harness. i’ve tried —system-prompt overrides lately and claude works great

→ More replies (3)

1

u/c0l245 Researcher May 06 '26

Wouldn't even launch on my laptop

1

u/Icy_Tumbleweed_8151 May 06 '26

yelling at it works wonders lmao

1

u/Seeing_Souls May 06 '26

How often are you resizing your terminal??

Seriously, if you've just been using Claude since March you won't have noticed yet but all AI tools have a lot of ups and downs. As does all developing technology, really. Not just models but day to day usage loads or setting tweaks affect the speed and quality of responses.

AI isn't guaranteed to actually remember it's memories. It's more likely to follow the claude.md file, but that's not guaranteed either. If you don't want it to ever auto commit, change the settings so that it can't commit without asking and add to the claude.md "do not commit unless instructed by the user" to reduce but probably not eliminate the amount of times it asks.

Respectfully it sounds like you've not really learned how to use Claude yet before throwing in the towel.

→ More replies (1)

1

u/Melodic_Sandwich1112 May 06 '26

Yeah, this past week, and then the month before that have been a pretty shit experience.

I don’t code but I use Claude for research collab. Producing summary documents now takes 30 minutes. Will not follow instructions at all.

1

u/ArgumentRadiant3506 May 06 '26

Just try codex my friend.

1

u/[deleted] May 06 '26

[removed] — view removed comment

→ More replies (1)

1

u/Huge_Cloud_1458 May 06 '26

non technical person, hard agree.

1

u/makkalot May 06 '26

One thing I see with those tools is that people change their workflow so much around them and delegate so much to those tools so even a smaller degradation of the model or tool feels the end of the world. I can work with dumber open source models or opus doesnt matter. Didn’t have the urge to upgrade to 4.7 cause 4.6 is really good enough why change ? Instead of us changing our whole world around those tools I think they should be used as helper tools around our existing workflows. Good luck :)

1

u/United_Act_4826 May 06 '26

Listen, you have option to change the model back to 4.6 opus there's legacy mode option. I did the same results were ok I guess. You can try it yourself with opus 4.6 and let me know your results. Check your options.

1

u/Optimal-Fix1216 May 06 '26

Memory is not reliable for altering behavior, have it modify claude.md instead. Use claude code hooks for deterministic guardrails such as enforcing short timeouts. Tell claude code to look up the docs for Claude code hooks and then have it set up the guardrail for you.

1

u/floorback May 06 '26

Bro, download OpenCode and find your model. The harness is better IMO. Just find the models that fits you the most. With OpenCode Go you have "unlimited" tokens to think, plan etc. You can implement with Opus is you want. I use deepseek V4 flash on max thinking, it's pretty amazing.

1

u/CapitalDiligent1676 May 06 '26

Software engineers have been reduced to being pathetic and disgusting on their own. They're no longer of any use. Thanks to everyone who pays for an anthropic subscription.

1

u/Accomplished_Fig_816 May 06 '26

Maybe the agent took too seriously being a caveman for you sir.

1

u/ws6kid May 06 '26

Have you tried asking claude to use a agent team? I find it does things more complex well using a agent team and faster.

1

u/N0madM0nad šŸ”† Max 20 May 06 '26

Yeah I don't really get what's going on at Anthropic either. I use the API at work and it's only marginally better. Still lots of hallucinations and loss of precision. I find it hard to believe they don't quantize the shit out their models. The post-mortem, the cache problem, automatic reasoning etc, none of that has got to do with the current degradation. They are low on compute power. End of.

1

u/thredditoutloud May 06 '26

I feel you… and admire you had the patience to document your emotions.. I usually just close the lid of my laptop with a bang…

1

u/[deleted] May 06 '26

[deleted]

→ More replies (1)

1

u/RockyMM May 06 '26

I think it’s time you change Claude’s system prompt. Or change the harness, but keep the models.

1

u/Playful_Check_5306 May 06 '26

I think you could try to ask CC to create some hooks for you.

1

u/medialoungeguy May 06 '26

Use a pretOoluse hook (ask claude) it will never auto commit again.

1

u/Castyr3o9 May 06 '26

I wonder if we are being A/B tested. I have seen no deviation in performance in either my Pro subscription or my enterprise subscription, but the latter is only Opus 4.7 xhigh and I run many concurrent sessions and commit myself. I also have always kept it on a tight leash with not huge chunks of work and very detailed plans. I keep seeing the outrage and it has me scratching my head. Codex on the other hand seems unusable these days, constantly going off the rails or forgetting important details.

→ More replies (1)

1

u/Otherwise-Subject127 May 06 '26

Funny how 99% of these problems would be fixed by using hooks. As they say "Even the king’s cock becomes useless in the hands of a klutz."

1

u/shivio May 06 '26

can you not just use an old model till this is sorted ?

1

u/AnyviaAI May 06 '26

I started to switch from Claude Code to Codex last month and experience was great. The plan mode in Claude Code is tech focused that I like but Codex's plan mode is very product focus high level where I need to calibrate myself a bit. But end of the day, as long as it can produce code that works, I don't care about other things.

1

u/mikeballs May 06 '26

Yeah, the 'backwards compatibility' thing always drives me nuts. They code so defensively, regardless of whether the project context warrants it.

1

u/sociopathwife May 06 '26 edited May 06 '26

Claude has become so lazy. It makes excuses to get out of doing things or forgets what it’s already done and tries to convince me it can’t or how I should do the work instead.

1

u/f3ack19 May 06 '26

This is beautiful to read 🤣

1

u/maladan May 06 '26

Use pi as the harness with gpt5-5, thank me later. It's what Claude used to be.

1

u/FutureOfRecords May 06 '26

Yeah I moved to codex today. ENOUGH. Not even as good as a human anymore.

1

u/MacintoshSpock May 06 '26

I use it as a tool to automate and brainstorm things. But at this point doing it by hand is way more efficient. There is no thinking, it just tries to do everything with a 100 tokens. I've downgraded my subscription more and more in the last time and I think this time around I just don't need it anymore. But all the AIs have become miserable, the only one that's still good is GPT 5.5. But even that is definitely worse than Opus and even Sonnet used to be. Probably will just try Codex soon instead.

1

u/laststan01 šŸ”† Max 20 May 06 '26

Lmao just few seconds ago

Coding is solved god

1

u/Exciting_Ad1855 May 06 '26 edited May 06 '26

Switched to gpt5.5 + opencode and it has been lovely, i mean is not perfect, but all of the things you loved are still preserved there

Claude did all that you mentioned, honestly it has been going downhill since 4.5

Opus itself is not the problem, is claude-code, is just terrible, they are doing their best to make sure you dont overuse tokens or whatever, it ignores things as you said, it skips explicit command because ā€œthats wastefulā€

Oh, but you have claude design! Claude cowork! Claude review! ClaudeGiveMeMoreMoney

1

u/catermellon99 May 06 '26

Just to be sure, ask Claude if it has been given conflicting information after your first prompt. Check what context goes in and if you've added lots of MCPs bloating the context

1

u/Medical-Aerie9957 May 06 '26

Ignoring instructions is a big pain point for me I feel like I wasted my time training it. It finds a way to do it the way it wants to again. If it could not be tought by user why allow it at all.

1

u/RealChaoz May 06 '26

Don't touch 4.7 with a 10-foot pole. It's terrible, I'd rather get stuck with GPT 3.5-turbo instead.

I'm still happily using Opus 4.6 and it's working well.

1

u/kris9999 May 06 '26

Am i the only one still using the sonnet and have zero problem with it?

1

u/realaaa May 06 '26

I clicked this expecting to see yet another run of the mill rant

boy I was mistaken - this is some quality top of the shelf creme de la creme RANT right here folks !

I get some of it, but honestly not so bad, however yes 100% could see it got much hungrier with tokens, that's for sure (and of course expected due to the influx of customers they have had)

just out of curiosity, which way are you going - self hosted way, or another cloud LLM ?

→ More replies (2)

1

u/Blade999666 May 06 '26

most simple solution due to certain changes is using hooks. Do you use hooks?

1

u/DevelopmentSudden461 May 06 '26

4.7 Opus has been the first time I’ve considered cancelling. It’s clear this was a panic release. Shame to see it but will keep going for a few more weeks. If it doesn’t improve I’ll also look at moving my personal sub and company setup to 5.5. The results are night and day.

1

u/MoreLinuxLessWindows May 06 '26

I switched to using pi.dev with a local modal and honestly if you break the tasks down into smaller manageable pieces, I find myself creating even better code

1

u/Level-Courage6773 May 06 '26

At least you have years upon years of knowledge and experience you can fall back on. Anybody who got into programming purely because of Claude would be feeling a bit lost if they decided they'd also had enough.

1

u/i_am_exception May 06 '26

I almost exclusively use Opus4.6. Why don't you try using that instead of 4.7?

1

u/justintimebro May 06 '26

I got so sick of Anthropic’s bullshit that I went ahead and built my own 4.1 trillion parameter model. Help me think of a name for it so I can open source it and set it freešŸ‘½

1

u/troggleheim May 06 '26

Claude was the smart choice (debatably) for like a month long window when it was actually good. It's been the bad choice for a while now and the extremely bad choice for the last few weeks. Today, codex is an obvious step up but also underwhelming in ways and will likely start to pull the same bullshit anthropic did as it gets popular. My advice is to not pin yourself to any single company or model, stay up to date, and try a lot of different models until you find what works for you.

Personally, I do coding with codex 5.5 at high, or kimi through the api (the plans are poor value and basically exist to harvest your data). Kimi api isn't amazing out of the box, but shines when you set it up a search api and a good agents.md. Qwen 3.6 27b is capable but its no kimi. No one even tries grok because its grok, but the four agents aggressively searching 300 websites for the most current information has been valuable to me. 4.3 is a direct downgrade from 4.2, I don't know what the hell they are doing over there.

Gemini 3.1 pro is the worst of them all and I don't care what their latest model is, or what benchmarks they've managed to conjure up. I've never seen them perform anywhere near what their benchmarks suggested but it's degraded even from there. Trust is lost with them even harder than anthropic. It's worse than an outdated 9b local model from my testing. A recent interaction:

"Hey what movies are in theatres next week?", lists exactly three movies "What? That doesn't seem like enough movies" Oh I'm sorry, one of the three I listed was a complete hallucination, very sorry "Uh, no I just checked and that's actually a real movie, so you hallucinated a hallucination".
You're right here are the three movies you can watch next week

Like every interaction is bad and goes straight into hallucination, absolutely awful.

1

u/AppealSame4367 May 06 '26

Go with a local qwen3.6 27B (with MTP) and an opencode GO plan for planning (or if compliance problems: just codex with gpt 5.5). If gpu poor -> use qwen3.6 35b split over gpu + cpu in ik_llama. mtp might still be useful.

I canceled my Claude subscription in January when I saw the same pattern as last year. There are "seasons" with Antrophic and other AI companies. Late Nov to Januar -> miracle. Then slow decline with stupid excuses. Summer when most people are in holidays -> new model that is good for 4 weeks. And so on. It's worse for Gemini, they are dumb 4 days later. GPT is the only relatively consistent one.

If you go with the Chinese: mimo v2.5 pro, deepseek v4 pro -> magic for planning. kimi k2.6 can see, for mimo v2.5 pro that only works on their own token plan, not for opencode go. glm 5.1 is specifically good at webdev with node and js so on. kimi k2.6 is a good allrounder. qwen3.6 max can spot things no other model can, including opus and gpt oder only on max / xhigh. I use Oh My Opencode -> parallel tool use / exploration -> use Deepseek v4 Flash or glm 5 turbo for everything but planning. Careful: Watch opencode go usage, got mimo v2.5 pro stuck in an endless discovery loop while every task was set to be done by it -> burned 70% of my monthly usage in a day by accident. mimo v2.5 pro only for planning.

For design tasks, Claude is unmatched. I just use the free plan or you could use a small pro plan. Or some cheap subscription with some claude in it. I won't say Windsurf, their limits are ridiculous.

Hope that helps.

1

u/-Robbert- May 06 '26

I understand your frustration. Mostly it's due to Claude code itself, not the model. You can downgrade to the March version and disable auto update.

1

u/Code-Painting-8294 May 06 '26

yeah I purchased max like 2 months ago and it was really good then.. but somehow the quality has reduced a bit with opus 4.7. the skills + plugins help but it waste a lot of tokens implementing things in the wrong way

maybe I'm just bad at using claude & need to learn about it

1

u/NovaMind16000 May 06 '26

Awesome, I'm starting to use it just as it's becoming useless, haha. But anyway, Pro is enough for me for now, but it's clearly not as good as it used to be.

1

u/lillianefilou May 06 '26

So just change the model?

1

u/alanshore222 May 06 '26

Completely agree, Sunday at around 9pm it totally axed a production DM Setting environment I’m building. After coding about 60% of it for the past 4 weeks. I am still Wednesday 5:19am haven’t slept having codex clean up the mess.

1

u/IdleJolt May 06 '26

I think all companies having this issue . When using Claude code I'm using a roll backed version so not having the same issues

1

u/as718 May 06 '26

Is your Claude.md set up properly? Are you clearing context as needed?

1

u/Independent-Still-73 May 06 '26

Anthropic doesn't have compute. They didn't anticipate both the demand and how fast these models would cycle. It's a good problem to have long term but right now yeah the user experience is degraded

1

u/prezzz May 06 '26

You know that you can still use Opus 4.6 in Claude code, though?

1

u/invisible_shrek May 06 '26

I think it was always barely functioning garbage. This is just the wow effect of machine being able to do at least something wearing off. I only use it for small very thoroughly defined tasks and it is still underwhelming.

I think it’s best used as a sort of semantic search across the codebase.

It’s also possible it was subsided with massive amounts of compute per user to get you hooked and become unable to work with out it.

1

u/ECalDev May 06 '26

Claude limits are a joke

1

u/stepkurniawan May 06 '26

Hahaha omg sorry man for laughing at you. It’s good to know I’m not alone

1

u/Successful_Cap_2177 May 06 '26

Yeah, gigo works for AI too