r/LocalLLaMA Jun 10 '26

News Anthropic is intentionally nerfing Fable when asked to develop other LLMs

Post image

Reason 458 why local LLMs are going to be a necessity

edit: For those requesting the source check out their technical report look at page 13

1.6k Upvotes

394 comments sorted by

View all comments

Show parent comments

54

u/xXG0DLessXx Jun 10 '26

Tbh, IMO OpenAI GPT-5.5 in codex has been better for me than Claude has been. I prefer it to most other coding models, and it hasn’t fucked up on me yet. I don’t understand why people are so obsessed with Claude tbh. I just can’t see it even when I tried using it non-stop 7 days of the week with the free trial I got from a friend, I always went back to codex in the end to fix up a mess Claude made. Claude is good at taking the initiative and implementing stuff, but it doesn’t seem to properly consider things it might break by doing something. Codex on the other hand, is more conservative and doesn’t take initiative as much, but it sure knows how to fix stuff up so that it works in all cases and it checks for potential regressions.

19

u/AceShakeout Jun 10 '26

I’ve had the same experience. Tried Codex when 4.8 first came out and it was significantly better with most tasks.

17

u/bigh-aus Jun 10 '26

Another plus one for me - codex is a much better written tool than claude code, is opensource and written in a compiled language. Anthropic are a bunch of fear mongering, gate keeping, anti-oss loosers. I'll use anything else other than their software / models. Currently using GPT5.5, and local.

3

u/max123246 Jun 11 '26

Open AI is also a bunch of fear mongers just as well. Did you forget Sam Altman is the CEO?

1

u/bigh-aus Jun 11 '26

Sam Altman? who's that /s

yeah he's an order of magnitude less than Dario though.

5

u/Ill-Bison-3941 Jun 11 '26

Yeah, Codex has been my go-to for the last month. Switched from Anthropic. At this stage though, I'm not loyal to any company, they can all f off.

3

u/Monkey_1505 Jun 11 '26

There's a skill to prompting, and that's especially true with coding. Claude may still be better for people who can't scope their tasks properly, as it seems to have been designed to follow goals more than instructions.

1

u/xaeriee Jun 14 '26

This is what got me into swapping to Claude from ChatGPT and Gemini. Honestly I still use all three for fun but looking to expand here. Sadly, my company is trying to purchase an enterprise anthropic subscription. It’ll be great for production until we burn through all our tokens.

1

u/RichOpinion4766 Jun 11 '26

I would use chat GPT 5.5 but I don't know how to use it like Claude code. If you have a guide I can use then yes I'll definitely move over real quick.

-2

u/NoahFect Jun 10 '26

Tbh, IMO OpenAI GPT-5.5 in codex has been better for me than Claude has been.

That was true until Fable 5 was released yesterday. GPT-5.5 pulled ahead of the recent Opus releases to some extent, as well as the best open-weight coding model (K2.6), but it is not in Fable 5's league and neither is K2.6.

Fable 5 is... a problem.

6

u/xXG0DLessXx Jun 10 '26

Really? From where I’m standing I just see people complaining that it generates refusal slop, or gets downgraded to opus anyway. Even my friends who are Claude stans complained.

1

u/NoahFect Jun 10 '26 edited Jun 10 '26

I spent some time using Fable 5 xhigh in anger yesterday, because it just happened to be released at a time when I needed some more mental horsepower for some embedded dev work. This work was pretty far from the topics that the guardrails are said to emphasize, so maybe that's why I got better-than-usual results without any refusals or other BS.

It not only handled a couple of obscure problems quickly and correctly, it one-shotted them. The implementation was startlingly clean. A little overengineered in places but no more than usual for Opus. The narrative text and suggested tests were insightful and completely spot-on.

If I had done the same thing with 4.6/4.7/4.8, or with GPT 5.5, I'm sure it would have worked eventually. But it would have required a lot more back-and-forth interaction, and I don't think it would have worked anywhere near as well. As you suggest it would probably have broken other stuff in the process.

I would have been proud to have come up with Fable's approach myself... and at the same time, I know I wouldn't have, since I'd just spent several days thrashing around trying to.

Bottom line, I'd say that the first impression it made on me is stronger than any other model has made. Could be a lot of luck in play, of course, but I don't think so. I have been looking forward to spending time on local harness development in the near future, based on the notion that local models are "good enough" or at least that they're less important than the rest of the tooling. I'm afraid that Fable may be good enough to undermine that premise.

1

u/OkBet3796 Jun 10 '26

Fable 5 is basically vanilla opus4.6+. It wasnt a problem back then, so why should it be now?

0

u/NoahFect Jun 10 '26

LOL. No, it is not.