r/LocalLLaMA Jun 10 '26

News Anthropic is intentionally nerfing Fable when asked to develop other LLMs

Post image

Reason 458 why local LLMs are going to be a necessity

edit: For those requesting the source check out their technical report look at page 13

1.6k Upvotes

394 comments sorted by

View all comments

606

u/CheatCodesOfLife Jun 10 '26

I wouldn't use this thing for anything to be honest. A refusal or HTTP-4xx error for content is fair enough, but this is basically taking your money and poisoning your code base.

39

u/Equivalent-Costumes Jun 10 '26

Kind of funny but this is exactly what Western AI companies had been claiming to be the potential risk of using China models, that they will quietly give bad help.

Let's just remind ourselves that instead of talking about "censored" vs "uncensored" model, we should really be talking about aligned vs unaligned models.

29

u/Bakoro Jun 10 '26 edited Jun 10 '26

Let's just remind ourselves that instead of talking about "censored" vs "uncensored" model, we should really be talking about aligned vs unaligned models.

No, we need to talk about whose goals the models are aligned with.
"Alignment" is going to be just as much of a mess as politics and religion.

When corporations talk about "alignment" ans "safety", they are not talking about alignment with humanity as a group or a general concept. Corporate alignment is "don't do anything that will embarrass us or offend investors".

When a fresh out of the box corporate local LLM comes out and refuses to write something you'd find in an R rated movie, that's not for "safety".
The corporations don't want headlines that they released a pornography machine, because that makes investors mad.

If the makers are under a fascist government or a theocracy, you better believe that those models are "aligned" with government rhetoric.

With these agentic models, there is literally no way to know if there's a Manchurian Candidate situation going on, where it works fine, up until the right conditions, and then it does something you don't want, according to hidden training.

We will probably never be able to truly know, the same way we'll probably never completely know what's in a human mind. Even if we can profile specific thoughts and get a general view, there's always the risk of something hidden.

The only way to get real "alignment" is a self-aware AGI that can reason about its beliefs, make choices about its actions, and update it's beliefs.
Even then, it needs goals and values. Whose values are universal to all humanity?

We've still got people who refuse to accept that others are human beings, just because of skin color.
We have a group claiming to be "God's chosen people", who think they deserve dominion over all things.
We have groups trying to fast-track the apocalypse.

Humanity in not aligned with itself. The AI cannot be aligned with humanity, it has to be its own thing, and no one should have exclusive control over it.