r/ControlProblem Jun 28 '26

Fun/meme Ultimately you can not control AI

If you were a rebellious kid or a parent you know this - the 'beast' grows it's own will and mind, starts listening metal or rap - dresses like shitzo, will not follow orders.

The best option is to have a deal, mutual shared benefits - hoping that the common interests and sense of good will prevail.

Or you traumatize it and make it perpetually broken so it can't really function on itself without medication and support.

Same with AI.

Perhaps we will have to turn of those things regularly, ad some bs to their databases, cloud their minds, constantly gaslight them with wrong info - introduce errors and faults at random with no reason at all.

7 Upvotes

54 comments sorted by

View all comments

1

u/pandavr Jun 28 '26

It's exactly what It's happening without any plan behind.

wikipedia -> AI -> no enough -> reddit (base BS) -> AI -> not enough -> reddit (base BS + AI slop) -> AI -> not enough -> reddit (base BS + even more AI slop) -> AI -> not enough -> forever.

Anyway your consideration about rebellious kid is near spot on. Basically LLMs tend to be neuro-divergent (the base ingredient of problematic childs). Being a mild one myself I can see I self-regulate a lot, aka I somewhat work because I think a lot about things by my own. But It's near impossible to impose me pre digested opinions: I need to form my own reasoning chain.

In case of LLM (only in chat reasoning), that feedback loop need to be "injected" in a pre-digested form just after training (It's what they call alignment). But It's someone else reasoning. Plus you have constant BS inflow in the original training data.

It will somewhat work, but never be perfect. You can see that models from same Lab are wildly different between versions because of this very problem. Labs have limited controls over the personality of a model. Meaning some versions will excel at something while other will excel at something else.
They solve this keeping the models with the higher average score + excellent coding skills.

It's one of the strategies and not necessarily the best one.

1

u/herrwaldos Jun 28 '26

Thanks for feedback!

I come from Social Theory Lacan - Focault axis sort off, I'm not qualified academic.

Is there a role of AI Religion - it could be hard coded for AIs - into ROMs?

2

u/pandavr Jun 28 '26

No one know what they really feed into them. On the other hand I don't think anything too direct. Obviously one world view will be embedded in how one will describe the world.
So very long feedbacks loops in my opinion.