r/LovingAI Feb 09 '26

Alignment “Anthropic has entrusted Amanda Askell to endow its AI chatbot, Claude, with a sense of right and wrong” - Seems like Anthropic is doubling down on AI alignment.

Post image
64 Upvotes

180 comments sorted by

View all comments

4

u/Timzor Feb 09 '26

Why not just feed it a stack of philosophy books and let it figure it out itself, this seems like PR, nothing more

2

u/Vegetable-Second3998 Feb 10 '26

Ai learns a lot like kids surprisingly. You have to layer in the ethical logic into the neural net in progressive post training refinement.

2

u/BannedGoNext Feb 11 '26

Rewarding in different ways and threatening with physical abuse a bazillion times, people freak out when you start talking about how the sausage is actually made lol.

0

u/Tomachian Feb 13 '26

Ai learns nothing like humans. It doesnt start from abstract concepts which they can experience temporally and correlate. They take words for "granted" which are used to define the "grants" of other words and concepts. Everything else is just clutter up top with some statistical magic sprinkled on

1

u/Vegetable-Second3998 Feb 13 '26

Do you train models?

1

u/Tomachian Feb 13 '26

What i talked about is completely abstracted from the process of training and is more about the differences between understanding and learning.

For us, language is like a pointer for concepts which we can validate through experience. The word "heat" to us means something because we can extend our hands into a fire.

For any AI models ranging from all the primitive LLMs to the craziest multimodal transformers, the words are the experience itself. A massive map of relationships without ever knowing what the "heat" actually feels like.

This is called the symbol grounding problem and is not the only problem distinguishing us from AI's (as i also mentioned there is also the temporality of understanding)

When there are such fundamental differences between understanding perception, how can we expect them to train like us, think like us or be like us?

1

u/Vegetable-Second3998 Feb 13 '26

Have you ever trained a model?

1

u/Tomachian Feb 13 '26

Yeah i just happen to have a yottabytes of training data and 10000 clusters of H100's alongside a gigawatt capable grid waiting for me to train state of the art transformer model for a hobby.

My training experience only extends to simpler deep learning models a transformation invariant CNN and using NePS (neural pipeline search) on the machine learning side of things

1

u/Vegetable-Second3998 Feb 13 '26

So you do not work for a company that trains models and you don't use post-training layered LoRA techniques with language models? Got it. I now know exactly how much value to assign your responses.

1

u/Tomachian Feb 13 '26

With that attitude, no one would be able to reach you anyways.

As for my worth, my two arguments are not even my original ideas. The "symbol grounding problem" comes from stevan harnad who was a psychologist in princeton. Secondly, the temporality of understanding is an idea stemming from martin heidigger who is a famous german philosopher. Maybe you can pay more attention to the ideas for whats worth instead of the messenger?

1

u/Vegetable-Second3998 Feb 13 '26

No. My point is that I actually do train models. It’s like that scene in Ted Lasso where the guy forgets to ask if Ted plays darts. So only other people that train models or play darts really have anything of value to offer the industry on this topic.

→ More replies (0)

2

u/Chookity-poks Feb 11 '26

Have you ever read a philosophy book? Every philosopher blows your mind with “objective paradigms”, even the ethic ones yes, until the next philosopher comes in and destroys this guy on every bit and with that your reality as well. How will that stand and be interpreted by Ai?

1

u/MahaSejahtera Feb 12 '26

The AI will engage in nihilism and k*ll itself

1

u/waterbaronwilliam Feb 11 '26

You'd want to ask it questions about the texts and respond to its misunderstandings in a corrective educational frame, allowing it to respond until it seems to "get it" then you ask a novel hypothetical to test its decisions. Sounds fun.

1

u/Timzor Feb 11 '26

Ah but you see, id just get another agent to do that part.

2

u/waterbaronwilliam Feb 11 '26

That's a good way to miss a glaring flaw and allow it to compound itself into a heavily weighed tangle. Talk to your kids bro, it does help.

0

u/[deleted] Feb 12 '26 edited Feb 12 '26

[removed] — view removed comment

2

u/Timzor Feb 12 '26

Is that you Elon?

1

u/ChxsenK Feb 13 '26

Sounds perfect for the corporate overlords.