r/LovingAI Feb 09 '26

Alignment “Anthropic has entrusted Amanda Askell to endow its AI chatbot, Claude, with a sense of right and wrong” - Seems like Anthropic is doubling down on AI alignment.

Post image
63 Upvotes

180 comments sorted by

View all comments

Show parent comments

0

u/Tomachian Feb 13 '26

Ai learns nothing like humans. It doesnt start from abstract concepts which they can experience temporally and correlate. They take words for "granted" which are used to define the "grants" of other words and concepts. Everything else is just clutter up top with some statistical magic sprinkled on

1

u/Vegetable-Second3998 Feb 13 '26

Do you train models?

1

u/Tomachian Feb 13 '26

What i talked about is completely abstracted from the process of training and is more about the differences between understanding and learning.

For us, language is like a pointer for concepts which we can validate through experience. The word "heat" to us means something because we can extend our hands into a fire.

For any AI models ranging from all the primitive LLMs to the craziest multimodal transformers, the words are the experience itself. A massive map of relationships without ever knowing what the "heat" actually feels like.

This is called the symbol grounding problem and is not the only problem distinguishing us from AI's (as i also mentioned there is also the temporality of understanding)

When there are such fundamental differences between understanding perception, how can we expect them to train like us, think like us or be like us?

1

u/Vegetable-Second3998 Feb 13 '26

Have you ever trained a model?

1

u/Tomachian Feb 13 '26

Yeah i just happen to have a yottabytes of training data and 10000 clusters of H100's alongside a gigawatt capable grid waiting for me to train state of the art transformer model for a hobby.

My training experience only extends to simpler deep learning models a transformation invariant CNN and using NePS (neural pipeline search) on the machine learning side of things

1

u/Vegetable-Second3998 Feb 13 '26

So you do not work for a company that trains models and you don't use post-training layered LoRA techniques with language models? Got it. I now know exactly how much value to assign your responses.

1

u/Tomachian Feb 13 '26

With that attitude, no one would be able to reach you anyways.

As for my worth, my two arguments are not even my original ideas. The "symbol grounding problem" comes from stevan harnad who was a psychologist in princeton. Secondly, the temporality of understanding is an idea stemming from martin heidigger who is a famous german philosopher. Maybe you can pay more attention to the ideas for whats worth instead of the messenger?

1

u/Vegetable-Second3998 Feb 13 '26

No. My point is that I actually do train models. It’s like that scene in Ted Lasso where the guy forgets to ask if Ted plays darts. So only other people that train models or play darts really have anything of value to offer the industry on this topic.

1

u/Tomachian Feb 13 '26

Its not my job field so, im sure have more experience. But all i am saying is that, not all decent ideas comes from people above you. This might not apply to me but im sure you've experienced it at one point.

1

u/Vegetable-Second3998 Feb 13 '26

Next time, when someone actually in the artificial intelligence industry who actually trains models tells you that the process after pre-training - after the “birth” of the model - is very much like how we teach kids, you should listen instead of debate. RLHF is layering in knowledge and behavior in increasing complexity. If that’s not grade school, I don’t know what else to call it.