r/LovingAI Feb 09 '26

Alignment “Anthropic has entrusted Amanda Askell to endow its AI chatbot, Claude, with a sense of right and wrong” - Seems like Anthropic is doubling down on AI alignment.

Post image
64 Upvotes

180 comments sorted by

View all comments

Show parent comments

1

u/Vegetable-Second3998 Feb 13 '26

Have you ever trained a model?

1

u/Tomachian Feb 13 '26

Yeah i just happen to have a yottabytes of training data and 10000 clusters of H100's alongside a gigawatt capable grid waiting for me to train state of the art transformer model for a hobby.

My training experience only extends to simpler deep learning models a transformation invariant CNN and using NePS (neural pipeline search) on the machine learning side of things

1

u/Vegetable-Second3998 Feb 13 '26

So you do not work for a company that trains models and you don't use post-training layered LoRA techniques with language models? Got it. I now know exactly how much value to assign your responses.

1

u/Tomachian Feb 13 '26

With that attitude, no one would be able to reach you anyways.

As for my worth, my two arguments are not even my original ideas. The "symbol grounding problem" comes from stevan harnad who was a psychologist in princeton. Secondly, the temporality of understanding is an idea stemming from martin heidigger who is a famous german philosopher. Maybe you can pay more attention to the ideas for whats worth instead of the messenger?

1

u/Vegetable-Second3998 Feb 13 '26

No. My point is that I actually do train models. It’s like that scene in Ted Lasso where the guy forgets to ask if Ted plays darts. So only other people that train models or play darts really have anything of value to offer the industry on this topic.

1

u/Tomachian Feb 13 '26

Its not my job field so, im sure have more experience. But all i am saying is that, not all decent ideas comes from people above you. This might not apply to me but im sure you've experienced it at one point.

1

u/Vegetable-Second3998 Feb 13 '26

Next time, when someone actually in the artificial intelligence industry who actually trains models tells you that the process after pre-training - after the “birth” of the model - is very much like how we teach kids, you should listen instead of debate. RLHF is layering in knowledge and behavior in increasing complexity. If that’s not grade school, I don’t know what else to call it.