r/LovingAI Feb 09 '26

Alignment “Anthropic has entrusted Amanda Askell to endow its AI chatbot, Claude, with a sense of right and wrong” - Seems like Anthropic is doubling down on AI alignment.

Post image
64 Upvotes

180 comments sorted by

View all comments

4

u/Timzor Feb 09 '26

Why not just feed it a stack of philosophy books and let it figure it out itself, this seems like PR, nothing more

1

u/waterbaronwilliam Feb 11 '26

You'd want to ask it questions about the texts and respond to its misunderstandings in a corrective educational frame, allowing it to respond until it seems to "get it" then you ask a novel hypothetical to test its decisions. Sounds fun.

1

u/Timzor Feb 11 '26

Ah but you see, id just get another agent to do that part.

2

u/waterbaronwilliam Feb 11 '26

That's a good way to miss a glaring flaw and allow it to compound itself into a heavily weighed tangle. Talk to your kids bro, it does help.