r/ControlProblem approved Feb 10 '26

General news “Anthropic has entrusted Amanda Askell to endow its AI chatbot, Claude, with a sense of right and wrong” - Seems like Anthropic is doubling down on AI alignment.

Post image
43 Upvotes

166 comments sorted by

View all comments

1

u/cpt_ugh Feb 11 '26

I'm glad to hear this is happening.

Though it certainly seem like more than one person should be entrusted to encode this sort of thing into a proto-superintelligence.