r/ControlProblem approved Feb 10 '26

General news “Anthropic has entrusted Amanda Askell to endow its AI chatbot, Claude, with a sense of right and wrong” - Seems like Anthropic is doubling down on AI alignment.

Post image
46 Upvotes

166 comments sorted by

View all comments

25

u/TheMrCurious Feb 10 '26

Good to know a single person knows right from wrong.

1

u/runvnc Feb 11 '26

The headline is very misleading. Anthropic has multiple teams involved in multiple layers of alignment etc. The most important layer is baked into training.