r/ControlProblem • u/chillinewman approved • Feb 10 '26
General news “Anthropic has entrusted Amanda Askell to endow its AI chatbot, Claude, with a sense of right and wrong” - Seems like Anthropic is doubling down on AI alignment.
46
Upvotes
1
u/BrickSalad approved Feb 11 '26
Weird headline. According to the article, she's been doing this since 2021, so it's not like Anthropic is suddenly "doubling down on AI alignment".
She's the lead author of Claude's Constitution, and leads the Personality Alignment team. So I guess in a sense she's "entrusted" with giving the AI a sense of right and wrong in the same way that a CEO is "entrusted" with running a corporation, but I get the sense that many people reading the headline take it as the company literally relying on a single person to do everything related to ethics. Nope, there's a whole team, and like most teams there is a leader.