A philosopher declined a $335,000 annual job offer from Anthropic to help shape the behavior and values of its Claude AI models, saying her personal principles outweighed the financial opportunity. The role was part of Anthropic’s broader effort to involve philosophers, ethicists, and scholars in the development of Claude’s behavior and value framework.
Explaining her decision, she argued that society is becoming overly dependent on artificial intelligence, warning that delegating more everyday tasks to AI risks eroding essential human abilities. “The more tasks we hand over to neural networks, the more skills we lose. And each time, we become lonelier and less intelligent,” she said.
Anthropic has increasingly incorporated experts from philosophy, ethics, religion, and other disciplines into Claude’s development as part of its constitutional AI approach. The company says these perspectives help shape the values and behaviors embedded in its models rather than relying solely on technical research.
The decision highlights the growing debate over AI’s impact on society. As companies race to build increasingly capable models, questions about how AI affects human judgment, critical thinking, creativity, and social interaction are becoming just as prominent as discussions around the technology itself.