Funding better evaluations of AI’s impact on wellbeing
Article image or reusable cover for Anthropic
Anthropic is allocating 5 million dollars to fund independent research on how AI models affect user wellbeing.
The program provides grants, access to Anthropic's models, and technical support to researchers building open evaluation tools. Since AI is now central to how people work, learn, and seek emotional support, the industry needs better standards for how models should behave—for example when a user seeks community or is going through a mental health crisis. Wellbeing is difficult to measure because it requires long-term contextual understanding; a good response in one context can cause harm in another.
But assessing wellbeing requires much more context. For example, a user in distress might not share thoughts of self-harm right away; the need for a more cautious response might only become clear over the course of a long conversation.
Vibekollen prepared this summary with AI from the original publication. The content belongs to Anthropic.