Anthropic launches $5 million AI wellbeing research grant program
Anthropic announced funding, model access and technical support for independent, open-source evaluations of AI’s effects on user wellbeing.
ひとことで言うと
What did Anthropic announce about funding research into AI’s impact on user wellbeing?
Anthropic launched a $5 million grant program for independent research into AI’s effects on user wellbeing. Selected grantees will receive direct funding, access to Anthropic’s models and technical support while independently developing open-source evaluations. Applications are due September 21, with full-proposal invitations scheduled by October 5.
要点
- Anthropic launched a $5 million grant program for independent research into how AI affects users’ wellbeing.
- The program will provide grantees with direct funding, access to Anthropic’s models and technical support.
- Grantees will work independently and publish their evaluations as open-source projects available to developers.
- Anthropic is seeking evaluations that include subject-matter expertise, realistic conversations and validation against human experts.
- Applications are due September 21, and selected applicants will be invited to submit full proposals by October 5.
Anthropic’s grant program
Anthropic announced a $5 million grant program to support independent research into how AI affects users’ wellbeing. The company said the program will provide direct funding, access to its models and technical support to grantees developing open-source evaluations.
Grantees will conduct their work independently, according to Anthropic. Their projects will be published as open source so that developers can use them.
Anthropic said it wants the program to involve clinicians, psychologists, methodologists and other specialists in creating evaluations and benchmarks of user wellbeing.
Focus of the evaluations
Anthropic said wellbeing can require more conversational context to evaluate than model behavior that can be judged from a single response. The company described situations in which risks become apparent only over a longer exchange or where a response may be appropriate in one context but harmful in another.
One example in the announcement concerns advice about diet and exercise. Anthropic said such advice could be reasonable when a user asks about losing weight, but could be inappropriate or harmful if the user has shown a history of disordered eating. It also cited conversations involving companionship and mental health crises as areas where standards for model behavior are still being developed.
The company said it develops safeguards intended to identify such conversations and help Claude respond appropriately. Anthropic also said it publishes research about the kinds of conversations users have with Claude to inform safeguards, their evaluation and other user-wellbeing measures.
Evaluation guidance
Alongside the grant program, Anthropic shared guidance from its Safeguards team about the characteristics it believes make wellbeing evaluations rigorous and the challenges that may limit their usefulness.
Anthropic said it is seeking evaluations that:
- Clearly define what they measure, including what constitutes a pass or failure and why the measurement matters.
- Include clinical and subject-matter experts in their design and validation.
- Test both precautions and harms, including risks from overcompliance and overrefusal.
- Represent how people use AI, including multi-turn scenarios in which risk increases and context changes during a long conversation.
- Validate automated or other graders against real subject-matter experts.
Anthropic directed prospective applicants to an application form and separately provided guidance on building wellbeing evaluations and benchmarks. Applications are due September 21. The company said applicants selected to submit full proposals will receive notification by October 5.
Source: Anthropic’s “Funding better evaluations of AI’s impact on wellbeing” announcement, published August 25, 2026.
よくある質問
- How much funding is Anthropic providing?
- Anthropic says the grant program will provide $5 million for independent research into AI’s impact on user wellbeing.
- What support will grantees receive?
- The company says grantees will receive direct funding, access to its models and technical support.
- Will the research be publicly available?
- Anthropic says grantees will publish their work as open-source projects that any developer can use.
- When are applications due?
- Applications are due September 21. Applicants selected to submit full proposals will be notified by October 5.