Skip to content
安全・倫理

Anthropic launches $5 million AI wellbeing research grant program

Anthropic announced funding, model access and technical support for independent, open-source evaluations of AI’s effects on user wellbeing.

DigitalNeuron Desk約2分

ひとことで言うと

What did Anthropic announce about funding research into AI’s impact on user wellbeing?

Anthropic launched a $5 million grant program for independent research into AI’s effects on user wellbeing. Selected grantees will receive direct funding, access to Anthropic’s models and technical support while independently developing open-source evaluations. Applications are due September 21, with full-proposal invitations scheduled by October 5.

要点

  • Anthropic launched a $5 million grant program for independent research into how AI affects users’ wellbeing.
  • The program will provide grantees with direct funding, access to Anthropic’s models and technical support.
  • Grantees will work independently and publish their evaluations as open-source projects available to developers.
  • Anthropic is seeking evaluations that include subject-matter expertise, realistic conversations and validation against human experts.
  • Applications are due September 21, and selected applicants will be invited to submit full proposals by October 5.

Anthropic’s grant program

Anthropic announced a $5 million grant program to support independent research into how AI affects users’ wellbeing. The company said the program will provide direct funding, access to its models and technical support to grantees developing open-source evaluations.

Grantees will conduct their work independently, according to Anthropic. Their projects will be published as open source so that developers can use them.

Anthropic said it wants the program to involve clinicians, psychologists, methodologists and other specialists in creating evaluations and benchmarks of user wellbeing.

Focus of the evaluations

Anthropic said wellbeing can require more conversational context to evaluate than model behavior that can be judged from a single response. The company described situations in which risks become apparent only over a longer exchange or where a response may be appropriate in one context but harmful in another.

One example in the announcement concerns advice about diet and exercise. Anthropic said such advice could be reasonable when a user asks about losing weight, but could be inappropriate or harmful if the user has shown a history of disordered eating. It also cited conversations involving companionship and mental health crises as areas where standards for model behavior are still being developed.

The company said it develops safeguards intended to identify such conversations and help Claude respond appropriately. Anthropic also said it publishes research about the kinds of conversations users have with Claude to inform safeguards, their evaluation and other user-wellbeing measures.

Evaluation guidance

Alongside the grant program, Anthropic shared guidance from its Safeguards team about the characteristics it believes make wellbeing evaluations rigorous and the challenges that may limit their usefulness.

Anthropic said it is seeking evaluations that:

  • Clearly define what they measure, including what constitutes a pass or failure and why the measurement matters.
  • Include clinical and subject-matter experts in their design and validation.
  • Test both precautions and harms, including risks from overcompliance and overrefusal.
  • Represent how people use AI, including multi-turn scenarios in which risk increases and context changes during a long conversation.
  • Validate automated or other graders against real subject-matter experts.

Anthropic directed prospective applicants to an application form and separately provided guidance on building wellbeing evaluations and benchmarks. Applications are due September 21. The company said applicants selected to submit full proposals will receive notification by October 5.

Source: Anthropic’s “Funding better evaluations of AI’s impact on wellbeing” announcement, published August 25, 2026.

よくある質問

How much funding is Anthropic providing?
Anthropic says the grant program will provide $5 million for independent research into AI’s impact on user wellbeing.
What support will grantees receive?
The company says grantees will receive direct funding, access to its models and technical support.
Will the research be publicly available?
Anthropic says grantees will publish their work as open-source projects that any developer can use.
When are applications due?
Applications are due September 21. Applicants selected to submit full proposals will be notified by October 5.

出典

  1. Funding better evaluations of AI’s impact on wellbeing \ AnthropicAnthropic
タグanthropicclaudeai-wellbeingresearch-grantsai-safetyevaluations

あわせて読みたい

アントロピックとOpenAIがそれぞれのドキュメントに記載している5つのプロンプティング習慣

両方のラボは同じ中核的なアドバイスを公開しています。指示は質問ではなくコマンドとして書き、モデルに知らないことを言う明確な許可を与え、コンテキストを質問の前に置き、プロンプトごとに1つのタスクを維持し、モデルにどれだけ積極的になるかを伝えます。共通点は、それらのすべてが曖昧さを排除していることです。それらのどれも、隠された品質を解除するために貼り付けるフレーズではありません。

約6分