「AIの方が思いやりがある」――プロの相談員より共感的だと評価された理由

読む前に予想してみよう!
「これはAIが書いた返事です」と明かされると、AIの評価は人間より下がった
作者を明かしても、AIの返答のほうが思いやりがあると評価された。
- 対象
- カナダなどの成人参加者
- 規模
- 556人(4実験の合計)
- 研究の種類
- 事前登録されたオンライン実験
- 確かさ
- ★★★☆☆
4つの実験で、第三者はAIの返答を人間の返答より「思いやりがある」と評価。危機相談のプロと比べても、AIだと明かしても(差は縮んだが)AIのほうが上だった。
ひとことで言うとAIの返事は、人間のプロより「やさしい」と感じられた!
つらいことがあったとき、誰かに話を聞いてほしい。でも、その返事を書いたのが人間ではなくAIだったら、あなたはどう感じますか?
2025年、カナダの心理学者たちが「AIの返事と人間の返事、どっちが思いやりを感じるか」を何度も確かめました。
結果は、何度やってもAIの勝ちでした。
01悩みへの返事を読み比べる4つの実験
返事を書いたのは、ChatGPT(GPT-4)と、一般の人、そして電話相談で危機対応の訓練を受けた相談員。どの比較でも、AIの返事のほうが好まれ、より思いやりがあると評価されました。返事の応答性(理解・承認・気づかい)も測った4つ目の実験では、その点でもAIが上回りました。
02「AIが書いた」とバラしても結果は同じ
「人間が書いたと思ったから高評価だったのでは?」という疑問もわきます。そこで研究チームは、誰が書いたかを明かした実験も行いました。
それでもAIの評価は高いまま。AIの返事は、相手の気持ちを理解し、受け止め、気づかっていると感じさせる力が強く、それが「思いやり」の評価を押し上げる一因になっていたことも分析から分かりました。

03なぜAIは「やさしく」見えるのか
研究チームは理由として、AIは疲れないこと、そして相手の文章の細かい部分を拾って返事に生かせることを挙げています。
人間の相談員は、何件も相談を受けると疲れてしまうし、重い相談に気持ちを揺さぶられることもあります。AIは毎回、相手の言葉を丁寧に受け止めるところから始められるのです。
2023年の研究では、ネット掲示板に寄せられた195件の医療相談に対する医師の回答とAIチャットボットの回答を、医療の専門家が読み比べました。その結果、延べ585回の評価のうち約8割(78.6%)でAIの回答の方が好まれ、「共感的」と評価される割合も高かったのです。AIの「やさしい書きぶり」は、医療の世界でも注目されています。
04まとめ
「思いやり」は、気持ちだけでなく伝え方の技術でもあるのかもしれません。
AIから学べるのは、相手の言葉をちゃんと拾い、否定せずに受け止めること。それって、人間同士でも今日から真似できる「やさしさのコツ」ですよね。
この研究を英語で読もう原田英語の English CornerCan a computer show more compassion than a person?
音声
4択クイズ9問A2・B1・B2タップして開く
原田英語の English Corner
この研究を英語で読もう
When we have a problem, we want someone to listen. In 2025, psychologists in Canada asked 556 people to read short stories about good and bad events.
Then they read replies written by AI and by people, including experts who help people in crisis. People rated the AI replies as kinder. This was true even when they knew which replies came from AI. But AI is not a replacement for real experts.
72 words ・ CEFR A2
和訳を見る
困ったことがあると、私たちはだれかに話を聞いてほしいと思います。2025年、カナダの心理学者たちは556人に、うれしい出来事やつらい出来事についての短い文章を読んでもらいました。
次に、AIと人間が書いた返事を読んでもらいました。人間の中には、困っている人を助ける危機相談の専門家も含まれていました。人々はAIの返事のほうをよりやさしいと評価しました。どの返事がAIのものか知っていても、結果は同じでした。しかし、AIは本当の専門家の代わりにはなりません。
4択リーディングクイズ
Q1. What is the passage mainly about?
AIの返事のほうがやさしいと評価された研究が中心なので D。
Q2. How many people took part in the study?
asked 556 people to read short stories とあるので A。
Q3. In this passage, "kinder" means ...
kind は形容詞で「親切な・やさしい」、kinder はその比較級。相手を思いやっているという意味なので C。
Can a computer show more compassion than a person? Psychologists at the University of Toronto tested this in four experiments with 556 participants, published in 2025. Participants read short stories about happy or painful experiences, along with replies written by ChatGPT, by ordinary people, and by volunteers trained to handle crisis calls. They then rated how compassionate each reply was.
In every comparison, the AI replies were rated as more compassionate, and one experiment also found them more responsive. Even when participants were told which replies were written by AI, the result did not change. The researchers suggest that AI does not get tired and is good at picking up small details in what people write. However, the people who rated the replies were not the ones sharing their problems.
130 words ・ CEFR B1
和訳を見る
コンピューターは人間よりも思いやりを示せるのでしょうか。トロント大学の心理学者たちは、556人が参加した4つの実験でこの問いを確かめ、2025年に発表しました。参加者は、うれしい経験やつらい経験についての短い文章と、それに対してChatGPT、一般の人、危機相談の電話に応じる訓練を受けたボランティアが書いた返事を読みました。そして、それぞれの返事がどれくらい思いやりがあるかを評価しました。
どの比較でも、AIの返事のほうが思いやりがあると評価され、1つの実験では、AIの返事のほうが相手にしっかり応えているとも評価されました。どの返事がAIによって書かれたかを参加者に伝えても、この結果は変わりませんでした。研究者たちは、AIは疲れず、人が書いた内容の細かい部分を拾うのが得意だからではないかと考えています。しかし、返事を評価した人たちは、悩みを打ち明けた本人ではありませんでした。
4択リーディングクイズ
Q1. Who wrote the replies in the experiments?
replies written by ChatGPT, by ordinary people, and by volunteers trained to handle crisis calls とあるので A。
Q2. What happened when participants were told which replies came from AI?
even when participants were told which replies were written by AI, the result did not change とあるので C。
Q3. What does the last sentence suggest?
評価したのは悩みを打ち明けた本人ではない=実際に悩む人の受け止め方は分からない、という限界を示しているので B。
Researchers at the University of Toronto asked people to read short accounts of good and bad personal experiences, each paired with a reply. Some replies were written by GPT-4, some by ordinary people, and some by expert crisis responders. Across four preregistered experiments with a total of 556 participants, published in Communications Psychology in 2025, third-party evaluators rated the AI replies as more compassionate in every comparison. This was partly because the AI seemed more responsive.
Its replies conveyed understanding, validation, and care. The advantage held even when evaluators knew which replies came from AI. Empathy has long been treated as a uniquely human strength, so the result challenges a common assumption. The judges, though, were outsiders, not the people actually seeking support. The study did not measure any long-term benefit, and the replies were not matched in length.
The researchers themselves note that sounding compassionate is different from giving the deep care that mental health problems require. As for why the AI did so well, the authors point to fatigue. An AI does not get tired and can attend closely to the details of each message, while human helpers may tire or be upset by hard cases.
198 words ・ CEFR B2
和訳を見る
トロント大学の研究者たちは、人々に、うれしい個人的経験やつらい個人的経験を記した短い文章を、それぞれへの返答とセットで読んでもらいました。返答の一部はGPT-4が、一部は一般の人が、一部は危機相談の専門家が書いたものです。2025年に『Communications Psychology』で発表された、合計556人が参加した事前登録済みの4つの実験で、第三者である評価者は、どの比較でもAIの返答のほうが思いやりがあると評価しました。その理由の一部は、AIのほうが応答的に見えたことにありました。
AIの返答は、理解・承認・気づかいを伝えていたのです。どの返答がAIのものかを評価者が知っていても、この優位は変わりませんでした。共感は長い間、人間ならではの強みとされてきたので、この結果はよくある前提に疑問を投げかけます。ただし、評価したのは第三者であり、実際に支援を求めている本人ではありません。長期的な効果は測定されておらず、返答の長さもそろえられていませんでした。
研究者自身も、思いやりがあるように聞こえることと、心の問題に必要な深いケアを提供することは違うと指摘しています。AIがこれほどうまくいった理由として、著者らは疲れを挙げています。AIは疲れることがなく、一つ一つのメッセージの細部に丁寧に注意を向けられますが、人間の支援者は疲れたり、つらい相談に心を乱されたりすることがあるのです。
4択リーディングクイズ
Q1. What assumption does the study challenge?
Empathy has long been treated as a uniquely human strength, so the result challenges a common assumption. とあるので D。
Q2. According to the analysis, why were the AI responses rated as more compassionate?
this was partly because the AI seemed more responsive. Its replies conveyed understanding, validation, and care. とあるので B。
Q3. Which of the following is mentioned as a limitation of the study?
The judges, though, were outsiders, not the people actually seeking support. とあるので A。実験は4つで、AIの返答が短かったとは書かれておらず(長さはそろえられていなかっただけ)、作者を知らせた条件もあったので B〜D は誤り。
重要単語
- compassionB2名詞思いやり、慈悲She showed great compassion for the victims.
com-(共に)+passion(苦しみ)=「共に苦しむ」。形容詞は compassionate。 - empathyB2名詞共感Empathy means understanding how others feel.
sympathy(同情)との違いが頻出。empathy は相手の立場に立って感じること。動詞は empathize with ~。 - third partyB2名詞(句)第三者A third party judged the two answers.
当事者(the parties concerned)以外の人。形容詞的に third-party evaluators「第三者の評価者」とも使う。 - responderC1名詞対応者、応答する人Crisis responders answer calls all night.
respond(応答する)+-er。first responder は「(事故現場に)最初に駆けつける救急隊員など」。 - preregisteredC1形容詞事前登録されたThe study was preregistered to prevent bias.
pre-(前もって)+registered(登録された)。結果を見てから仮説を変えるのを防ぐ、信頼性を高める工夫。 - validationC1名詞(気持ちの)承認、妥当性の確認Validation tells people that their feelings make sense.
validate(正しいと認める)の名詞形。心理学では「相手の気持ちをもっともだと受け止めること」。
動画でもっと知る
動画は各チャンネルの公式埋め込みで表示しています(YouTube)。
この話題を TED で聞く
同じテーマの TED・TED-Ed を、
TED HACK で。英語の見どころ・使える表現・学習ミッションつきの日本語ガイドです。(運営:原田英語)
公式・一次情報をチェック
このネタ、誰かに教えたくなった?
悩み相談への返事を読み比べたら、AIの方が「思いやりがある」と評価されたらしい。相手が危機相談のプロでも、AIが書いたと明かしても結果は同じ。理由は「疲れない」「細部を拾う」から?
押すと「マジかリスト」に保存されます
出典・参考文献(3件)タップで表示
- Ovsyannikova, de Mello & Inzlicht (2025) Third-party evaluators perceive AI as more compassionate than expert humans, Communications Psychologydoi.org
- Rubin et al. (2025) Comparing the value of perceived human versus AI-generated empathy, Nature Human Behaviourdoi.org
- Ayers et al. (2023) Comparing Physician and Artificial Intelligence Chatbot Responses to Patient Questions, JAMA Internal Medicinedoi.org
※本記事は上記の研究・公開資料をもとに、原田英語が独自の言葉で解説したものです。引用は著作権法第32条の範囲で行い、出典を明記しています。画像はすべてオリジナルのイメージです。研究結果には限界や個人差があります。誤りのご指摘は原田英語のお問い合わせフォームから。













