日本フィジカルAI新聞

世界のフィジカルAIを、日本語で。

週刊ニュースレター購読
HCI/共在コラボレーションarXiv:2608.26185

代わりに言ってくれませんか?共在ディスカッションにおける代理発言

Can You Say This for Me? Speaking Up by Proxy in Co-Located Discussion

シェア:XThreadsFacebookLINEはてブBluesky

共在ディスカッションで発言しにくい意見を、仮想プロキシを通じて代弁するMRシステム「SecondVoice」を提案し、テキストボードと比較して代理発言が会話に参加しやすいことを示した。

詳しい要約

1. どんなもの?

SecondVoiceは、共同作業中の対面討論において、発言者が直接声に出さずに、具現化された仮想プロキシ(embodied virtual proxy)を通じて発言できるようにするmixed-realityシステムです。発言内容と発言者を分離することで、社会的リスクを伴う発言をライブの音声討論に取り込みます。ユーザーはプライベートなオーバーレイで意図を構造化された仕様として入力し、システムがそれを言い換えてプロキシを通じて会話に発話します。

2. 先行研究と比べてどこがすごい?

先行研究では、匿名のテキストボードやチャットなど、発言を音声から分離する方法が提案されてきたが、それらは発言が音声の議論の流れに乗りにくいという問題があった。SecondVoiceは、プロキシを通じて発言を音声として発話することで、発言が議論のフロアに直接入り、グループの多ターンにわたる関与を引き出す点で優れている。また、ユーザーが完全な発話を構成するのではなく、構造化された仕様入力を行うことで、発言のハードルを下げている。

3. 技術・手法の肝は?

手法の肝は、発言内容と発言者を分離するためのmixed-realityシステムの設計にある。ユーザーはプライベートなオーバーレイで意図を構造化された仕様として入力し、システムがそれを言い換えて、具現化された仮想プロキシを通じて音声として発話する。これにより、ユーザーは直接声を出さずに発言でき、プロキシが発言の主体となる。また、社会的リスク下での参加チャネルの設計空間を特徴づけている。

4. どうやって有効だと検証した?

予備的な被験者内研究(N=16)を実施し、完全なSecondVoiceシステムと匿名テキストボードチャネルを、2つのグループ討論タスクで比較した。その結果、参加者の半数がSecondVoiceを声に出さなかった発言に使用したのに対し、テキストボードでは18.8%だった。プロキシによる発言は音声の議論に取り込まれ、多ターンにわたるグループの関与を引き出したが、テキストボードの投稿後にはそのような関与は観察されなかった。

5. 議論はある?

参加者は、このチャネルが状況によっては価値があると述べたが、タイミング、所有権、言い換えへの信頼に関するトレードオフを指摘した。要旨からは、これらのトレードオフの詳細や、システムの有効性に関する限界は不明である。また、予備的研究であり、サンプルサイズが小さいことや、長期的な影響については議論されていない。

6. 次に読むべき論文は?

要旨で参照されている研究は明示されていないが、関連する分野として、CSCW(Computer-Supported Cooperative Work)における参加支援技術、mixed-realityシステム、匿名コミュニケーション、社会的リスク下での発言に関する研究が挙げられる。具体的には、匿名テキストボードやチャットシステム、embodied agentsを用いた議論支援システムなどが関連する。

※ AIが要旨から生成した要約です。正確性は原文をご確認ください。

著者: Yue Shen, Rehema Abulikemu, Ryan P. McMahan, Yan Chen

分類: cs.AI, cs.ET, cs.HC

原文アブストラクト

Equal participation in co-located discussion is important for effective collaboration, yet people often hold back when they anticipate negative interpersonal or professional consequences, especially when raising a point requires voicing it themselves. We present SecondVoice, a mixed-reality system that enables people to speak up through an embodied virtual proxy. By separating what is said from who says it, SecondVoice brings hesitant points into the live spoken discussion without putting the speaker on the spot. Using a private overlay, users specify their intent through a structured specification process rather than composing a full utterance. The system reformulates the input and voices it into the conversation through the proxy. We characterize a design space of participation channels under social risk. In a preliminary within-subject study (N = 16), we compare the complete SecondVoice system with an anonymous text-board channel across two group discussion tasks. Half of participants reported using SecondVoice for a point they did not say aloud, compared with 18.8% for the text board. Proxy-delivered points entered the spoken floor and were followed by multi-turn group engagement, which we did not observe after text-board posts. Participants described the channel as situationally valuable but identified tradeoffs around timing, ownership, and trust in reformulation.