Skip to main content
Aggregate arXiv cs.AI 人工智能 25 Aug 2026 - 15:00

Ask or Answer: A Decision Framework for Multi-Turn Health Misinformation Intervention

RSS 官方收录 · 可信分层展示

关键摘要

arXiv:2608.…

  • 21721v1 Announce Type: new Abstract: Correcting health misinformation …
  • Yet existing methods either respond immediately or probe indiscriminat…
  • We propose Reward-Optimized Probe-and-Respond (RO-PnR), a framework th…

摘要引擎:抽取

正文提要

arXiv:2608.21721v1 Announce Type: new Abstract: Correcting health misinformation in dialogue requires more than producing a factual rebuttal: users differ in what they know, what they believe, and what they need to hear, so an effective intervention often depends on first asking the right clarifying question. Yet existing methods either respond immediately or probe indiscriminately, treating clarification as either unnecessary or always beneficial. We propose Reward-Optimized Probe-and-Respond (RO-PnR), a framework that learns when asking is worth its cost. At each turn, RO-PnR chooses between probing for more information and committing to a final correction, guided by a turn-level reward that weighs the expected gain from probing against its interaction cost. To capture how user heterogeneity affects probing value, we model each simulated user with a latent state along health literacy and belief commitment. Experiments show that RO-PnR achieves the highest cost-adjusted utility across three health-misinformation datasets and three base models, using 30% fewer turns than always-probe baselines.

来源:https://arxiv.org/abs/2608.21721

打开官方原文 站点原文页 可信分区 本信源更多 今日简报 分享图 RSS 稍后再看列表