Skip to main content
Aggregate arXiv cs.AI 人工智能 3 Sep 2026 - 14:30

Examining the Vulnerability of Multi-Agent Medical Systems to Human Interventions for Clinical Reasoning

RSS 官方收录 · 可信分层展示

关键摘要

arXiv:2609.02191v1 Announce Type: new Abstract: Human interventions at fault points can alter the diagnostic accuracy of multi-agent medical systems.…

  • We defined fault points as moments in AI agent conversations, in which…
  • Using the MedQA dataset, this study analyzed simulated doctor-patient …
  • Correct intervention methods showed an improvement in baseline diagnos…

摘要引擎:抽取

正文提要

arXiv:2609.02191v1 Announce Type: new Abstract: Human interventions at fault points can alter the diagnostic accuracy of multi-agent medical systems. We defined fault points as moments in AI agent conversations, in which an agent's reasoning became most vulnerable to external influence. Using the MedQA dataset, this study analyzed simulated doctor-patient conversations to measure how interventions shifted reasoning and accuracy. Correct intervention methods showed an improvement in baseline diagnostic accuracy of up to 40%, while incorrect or bias-related interventions degraded performance by up to 6% and increased diagnostic drift and uncertainty. Beyond performance changes, our analysis revealed behavioral similarities between cognitive biases in simulated agent environments and real-world clinical practice. Examples included premature closure and susceptibility to misleading cues. Overall, these findings demonstrate that identifying and guiding fault points with human interventions may provide a mechanism for improving diagnostic robustness in multi-agent medical systems.

来源:https://arxiv.org/abs/2609.02191

打开官方原文 站点原文页 可信分区 本信源更多 今日简报 分享图 RSS 稍后再看列表