中文MedVIGIL:在视觉证据受损情况下评估可信的医疗VLM
ENMedVIGIL: Evaluating Trustworthy Medical VLMs Under Broken Visual Evidence
医学视觉语言模型在证据被破坏(如假前提、措辞扰动、知识改写或ROI损坏图像)时,仍会给出流利但错误的回答(静默失败)。本研究引入medvigil,包含300个评估案例,旨在检验模型识别证据失效的能力,对临床可信赖性至关重要。
arXiv:2605.07919v2 Announce Type: replace Abstract: Medical vision--language models (VLMs) are usually evaluated on intact image--question pairs, but trustworthy clinical use requires a stronger property: a model must recognise when the evidential basis for an answer has failed. We study this through silent failures under perturbed evidence, where a vision-required medical question is paired with a false premise, wording perturbation, knowledge-only rewrite, or ROI-corrupted image, yet the model returns a fluent non-refusal answer. We introduce medvigil, a 300-case evaluation suite drawn from