← Back to all articles
arXiv cs.CLSeptember 23, 2026

MultiViewDx: Evidence-Linked Multi-View Clinical Diagnosis

Excerpt

arXiv:2410.14948v2 Announce Type: replace Abstract: Medical multimodal large language models (MLLMs) can perform well on existing medical visual question answering (MedVQA) benchmarks, but their training data often does not match clinical diagnosis. Most supervision is organized around isolated images or short QA pairs, leaving two structures weakly specified: how evidence leads to a decision, and how views, series, modalities, and patient context from the same case are linked. We introduce Mult