SMILE: Self-Explainable Multimodal Information Bottleneck for Medical Diagnosis

📅 2026-09-04
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
本文通过信息瓶颈框架提出了一种自解释多模态诊断方法,优化预测性能和模ality特定解释性,提高医疗诊断的透明度和准确性。
📝 Abstract
Explainability is increasingly seen as a crucial requirement in AI-based medical diagnosis, particularly in safety-critical clinical decision-making. Most existing explainability methods in healthcare operate in a post-hoc manner and are predominantly designed for unimodal data, which limits their applicability in increasingly prevalent multimodal diagnostic settings. This paper addresses the problem of self-explainable multimodal diagnosis by formulating it within the information bottleneck (IB) framework. We propose a unified learning paradigm that jointly optimizes predictive performance and modality-specific explainability by identifying the most informative elements inside each modality that contribute to diagnostic decisions. To enable tractable and stable optimization, we employ a matrix-based Renyi's $α$-order entropy functional under the assumption of sufficiently expressive encoders. Extensive experiments on representative medical datasets spanning heterogeneous modalities demonstrate that the proposed method consistently achieves strong diagnostic performance, including an absolute accuracy improvement of 9.1 percentage points on the iCTCF dataset. Moreover, the learned explanations provide transparent and modality-aware insights into feature relevance, thereby improving both the explainability and generalization.
Problem

Research questions and friction points this paper is trying to address.

Explainability
Multimodal Diagnosis
Information Bottleneck
Medical AI
Clinical Decision-Making
Innovation

Methods, ideas, or system contributions that make the work stand out.

self-explainable
multimodal information bottleneck
matrix-based Renyi's α-order entropy
predictive performance and explainability
medical diagnosis
💼 Related Jobs
No related jobs found.
Y
Yuqing Yang
Independent Researcher
A
Alexander Schmatz
Leiden Institute of Advanced Computer Science (LIACS), Leiden University, Leiden, The Netherlands
Z
Zhaozhao Ma
School of Applied and Creative Computing, Purdue University, West Lafayette, IN, USA
Changkyu Choi
Changkyu Choi
University of Tromsø - The Arctic University of Norway
Information-Theoretic LearningHuman-Centric AIMarine Intelligence
Robert Jenssen
Robert Jenssen
Visual Intelligence, UiT The Arctic University of Norway & Norw. Comp. Center & P1 Centre AI, UCPH
Machine learninginformation theoretic learningkernel methodsdeep learninghealth data analytics
S
Shujian Yu
Machine Learning Group, UiT — The Arctic University of Norway, Tromsø, Norway; Quantitative Data Analytics Group, Vrije Universiteit Amsterdam, Amsterdam, The Netherlands