<p>Concept Bottleneck Models (CBMs) aim to improve interpretability by forcing predictions to pass through human-interpretable concepts. However, many CBMs achieve high predictive accuracy by bypassing the bottleneck through latent shortcut features, a phenomenon known as <i>concept leakage</i>. Such behavior weakens the reliability of concept-based explanations, particularly in high-stakes clinical applications. To address this, we propose LCBM (Leakage-Constrained Bottleneck Model), a purified concept bottleneck architecture that decomposes latent representations into two structurally decorrelated subspaces: a concept-aligned semantic space (<InlineEquation ID="IEq1"> <EquationSource Format="TEX">\(z_{sem}\)</EquationSource> </InlineEquation>) containing clinically meaningful features and a residual space (<InlineEquation ID="IEq2"> <EquationSource Format="TEX">\(z_{res}\)</EquationSource> </InlineEquation>) capturing non-diagnostic artifacts. Through orthogonality regularization and an adversarial Gradient Reversal Layer (GRL), the framework minimizes the propagation of diagnostic information through non-semantic pathways. The framework is evaluated on the Derm7pt and HAM10000 dermatology benchmarks using ResNet and Vision Transformer backbones. Our results demonstrate that LCBM significantly improves the semantic reliability of CBMs, reducing the leakage gap from 0.4487 to 0.0539 under ground-truth concept intervention. The model achieves competitive AUC scores of 0.8358 (ResNet-50) and 0.8107 (ViT-B/16) on Derm7pt, while reaching an AUC of 0.94 on HAM10000. These results suggest that structured latent purification through decorrelation can enhance faithful concept-based reasoning in clinical AI systems, providing a more transparent foundation for diagnostic decision support.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

A vision transformer based leakage free concept bottleneck model for clinical skin lesion diagnosis

  • Kailash Chandra Kandpal,
  • Prabhat Verma

摘要

Concept Bottleneck Models (CBMs) aim to improve interpretability by forcing predictions to pass through human-interpretable concepts. However, many CBMs achieve high predictive accuracy by bypassing the bottleneck through latent shortcut features, a phenomenon known as concept leakage. Such behavior weakens the reliability of concept-based explanations, particularly in high-stakes clinical applications. To address this, we propose LCBM (Leakage-Constrained Bottleneck Model), a purified concept bottleneck architecture that decomposes latent representations into two structurally decorrelated subspaces: a concept-aligned semantic space ( \(z_{sem}\) ) containing clinically meaningful features and a residual space ( \(z_{res}\) ) capturing non-diagnostic artifacts. Through orthogonality regularization and an adversarial Gradient Reversal Layer (GRL), the framework minimizes the propagation of diagnostic information through non-semantic pathways. The framework is evaluated on the Derm7pt and HAM10000 dermatology benchmarks using ResNet and Vision Transformer backbones. Our results demonstrate that LCBM significantly improves the semantic reliability of CBMs, reducing the leakage gap from 0.4487 to 0.0539 under ground-truth concept intervention. The model achieves competitive AUC scores of 0.8358 (ResNet-50) and 0.8107 (ViT-B/16) on Derm7pt, while reaching an AUC of 0.94 on HAM10000. These results suggest that structured latent purification through decorrelation can enhance faithful concept-based reasoning in clinical AI systems, providing a more transparent foundation for diagnostic decision support.