Medical AI Encodes a "Feeling of Error": Verifying Cancer Segmentation via Internal Concepts
Ranking
Overall
85
Content
95
Popularity
61
Observed public metrics from 1 member.
Merged summary
TL;DR - This paper finds that cancer segmentation models encode internal activation patterns that distinguish successful predictions from failures. Detecting this latent “feeling of error” could flag unreliable masks without sacrificing segmentation quality.
- Sparse autoencoders decompose internal neural activations into human-interpretable concepts.
- Failed segmentations exhibit fewer active concepts and lower activation magnitudes than successful cases.
- A classifier trained on these concept activations detects failures and provides explanations tied to internal model concepts.
- Across prostate, pancreatic, and brain cancer segmentation, the method outperforms output-based failure detection approaches while preserving segmentation quality.
Sources (1)
Medical AI Encodes a "Feeling of Error": Verifying Cancer Segmentation via Internal Concepts
Public signals
Semantic Scholar citations 1 · Semantic Scholar influential citations 0
TL;DR - This paper finds that cancer segmentation models encode internal activation patterns that distinguish successful predictions from failures. Detecting this latent “feeling of error” could flag unreliable masks without sacrificing segmentation quality.
- Sparse autoencoders decompose internal neural activations into human-interpretable concepts.
- Failed segmentations exhibit fewer active concepts and lower activation magnitudes than successful cases.
- A classifier trained on these concept activations detects failures and provides explanations tied to internal model concepts.
- Across prostate, pancreatic, and brain cancer segmentation, the method outperforms output-based failure detection approaches while preserving segmentation quality.