Let’s Agree to Disagree: Learning Highly Debatable Multirater Labelling

Carole H. Sudre, Beatriz Gomez Anson,Silvia Ingala, Chris D. Lane,Daniel Jimenez,Lukas Haider,Thomas Varsavsky,Ryutaro Tanno,Lorna Smith,Sébastien Ourselin, Rolf H. Jäger,M. Jorge Cardoso

MEDICAL IMAGE COMPUTING AND COMPUTER ASSISTED INTERVENTION - MICCAI 2019, PT IV（2019）

引用 21|浏览0

暂无评分

摘要

Classification and differentiation of small pathological objects may greatly vary among human raters due to differences in training, expertise and their consistency over time. In a radiological setting, objects commonly have high within-class appearance variability whilst sharing certain characteristics across different classes, making their distinction even more difficult. As an example, markers of cerebral small vessel disease, such as enlarged perivascular spaces (EPVS) and lacunes, can be very varied in their appearance while exhibiting high inter-class similarity, making this task highly challenging for human raters. In this work, we investigate joint models of individual rater behaviour and multi-rater consensus in a deep learning setting, and apply it to a brain lesion object-detection task. Results show that jointly modelling both individual and consensus estimates leads to significant improvements in performance when compared to directly predicting consensus labels, while also allowing the characterization of human-rater consistency.

查看译文

关键词

Deep learning,Noisy labels,Classification

AI 理解论文

溯源树

样例

生成溯源树，研究论文发展脉络

Chat Paper

正在生成论文摘要