Abstract
Improving the retrieval relevance on noisy datasets is an emerging need for the curation of a large-scale clean dataset in the medical domain. While existing methods can be applied for class-wise retrieval (aka. inter-class), they cannot distinguish the granularity of likeness within the same class (aka. intra-class). The problem is exacerbated on medical external datasets, where noisy samples of the same class are treated equally during training. Our goal is to identify both intra/inter-class similarities for fine-grained retrieval. To achieve this, we propose an Outlier-Sensitive Content-based rAdiologhy Retrieval System (OSCARS), consisting of two steps. First, we train an outlier detector on a clean internal dataset in an unsupervised manner. Then we use the trained detector to generate the anomaly scores on the external dataset, whose distribution will be used to bin intra-class variations. Second, we propose a quadruplet (a, p, nintra, ninter) sampling strategy, where intra-class negatives nintra are sampled from bins of the same class other than the bin anchor a belongs to, while n_inter are randomly sampled from inter-classes. We suggest a weighted metric learning objective to balance the intra and inter-class feature learning. We experimented on two representative public radiography datasets. Experiments show the effectiveness of our approach. The training and evaluation code can be found in https://github.com/XiaoyuanGuo/oscars.
Author supplied keywords
Cite
CITATION STYLE
Guo, X., Duan, J., Purkayastha, S., Trivedi, H., Gichoya, J. W., & Banerjee, I. (2022). OSCARS: An Outlier-Sensitive Content-Based Radiography Retrieval System. In ICMR 2022 - Proceedings of the 2022 International Conference on Multimedia Retrieval (pp. 11–18). Association for Computing Machinery, Inc. https://doi.org/10.1145/3512527.3531425
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.