Abstract
We study the problem of object detection when training and test objects are disjoint, i.e. no training examples of the target classes are available. Existing unseen object detection approaches usually combine generic detection frameworks with a single-path unseen classifier, by aligning object regions with semantic class embeddings. In this paper, inspired from human cognitive experience, we propose a simple but effective dual-path detection model that further explores associative semantics to supplement the basic visual-semantic knowledge transfer. We use a novel target-centric multiple-association strategy to establish concept associations, to ensure that the predictor generalized to unseen domain can be learned during training. In this way, through a reasonable inference fusion mechanism, those two parallel reasoning paths can strengthen the correlation between seen and unseen objects, thus improving detection performance. Experiments show that our inductive method can significantly boost the performance by 7.42% over inductive models, and even 5.25% over transductive models on MSCOCO dataset.
Cite
CITATION STYLE
Li, Y., Li, P., Cui, H., & Wang, D. (2021). Inference Fusion with Associative Semantics for Unseen Object Detection. In 35th AAAI Conference on Artificial Intelligence, AAAI 2021 (Vol. 3A, pp. 1993–2001). Association for the Advancement of Artificial Intelligence. https://doi.org/10.1609/aaai.v35i3.16295
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.