Visual context embeddings for zero-shot recognition

Citations

WEB OF SCIENCE

0
Citations

SCOPUS

0

초록

Existing word-embeddings have performed well in various downstream tasks, but there may be a bias towards the text domain because they are learned from a text corpus. When word-embeddings are used in the Zero-Shot Recognition(ZSR) task, the task becomes a mapping problem between two completely different heterogeneous domains, a low-level visual feature domain, and a word embedding domain, and due to the bias of word-embeddings, it was not easy to learn this mapping function. However, if the context of the visual domain can be learned and embedded, the mapping function of ZSR will be much easier to converge because it only needs to learn the mapping between domains that are more correlated to each other. Therefore, in this paper, we propose a new methodology for embedding the context contained in the visual domain using the annotation information collected from the image dataset. In addition, to utilize the annotations collected from the image dataset for embedding, we proposed a new distance formula to measure the contextual distance between the bounding boxes of objects. Finally, it was verified through various experiments on two datasets that the embeddings learned by our new methodology performed well when applied to ZSR.

키워드

semantic embeddingszero-shot learningzero-shot recognitionComputer visionMappingSemanticsDown-streamEmbeddingsImage datasetsLearn+Mapping functionsSemantic embeddingText corporaVisual contextZero-shot learningZero-shot recognitionEmbeddings
제목
Visual context embeddings for zero-shot recognition
저자
Cho, GunheeChoi, Yong Suk
DOI
10.1145/3477314.3507071
발행일
2022-04
유형
Proceedings Paper
저널명
37TH ANNUAL ACM SYMPOSIUM ON APPLIED COMPUTING
페이지
1039 ~ 1047