Proceedings of the 28th ACM International Conference on Multimedia 2020
DOI: 10.1145/3394171.3413638
|View full text |Cite
|
Sign up to set email alerts
|

Text-Embedded Bilinear Model for Fine-Grained Visual Recognition

Abstract: Fine-grained visual recognition, which aims to identify subcategories of the same base-level category, is a challenging task because of its large intra-class variances and small inter-class variances. Human beings can perform object recognition task based on not only the visual appearance but also the knowledge from texts, as texts can point out the discriminative parts or characteristics which are always the key to distinguishing different subcategories. This is an involuntary transfer from human textual atte… Show more

Help me understand this report

Search citation statements

Order By: Relevance

Paper Sections

Select...

Citation Types

0
0
0

Year Published

2021
2021
2024
2024

Publication Types

Select...
3
2

Relationship

0
5

Authors

Journals

citations
Cited by 7 publications
references
References 51 publications
0
0
0
Order By: Relevance