학술논문

Generalized Orderless Pooling Performs Implicit Salient Matching

Document Type

Conference

Author

Simon, Marcel; Gao, Yang; Darrell, Trevor; Denzler, Joachim; Rodner, Erik

Source

2017 IEEE International Conference on Computer Vision (ICCV) ICCV Computer Vision (ICCV), 2017 IEEE International Conference on. :4970-4979 Oct, 2017

Subject

Computing and Processing
Training
Encoding
Birds
Visualization
Head
Kernel

Language

ISSN

2380-7504

Abstract

Most recent CNN architectures use average pooling as a final feature encoding step. In the field of fine-grained recognition, however, recent global representations like bilinear pooling offer improved performance. In this paper, we generalize average and bilinear pooling to “α-pooling”, allowing for learning the pooling strategy during training. In addition, we present a novel way to visualize decisions made by these approaches. We identify parts of training images having the highest influence on the prediction of a given test image. This allows for justifying decisions to users and also for analyzing the influence of semantic parts. For example, we can show that the higher capacity VGG16 model focuses much more on the bird's head than, e.g., the lower-capacity VGG-M model when recognizing fine-grained bird categories. Both contributions allow us to analyze the difference when moving between average and bilinear pooling. In addition, experiments show that our generalized approach can outperform both across a variety of standard datasets.

Online Access

Full Text (IEEE) Find it@PNU

이메일

부산대학교 도서관

Online Access

메일 발송