Learning Local Descriptors by Optimizing the Keypoint-Correspondence Criterion: Applications to Face Matching, Learning From Unlabeled Videos and 3D-Shape Retrieval.

Nenad Markus,Igor Pandzic,Jorgen Ahlberg

doi:10.1109/tip.2018.2867270

Nenad Markus, Igor Pandzic + Show 1 more

Open Access

https://doi.org/10.1109/tip.2018.2867270

Copy DOI

Journal: IEEE Transactions on Image Processing	Publication Date: May 22, 2018
Citations: 35	License type: public-domain

Affiliation: University of Zagreb

Abstract

Current best local descriptors are learned on a large data set of matching and non-matching keypoint pairs. However, data of this kind are not always available, since the detailed keypoint correspondences can be hard to establish. On the other hand, we can often obtain labels for pairs of keypoint bags. For example, keypoint bags extracted from two images of the same object under different views form a matching pair, and keypoint bags extracted from images of different objects form a non-matching pair. On average, matching pairs should contain more corresponding keypoints than non-matching pairs. We describe an end-to-end differentiable architecture that enables the learning of local keypoint descriptors from such weakly labeled data. In addition, we discuss how to improve the method by incorporating the procedure of mining hard negatives. We also show how our approach can be used to learn convolutional features from unlabeled video signals and 3D models.

Full Text