Search By Image: Deeply Exploring Beneficial Features for Beauty Product Retrieval
It addresses a domain-specific problem of image-based beauty product retrieval, which is incremental as it builds on existing neural network methods with a novel feature combination approach.
The paper tackles the challenge of beauty product retrieval by developing a variable-attention neural network (VM-Net) that combines multiple image features to suppress data variations and distinguish similar products, achieving clear improvements in MAP@7 on the Perfect-500K benchmark dataset.
Searching by image is popular yet still challenging due to the extensive interference arose from i) data variations (e.g., background, pose, visual angle, brightness) of real-world captured images and ii) similar images in the query dataset. This paper studies a practically meaningful problem of beauty product retrieval (BPR) by neural networks. We broadly extract different types of image features, and raise an intriguing question that whether these features are beneficial to i) suppress data variations of real-world captured images, and ii) distinguish one image from others which look very similar but are intrinsically different beauty products in the dataset, therefore leading to an enhanced capability of BPR. To answer it, we present a novel variable-attention neural network to understand the combination of multiple features (termed VM-Net) of beauty product images. Considering that there are few publicly released training datasets for BPR, we establish a new dataset with more than one million images classified into more than 20K categories to improve both the generalization and anti-interference abilities of VM-Net and other methods. We verify the performance of VM-Net and its competitors on the benchmark dataset Perfect-500K, where VM-Net shows clear improvements over the competitors in terms of MAP@7. The source code and dataset will be released upon publication.