Abstract:Abstract Real life applications of deep learning (DL) are often limited by the lack of expert labeled data required to effectively train DL models. Creation of such data usually requires substantial amount of time for manual categorization, which is costly and is considered to be one of the major impediments in development of DL methods in many areas. This work proposes a classification approach which completely removes the need for costly expert labeled data and utilizes noisy web data created by the users who are not subject matter experts. The experiments are performed with two well-known Convolutional Neural Network (CNN) architectures: VGG16 and ResNet50 trained on three randomly collected Instagram-based sets of images from three distinct domains: metropolitan cities, popular food and common objects - the last two sets were compiled by the authors and made freely available to the research community. The dataset containing common objects is a webly counterpart of PascalVOC2007 set. It is demonstrated that despite significant amount of label noise in the training data, application of proposed approach paired with standard training CNN protocol leads to high classification accuracy on representative data in all three above-mentioned domains. Additionally, two straightforward procedures of automatic cleaning of the data, before its use in the training process, are proposed. Apparently, data cleaning does not lead to improvement of results which suggests that the presence of noise in webly data is actually helpful in learning meaningful and robust class representations. Manual inspection of a subset of web-based test data shows that labels assigned to many images are ambiguous even for humans. It is our conclusion that for the datasets and CNN architectures used in this paper, in case of training with webly data, a major factor contributing to the final classification accuracy is representativeness of test data rather than application of data cleaning procedures.

Extracting Useful Knowledge from Noisy Web Images Via Data Purification for Fine-Grained Recognition.

Exploiting Web Images for Fine-Grained Visual Recognition by Eliminating Open-Set Noise and Utilizing Hard Examples

Data-driven Meta-set Based Fine-Grained Visual Classification

FGCM: Noisy Label Learning via Fine-Grained Confidence Modeling

Group Benefits Instances Selection for Data Purification

Group benefits instance for data purification

Robust fine‐grained visual recognition with images based on internet of things

Remote Sensing Image Scene Classification with Noisy Label Distillation

Training CNN Classifiers Solely on Webly Data

Web-Supervised Network for Fine-Grained Visual Classification.

Learning to Purification for Unsupervised Person Re-identification

Data reweighting net for web fine-grained image classification

HCL: Hierarchical Consistency Learning for Webly Supervised Fine-Grained Recognition

Noisy Label Processing for Classification: A Survey

Webly Supervised Fine-Grained Recognition: Benchmark Datasets and An Approach

Multi-Label and Evolvable Dataset Preparation for Web-Based Object Detection

An accurate detection is not all you need to combat label noise in web-noisy datasets

Mining Weakly Labeled Web Facial Images for Search-Based Face Annotation

Web Image Annotation Based On Automatically Obtained Noisy Training Set

Learning From Large-Scale Noisy Web Data With Ubiquitous Reweighting for Image Classification

Purifying real images with an attention-guided style transfer network for gaze estimation