Abstract:Traditional active learning methods require the labeler to provide a class label for each queried instance. The labelers are normally highly skilled domain experts to ensure the correctness of the provided labels, which in turn results in expensive labeling cost. To reduce labeling cost, an alternative solution is to allow nonexpert labelers to carry out the labeling task without explicitly telling the class label of each queried instance. In this paper, we propose a new active learning paradigm, in which a nonexpert labeler is only asked “whether a pair of instances belong to the same class”, namely, a pairwise label homogeneity. Under such circumstances, our active learning goal is twofold: (1) decide which pair of instances should be selected for query, and (2) how to make use of the pairwise homogeneity information to improve the active learner. To achieve the goal, we propose a “Pairwise Query on Max-flow Paths” strategy to query pairwise label homogeneity from a nonexpert labeler, whose query results are further used to dynamically update a Min-cut model (to differentiate instances in different classes). In addition, a “Confidence-based Data Selection” measure is used to evaluate data utility based on the Min-cut model’s prediction results. The selected instances, with inferred class labels, are included into the labeled set to form a closed-loop active learning process. Experimental results and comparisons with state-of-the-art methods demonstrate that our new active learning paradigm can result in good performance with nonexpert labelers.

Active Learning Through Label Error Statistical Methods

Uncertainty-aware Complementary Label Queries for Active Learning

Distributed Active Learning.

REAL: A Representative Error-Driven Approach for Active Learning

Cost-Accuracy Aware Adaptive Labeling for Active Learning

ActiveLab: Active Learning with Re-Labeling by Multiple Annotators

Combining Topological Analysis Matrices-Based Active Learning on Networked Data Classification

Learning to Label with Active Learning and Reinforcement Learning.

Active Learning Without Knowing Individual Instance Labels: A Pairwise Label Homogeneity Query Approach.

New Balanced Active Learning Model and Optimization Algorithm.

Cost-Effective Active Learning from Diverse Labelers.

Data : Labeler 1 : Labeler 2 : Labeler 3 : Figure

Combining Clustering Coefficient-Based Active Learning and Semi-Supervised Learning on Networked Data

Active Learning from Weak and Strong Labelers

Composite Active Learning: Towards Multi-Domain Active Learning with Theoretical Guarantees

Active Learning without Knowing Individual Instance Labels

A Survey on Active Learning: State-of-the-Art, Practical Challenges and Research Directions

A meta-active learning approach exploiting instance importance

Active Multi-label Learning with Optimal Label Subset Selection

AnchorAL: Computationally Efficient Active Learning for Large and Imbalanced Datasets

Addressing practical challenges in Active Learning via a hybrid query strategy