Abstract:Author name ambiguity causes inadequacy and inconvenience in academic information retrieval, which raises the necessity of author name disambiguation (AND). Existing AND methods can be divided into two categories: the models focusing on content information to distinguish whether two papers are written by the same author, the models focusing on relation information to represent information as edges on the network and to quantify the similarity among papers. However, the former requires adequate labeled samples and informative negative samples, and are also ineffective in measuring the high-order connections among papers, while the latter needs complicated feature engineering or supervision to construct the network. We propose a novel generative adversarial framework to grow the two categories of models together: (i) the discriminative module distinguishes whether two papers are from the same author, and (ii) the generative module selects possibly homogeneous papers directly from the heterogeneous information network, which eliminates the complicated feature engineering. In such a way, the discriminative module guides the generative module to select homogeneous papers, and the generative module generates high-quality negative samples to train the discriminative module to make it aware of high-order connections among papers. Furthermore, a self-training strategy for the discriminative module and a random walk based generating algorithm are designed to make the training stable and efficient. Extensive experiments on two real-world AND benchmarks demonstrate that our model provides significant performance improvement over the state-of-the-art methods.

A Graph-Based Author Name Disambiguation Method and Analysis via Information Theory

Author Name Disambiguation Based on Heterogeneous Graph

Exploring Graph Based Approaches for Author Name Disambiguation

Author Name Disambiguation Using Multiple Graph Attention Networks

Author Name Disambiguation via Heterogeneous Network Embedding from Structural and Semantic Perspectives

On Graph-Based Name Disambiguation.

An Effective Approach for Automatic Author Name Disambiguation Based on Multiple Strategies

Author Name Disambiguation Based on Rule and Graph Model

Author Name Disambiguation on Heterogeneous Information Network with Adversarial Representation Learning

Unsupervised Author Disambiguation Using Dempster–Shafer Theory

A Unified Framework for Name Disambiguation

High‐degree penalty based global statistical network embedding for name disambiguation in anonymized graph

A Constraint-Based Probabilistic Framework for Name Disambiguation

Name Disambiguation Using Atomic Clusters.

On Disambiguating Authors: Collaboration Network Reconstruction in a Bottom-up Manner.

Name Disambiguation By Collective Classification

Name Disambiguation in AMiner: Clustering, Maintenance, and Human in the Loop.

A Unified Semi-supervised Framework for Author Disambiguation in Academic Social Network.

A Bayesian Learning, Greedy agglomerative clustering approach and evaluation techniques for Author Name Disambiguation Problem

MORE: Toward Improving Author Name Disambiguation in Academic Knowledge Graphs

ADANA: Active Name Disambiguation