Abstract:Motivation: The explosive increase of biomedical literature has made information extraction an increasingly important tool for biomedical research. A fundamental task is the recognition of biomedical named entities in text (BNER) such as genes/proteins, diseases and species. Recently, a domain-independent method based on deep learning and statistical word embeddings, called long short-term memory network-conditional random field (LSTM-CRF), has been shown to outperform state-of-the-art entity-specific BNER tools. However, this method is dependent on gold-standard corpora (GSCs) consisting of hand-labeled entities, which tend to be small but highly reliable. An alternative to GSCs are silver-standard corpora (SSCs), which are generated by harmonizing the annotations made by several automatic annotation systems. SSCs typically contain more noise than GSCs but have the advantage of containing many more training examples. Ideally, these corpora could be combined to achieve the benefits of both, which is an opportunity for transfer learning. In this work, we analyze to what extent transfer learning improves upon state-of-the-art results for BNER.Results: We demonstrate that transferring a deep neural network (DNN) trained on a large, noisy SSC to a smaller, but more reliable GSC significantly improves upon state-of-the-art results for BNER. Compared to a state-of-the-art baseline evaluated on 23 GSCs covering four different entity classes, transfer learning results in an average reduction in error of approximately 11%. We found transfer learning to be especially beneficial for target datasets with a small number of labels (approximately 6000 or less).Availability and implementation: Source code for the LSTM-CRF is available at https://github.com/Franck-Dernoncourt/NeuroNER/ and links to the corpora are available at https://github.com/BaderLab/Transfer-Learning-BNER-Bioinformatics-2018/.Supplementary information: Supplementary data are available at Bioinformatics online.

Noise Reduction Learning Based on XLNet-CRF for Biomedical Named Entity Recognition

Clinical Named Entity Recognition from Chinese Electronic Medical Records Based on Deep Learning Pretraining

Improving Biomedical Named Entity Recognition with a Unified Multi-Task MRC Framework

Named Entity Recognition by Using XLNet-BiLSTM-CRF

Language model based on deep learning network for biomedical named entity recognition

Integrating Language Model and Reading Control Gate in BLSTM-CRF for Biomedical Named Entity Recognition

Long short-term memory RNN for biomedical named entity recognition

Hierarchical shared transfer learning for biomedical named entity recognition

A BIGRU-Based Stacked Attention Network for Biomedical Named Entity Recognition with Chinese EMRs

Nested Named Entity Recognition from Medical Texts: An Adaptive Shared Network Architecture with Attentive CRF

An attention-based deep learning model for clinical named entity recognition of Chinese electronic medical records

Research on Chinese medical named entity recognition based on collaborative cooperation of multiple neural network models

A Hybrid Model Based on Deep Convolutional Network for Medical Named Entity Recognition

Named Entity Recognition Via Noise Aware Training Mechanism with Data Filter.

Biomedical named entity recognition with the combined feature attention and fully-shared multi-task learning

Biomedical named entity recognition using deep neural networks with contextual information

Online biomedical named entities recognition by data and knowledge-driven model

Transfer learning for biomedical named entity recognition with neural networks

Named entity recognition of Chinese electronic medical records based on a hybrid neural network and medical MC-BERT

Named entity recognition model based on Multi‐BiLSTM and competition mechanism

Serial and Parallel Recurrent Convolutional Neural Networks for Biomedical Named Entity Recognition