Astro-NER -- Astronomy Named Entity Recognition: Is GPT a Good Domain Expert Annotator?

Julia Evans,Sameer Sadruddin,Jennifer D'Souza
2024-05-04
Abstract:In this study, we address one of the challenges of developing NER models for scholarly domains, namely the scarcity of suitable labeled data. We experiment with an approach using predictions from a fine-tuned LLM model to aid non-domain experts in annotating scientific entities within astronomy literature, with the goal of uncovering whether such a collaborative process can approximate domain expertise. Our results reveal moderate agreement between a domain expert and the LLM-assisted non-experts, as well as fair agreement between the domain expert and the LLM model's predictions. In an additional experiment, we compare the performance of finetuned and default LLMs on this task. We have also introduced a specialized scientific entity annotation scheme for astronomy, validated by a domain expert. Our approach adopts a scholarly research contribution-centric perspective, focusing exclusively on scientific entities relevant to the research theme. The resultant dataset, containing 5,000 annotated astronomy article titles, is made publicly available.
Computation and Language,Artificial Intelligence,Information Theory
What problem does this paper attempt to address?
The problem that this paper attempts to solve is the challenge of scarce labeled data when developing Named Entity Recognition (NER) models in astronomical literature. Specifically, the authors experimented with a method that uses the predictions of fine - tuned large - scale language models (LLMs) to assist non - domain experts in labeling scientific entities, aiming to explore whether this collaborative process can approach the level of domain experts. In addition, they compared the performance of LLMs before and after fine - tuning on this task and introduced a scientific entity - labeling scheme specifically for astronomy. Through this method, the research aims to explore how to use LLMs to support or replace labelers in the absence of domain experts, thereby reducing the burden of generating labeled text data. The paper also released a data set containing 5,000 labeled astronomy article titles to promote research in related fields.